5 steps to ensure startups successfully deploy LLMs

5:35 AM PST • January 5, 2024

Computer Processor Processing Artificial Intelligence Data. Glowing Chip. Computer And Technology Related 3D Illustration Render. — **Image Credits:** yucelyilmaz / Getty Images

Lu Zhang

Contributor

Lu Zhang, the founder and managing partner of Fusion Fund, is a renowned Silicon Valley–based investor and a serial entrepreneur in healthcare.

ChatGPT’s launch ushered in the age of large language models. In addition to OpenAI’s offerings, other LLMs include Google’s LaMDA family of LLMs (including Bard), the BLOOM project (a collaboration between groups at Microsoft, Nvidia, and other organizations), Meta’s LLaMA, and Anthropic’s Claude.

More will no doubt be created. In fact, an April 2023 Arize survey found that 53% of respondents planned to deploy LLMs within the next year or sooner. One approach to doing this is to create a “vertical” LLM that starts with an existing LLM and carefully retrains it on knowledge specific to a particular domain. This tactic can work for life sciences, pharmaceuticals, insurance, finance, and other business sectors.

Deploying an LLM can provide a powerful competitive advantage — but only if it’s done well.

LLMs have already led to newsworthy issues, such as their tendency to “hallucinate” incorrect information. That’s a severe problem, and it can distract leadership from essential concerns with the processes that generate those outputs, which can be similarly problematic.

The challenges of training and deploying an LLM

One issue with using LLMs is their tremendous operating expense because the computational demand to train and run them is so intense (they’re not called large language models for nothing).

First, the hardware to run the models on is costly. The H100 GPU from Nvidia, a popular choice for LLMs, has been selling on the secondary market for about $40,000 per chip. One source estimated it would take roughly 6,000 chips to train an LLM comparable to ChatGPT-3.5. That’s roughly $240 million on GPUs alone.

Another significant expense is powering those chips. Merely training a model is estimated to require about 10 gigawatt-hours (GWh) of power, equivalent to 1,000 U.S. homes’ yearly electrical use. Once the model is trained, its electricity cost will vary but can get exorbitant. That source estimated that the power consumption to run ChatGPT-3.5 is about 1 GWh a day, or the combined daily energy usage of 33,000 households.

Power consumption can also be a potential pitfall for user experience when running LLMs on portable devices. That’s because heavy use on a device could drain its battery very quickly, which would be a significant barrier to consumer adoption.

Integrating LLMs into devices presents another critical challenge to the user experience: effective communication between the LLM and the device. If the channel has a high latency, users will be frustrated by long lags between queries and responses.

Finally, privacy is a crucial component of offering an LLM-based service that conforms to privacy regulations that customers want to use. Given that LLMs tend to memorize their training data, there is a risk of exposing sensitive data when users query the model. User interactions are also logged, which means that users’ questions — sometimes containing private information — may be vulnerable to acquisition by hackers.

The threat of data theft is not merely theoretical; several feasible backdoor attacks on LLMs are already under scrutiny. So, it’s unsurprising that over 75% of enterprises are holding off on adopting LLMs out of privacy concerns.

For all the above reasons, including bankrupting their companies or creating catastrophic reputational damage, business leaders are concerned about taking advantage of the early days of LLMs. To succeed, they must approach things holistically because the challenges need to be simultaneously conquered before launching a viable LLM-based offering.

It’s often difficult to know where to start. Here are five crucial points tech leaders and startup founders should consider when planning a transition to LLMs:

1. Keep an eye out for new hardware optimizations

Although training and running an LLM is expensive now, market competition is already driving innovations that reduce power consumption and boost efficiency, which should reduce costs. One of these solutions is Qualcomm’s Cloud AI 100. The organization claims it’s designed for “deep learning with low power consumption.”

Leaders need to empower management to stay abreast of developments in hardware to reduce power consumption and, therefore, costs. What may not be within reach currently may soon become feasible with the next wave of efficiency breakthroughs.

2. Explore a distributed data analysis approach

Sometimes the infrastructure supporting an LLM could combine edge and cloud computing for distributed data analysis. This would be appropriate for several use cases, such as when one has critical and highly time-sensitive data on an edge device while leaving less time-sensitive data to be processed in the cloud. This approach enables much lower latency for users interacting with the LLM than if all computations were done in the cloud.

On the other hand, offloading computations to the cloud will help preserve a device’s battery power, so there are critical trade-offs to consider with a distributed data analysis approach. Decision-makers must determine the optimized proportion of computations done by each processor given the needs at that moment.

3. Stay flexible regarding which model to use

It’s essential to be flexible on which underlying model to use in building a vertical LLM because each has its pros and cons for any particular use case. That flexibility should not only be at the outset when selecting a model but should also remain a critical factor throughout the use of the model, as needs could change. In particular, open source options are worth considering because these models can be smaller and less expensive.

Building an infrastructure that can accommodate switching to a new model without operational disruption is essential. Some companies now offer “multi-LLM” solutions, such as Merlin, whose DiscoveryPartner generative AI platform uses LLMs from OpenAI, Microsoft, and Anthropic for document analysis.

4. Make data privacy a priority

In an era of increasing regulation for data and data breaches, data privacy must be a priority. One approach is to use sandboxing, in which a controlled computational environment confines data to a restricted system.

Another is data obfuscation (such as with data masking, tokenization, or encryption), which allows the LLM to understand the data while making it unintelligible to anyone who might tap into it. These and other techniques can assure users that privacy is baked into your LLMs.

5. Looking ahead, consider analog computing

An even more radical approach to deploying hardware for LLMs is to move away from digital computing. Once considered more of a curiosity in the IT world, analog computing could ultimately prove to be a boon to LLM adoption because it could reduce the energy consumption required to train and run LLMs.

This is more than just theoretical. For example, IBM has been developing an “analog AI” chip that could be 40 to 140 times more energy efficient than GPUs for training LLMs. As similar chips enter the market from competing vendors, we will see market forces bring down their prices.

The LLM future is here — are you ready?

LLMs are exciting, but developing and adopting them requires overcoming several feasibility hurdles. Fortunately, an increasing number of tools and approaches are bringing down costs, making systems more challenging to hack and ensuring a positive user experience.

So, don’t hesitate to explore how LLMs might turbocharge your business. With the right approach, your organization can be well positioned to take advantage of everything this new era offers. You’ll be glad you got started now.

More TechCrunch

Presti is using GenAI to replace costly furniture industry photo shoots

Romain Dillet

2 hours ago

If you’ve ever bought a sofa online, have you thought about the homes you can see in the background of the product shots? When it’s time to release a new…

Presti is using GenAI to replace costly furniture industry photo shoots

Startups

Google backs Indian open-source Uber rival

Manish Singh

3 hours ago

Google has joined investors backing Moving Tech, the parent firm of open-source ride-sharing app Namma Yatri in India that is eroding market share from Uber and Ola with its no-commission…

Google backs Indian open-source Uber rival

Apps

At last, Apple’s Messages app will support RCS and scheduling texts

Sarah Perez

5 hours ago

These messaging features, announced at WWDC 2024, will have a significant impact on how people communicate every day.

At last, Apple’s Messages app will support RCS and scheduling texts

Apps

Here are all the devices compatible with iOS 18

Lauren Forristal

9 hours ago

iOS 18 will be available in the fall as a free software update.

Here are all the devices compatible with iOS 18

Commerce

TikTok glitch allows Shop to appear to users under 18, despite adults-only policy

Sarah Perez

9 hours ago

The tests indicate there are loopholes in TikTok’s ability to apply its parental controls and policies effectively in a situation where the teen user originally lied about their age, as…

TikTok glitch allows Shop to appear to users under 18, despite adults-only policy

Startups

Lhoopa raises $80M to spur more affordable housing in the Philippines

Jagmeet Singh

9 hours ago

Lhoopa has raised $80 million to address the lack of affordable housing in Southeast Asian markets, starting with the Philippines.

Lhoopa raises $80M to spur more affordable housing in the Philippines

Venture

Trump’s VP candidate JD Vance has long ties to Silicon Valley, and was a VC himself

Marina Temkin

9 hours ago

Former President Donald Trump picked Ohio Senator J.D. Vance as his running mate on Monday, as he runs to reclaim the office he lost to President Joe Biden in 2020.…

Trump’s VP candidate JD Vance has long ties to Silicon Valley, and was a VC himself

Space

TechCrunch Space: Space cowboys

Aria Alamalhodaei

11 hours ago

Hello and welcome back to TechCrunch Space. Is it just me, or is the news cycle only accelerating this summer?!

Apps

Without Apple Intelligence, iOS 18 beta feels like a TV show that’s waiting for the finale

Ivan Mehta

11 hours ago

Apple Intelligence features are not available in the developer beta, which is out now.

Without Apple Intelligence, iOS 18 beta feels like a TV show that’s waiting for the finale

Apps

Apple’s public betas for iOS 18 are here to test out

Maxwell Zeff

11 hours ago

Apple released the public betas for its next generation of software on the iPhone, Mac, iPad and Apple Watch on Monday. You can now test out iOS 18 and many…

Apple’s public betas for iOS 18 are here to test out

Transportation

Fisker has one major objector to its Ocean SUV fire sale

Sean O'Kane

13 hours ago

One major dissenter threatens to upend Fisker’s apparent best chance at offloading its unsold EVs, a deal that would keep the startup’s bankruptcy proceeding alive and pave the way for…

Fisker has one major objector to its Ocean SUV fire sale

Venture

Major Stripe investor Sequoia confirms $70B valuation, offers its investors a payday

Mary Ann Azevedo

14 hours ago

Payments giant Stripe has delayed going public for so long that its major investor Sequoia Capital is getting creative to offer returns to its limited partners. The venture firm emailed…

Major Stripe investor Sequoia confirms $70B valuation, offers its investors a payday

Security

Google’s Kurian approached Wiz, $23B deal could take a week to land, source says

Ingrid Lunden

Marina Temkin

14 hours ago

Alphabet, Google’s parent company, is in advanced talks to acquire Wiz for $23 billion, a person close to the company told TechCrunch. The deal discussions were previously reported by The…

Google’s Kurian approached Wiz, $23B deal could take a week to land, source says

Hardware

Bird Buddy’s new AI feature lets people name and identify individual birds

Brian Heater

14 hours ago

Name That Bird determines individual members of a species by identifying distinguishing characteristics that most humans would be hard-pressed to spot.

Bird Buddy’s new AI feature lets people name and identify individual birds

Apps

YouTube Music is testing an AI-generated radio feature and adding a song recognition tool

Aisha Malik

14 hours ago

YouTube Music is introducing two new ways to boost song discovery on its platform. YouTube announced on Monday that it’s experimenting with an AI-generated conversational radio feature, and rolling out…

Transportation

Elon Musk confirms Tesla ‘robotaxi’ event delayed due to design change

Sean O'Kane

15 hours ago

Tesla had internally planned to build the dedicated robotaxi and the $25,000 car, often referred to as the Model 2, on the same platform.

Elon Musk confirms Tesla ‘robotaxi’ event delayed due to design change

Space

Moon cave! Discovery could redirect lunar colony and startup plays

Devin Coldewey

15 hours ago

What this means for the space industry is that theory has become reality: The possibility of designing a habitation within a lunar tunnel is a reasonable proposition.

Moon cave! Discovery could redirect lunar colony and startup plays

TechCrunch Disrupt 2024

Disrupt Deal Days are here: Prime savings for TechCrunch Disrupt 2024!

TechCrunch Events

18 hours ago

Get ready for a prime week of savings at TechCrunch Disrupt 2024 with the launch of Disrupt Deal Days! From now to July 19 at 11:59 p.m. PT, we’re going…

Disrupt Deal Days are here: Prime savings for TechCrunch Disrupt 2024!

Apps

Deezer chases Spotify and Amazon Music with its own AI playlist generator

Aisha Malik

18 hours ago

Deezer is the latest music streaming app to introduce an AI playlist feature. The company announced on Monday that a select number of paid users will be able to create…

Deezer chases Spotify and Amazon Music with its own AI playlist generator

Fintech

Caliza lands $8.5 million to bring real-time money transfers to Latin America using USDC

Anna Heim

20 hours ago

Real-time payments are becoming commonplace for individuals and businesses, but not yet for cross-border transactions. That’s what Caliza is hoping to change, starting with Latin America. Founded in 2021 by…

Caliza lands $8.5 million to bring real-time money transfers to Latin America using USDC

Adaptive builds automation tools to speed up construction payments

Kyle Wiggers

20 hours ago

Adaptive is a platform that provides tools designed to simplify payments and accounting for general construction contractors.

Adaptive builds automation tools to speed up construction payments

Transportation

How VanMoof’s new owners plan to win over its old customers

Rebecca Bellan

24 hours ago

When VanMoof declared bankruptcy last year, it left around 5,000 customers who had preordered e-bikes in the lurch. Now VanMoof is up and running under new management, and the company’s…

How VanMoof’s new owners plan to win over its old customers

Climate

Mitti Labs aims to make rice farming less harmful to the climate, starting in India

Jagmeet Singh

1 day ago

Mitti Labs aims to transform rice farming in India and other South Asian markets by reducing methane emissions by 50% and water consumption by 30%.

Mitti Labs aims to make rice farming less harmful to the climate, starting in India

Security

How to tell if your online accounts have been hacked

Lorenzo Franceschi-Bicchierai

2 days ago

This is a guide on how to check whether someone compromised your online accounts.

How to tell if your online accounts have been hacked

The AI financial results paradox

Ron Miller

2 days ago

There is a general consensus today that generative AI is going to transform business in a profound way, and companies and individuals who don’t get on board will be quickly…

Security

Google reportedly in talks to acquire cloud security company Wiz for $23B

Anthony Ha

2 days ago

Google’s parent company Alphabet might be on the verge of making its biggest acquisition ever. The Wall Street Journal reports that Alphabet is in advanced talks to acquire Wiz for…

Google reportedly in talks to acquire cloud security company Wiz for $23B

Featured Article

Hank Green reckons with the power — and the powerlessness — of the creator

Hank Green has had a while to think about how social media has changed us. He started making YouTube videos in 2007 with his brother, novelist John Green, at a time when the first iPhone was in development, Myspace was still relevant and Instagram didn’t exist. Seventeen years later, posting…

Amanda Silberling

2 days ago

Hank Green reckons with the power — and the powerlessness — of the creator

Fintech

Synapse’s collapse has frozen nearly $160M from fintech users — here’s how it happened

Mary Ann Azevedo

2 days ago

Here is a timeline of Synapse’s troubles and the ongoing impact it is having on banking consumers.

Synapse’s collapse has frozen nearly $160M from fintech users — here’s how it happened

Featured Article

Helixx wants to bring fast-food economics and Netflix pricing to EVs

When Helixx co-founder and CEO Steve Pegg looks at Daisy — the startup’s 3D-printed prototype delivery van — he sees a second chance. And he’s pulling inspiration from McDonald’s to get there. The prototype, which made its global debut this week at the Goodwood Festival of Speed, is an interesting proof…

Tim Stevens

2 days ago

Helixx wants to bring fast-food economics and Netflix pricing to EVs

Featured Article

India clings to cheap feature phones as brands struggle to tap new smartphone buyers

India is struggling to get new smartphone buyers, as millions of Indians don’t go for an upgrade and continue to be on feature phones.

Jagmeet Singh

2 days ago

5 steps to ensure startups successfully deploy LLMs

Lu Zhang

The challenges of training and deploying an LLM

1. Keep an eye out for new hardware optimizations

2. Explore a distributed data analysis approach

3. Stay flexible regarding which model to use

4. Make data privacy a priority

5. Looking ahead, consider analog computing

The LLM future is here — are you ready?

More TechCrunch

Get the industry’s biggest tech news

Tags