AI

This Week in AI: Companies are growing skeptical of AI’s ROI

Comment

Big Data concept picture with ones and zeros going off into infinity.
Image Credits: Darpa under a Public Domain (opens in a new window) license.

Hiya, folks, welcome to TechCrunch’s regular AI newsletter.

This week in AI, Gartner released a report suggesting that around a third of generative AI projects in the enterprise will be abandoned after the proof-of-concept phase by year-end 2025. The reasons are many — poor data quality, inadequate risk controls, escalating infrastructure costs and so on.

But one of the biggest barriers to generative AI adoption is the unclear business value, per the report.

Embracing generative AI organization-wide comes with significant costs, ranging from $5 million to a whopping $20 million, estimates Gartner. A simple coding assistant has an upfront cost between $100,000 and $200,000 and recurring costs upward of $550 per user per year, while an AI-powered document search tool can cost $1 million upfront and between $1.3 million and $11 million per user annually, finds the report.

Those steep price tags are hard for corporations to swallow when the benefits are difficult to quantify and could take years to materialize — if, indeed, they ever materialize.

A survey from Upwork this month reveals that AI, rather than enhancing productivity, has actually proven to be a burden for many of the workers using it. According to the survey, which interviewed 2,500 C-suite execs, full-time staffers and freelancers, nearly half (47%) of workers using AI say that they have no idea how to achieve the productivity gains their employers expect while over three-fourths (77%) believe that AI tools have decreased productivity and added to their workload in at least one way.

It seems the honeymoon phase of AI may well be ending, despite robust activity on the VC side. And that’s not shocking. Anecdote after anecdote reveals how generative AI, which has unsolved fundamental technical issues, is frequently more trouble than it’s worth.

Just Tuesday, Bloomberg published a piece about a Google-powered tool that uses AI to analyze patient medical records, now in testing at HCA hospitals in Florida. Users of the tool Bloomberg spoke with said that it can’t consistently deliver reliable health information; in once instance, it failed to note whether a patient had any drug allergies.

Companies are beginning to expect more of AI. Barring research breakthroughs that address the worst of its limitations, it’s incumbent on vendors to manage expectations.

We’ll see if they have the humility to do so.

News

SearchGPT: OpenAI last Thursday announced SearchGPT, a search feature designed to give “timely answers” to questions, drawing from web sources.

Bing gets more AI: Not to be outdone, Microsoft last week previewed its own AI-powered search experience, called Bing generative search. Available for only a “small percentage” of users at the moment, Bing generative search — like SearchGPT — aggregates info from around the web and generates a summary in response to search queries.

X opts users in: X, formerly Twitter, quietly pushed out a change that appears to default user data into its training pool for X’s chatbot Grok, a move that was spotted by users of the platform on Friday. EU regulators and others quickly cried foul. (Wondering how to opt out? Here’s a guide.)

EU calls for help with AI: The European Union has kicked off a consultation on rules that will apply to providers of general-purpose AI models under the bloc’s AI Act, its risk-based framework for regulating applications of AI.

Perplexity details publisher licensing: AI search engine Perplexity will soon start sharing advertising revenue with news publishers when its chatbot surfaces their content in response to a query, a move that appears to be designed to assuage critics that’ve accused Perplexity of plagiarism and unethical web scraping. 

Meta rolls out AI Studio: Meta said Monday that it’s rolling out its AI Studio tool to all creators in the U.S. to let them make personalized AI-powered chatbots. The company first unveiled AI Studio last year and started testing it with select creators in June.

Commerce Department endorses “open” models: The U.S. Commerce Department on Monday issued a report in support of “open-weight” generative AI models like Meta’s Llama 3.1, but recommended the government develop “new capabilities” to monitor such models for potential risks.

$99 Friend: Avi Schiffmann, a Harvard dropout, is working on a $99 AI-powered device called Friend. As the name suggests, the neck-worn pendant is designed to be treated as a companion of sorts. But it’s not clear yet whether it works quite as advertised.

Research paper of the week

Reinforcement learning from human feedback (RLHF) is the dominant technique for ensuring that generative AI models follow instructions and adhere to safety guidelines. But RLHF requires recruiting a large number of people to rate a model’s responses and provide feedback, a time-consuming and expensive process.

So OpenAI is embracing alternatives.

In a new paper, researchers at OpenAI describe what they call rule-based rewards (RBRs), which use a set of step-by-step rules to evaluate and guide a model’s responses to prompts. RBRs break down desired behaviors into specific rules that are then used to train a “reward model,” which steers the AI — “teaching” it, in a sense — about how it should behave and respond in specific situations.

OpenAI claims that RBR-trained models demonstrate better safety performance than those trained with human feedback alone while reducing the need for large amounts of human feedback data. In fact, the company says it’s used RBRs as part of its safety stack since the launch of GPT-4 and plans to implement RBRs in future models.

Model of the week

Google’s DeepMind is making progress in its quest to tackle complex math problems with AI.

A few days ago, DeepMind announced that it trained two AI systems to solve four out of the six problems from this year’s International Mathematical Olympiad (IMO), the prestigious high school math competition. DeepMind claims the systems, AlphaProof and AlphaGeometry 2 (the successor to January’s AlphaGeometry), demonstrated an aptitude for forming and drawing on abstractions and complex hierarchical planning — all of which have been historically challenging for AI systems to do.

AlphaProof and AlphaGeometry 2 worked together to solve two algebra problems and a number theory problem. (The two remaining questions on combinatorics were left unsolved). The results were verified by mathematicians; it’s the first time AI systems have been able to achieve silver medal-level performance on IMO questions.

There are a few caveats, however. It took days for the models to solve some of the problems. And while their reasoning capabilities are impressive, AlphaProof and AlphaGeometry 2 can’t necessarily help with open-ended problems that have many possible solutions, unlike those with one right answer.

We’ll see what the next generation brings.

Grab bag

AI startup Stability AI has released a generative AI model that turns a video of an object into multiple clips that look as though they were captured from different angles.

Called Stable Video 4D, the model could have applications in game development and video editing, Stability says, as well as virtual reality. “We anticipate that companies will adopt our model, fine-tuning it further to suit their unique requirements,” the company wrote in a blog post.

Stability AI Stable Video 4D
Image Credits: Stability AI

To use Stable Video 4D, users upload footage and specify their desired camera angles. After about 40 seconds, the model then generates eight five-frame videos (although “optimization” can take another 25 minutes).

Stability says that it’s actively working on refining the model, optimizing it to handle a wider range of real-world videos beyond the current synthetic datasets it was trained on. “The potential for this technology in creating realistic, multi-angle videos is vast, and we are excited to see how it will evolve with ongoing research and development,” the company continued.

More TechCrunch

The pharma giant won’t say how many patients were affected by its February data breach. A count by TechCrunch confirms that over a million people are affected.

Pharma giant Cencora is alerting millions about its data breach

Self-driving technology company Aurora Innovation is looking to raise hundreds of millions in additional capital as it races toward a driverless commercial launch by the end of 2024.  Aurora is…

Self-driving truck startup Aurora Innovation to sell up to $420M in shares ahead of commercial launch

Payments infrastructure firm Infibeam Avenues has acquired a majority 54% stake in Rediff.com for up to $3 million, a dramatic twist of fate for the 28-year-old business that was the…

Rediff, once an internet pioneer in India, sells majority stake for $3M

The ruling confirmed an earlier decision in April from the High Court of Podgorica which rejected a request to extradite the crypto fugitive to the United States.

Terraform Labs co-founder and crypto fugitive Do Kwon set for extradition to South Korea

A day after Meta CEO Mark Zuckerberg talked about his newest social media experiment Threads reaching “almost” 200 million users on the company’s Q2 2024 earnings call, the platform has…

Meta’s Threads crosses 200 million active users

TechCrunch Disrupt 2024 will be in San Francisco on October 28–30, and we’re already excited! Disrupt brings innovation for every stage of your startup journey, and we could not bring you this…

Connect with Google Cloud, Aerospace, Qualcomm and more at Disrupt 2024

Featured Article

A comprehensive list of 2024 tech layoffs

The tech layoff wave is still going strong in 2024. Following significant workforce reductions in 2022 and 2023, this year has already seen 60,000 job cuts across 254 companies, according to independent layoffs tracker Layoffs.fyi. Companies like Tesla, Amazon, Google, TikTok, Snap and Microsoft have conducted sizable layoffs in the…

A comprehensive list of 2024 tech layoffs

Intel announced it would layoff more than 15% of its staff, or 15,000 employees, in a memo to employees on Thursday. The massive headcount is part of a large plan…

Intel to lay off 15,000 employees

Following the recent lawsuit filed by the Recording Industry Association of America (RIAA) against music generation startups Udio and Suno, Suno admitted in a court filing on Thursday that it did, in…

AI music startup Suno claims training model on copyrighted music is ‘fair use’

In spite of a drop for the quarter, iPhone remained Apple’s most important category by a wide margin.

iPad sales help bail out Apple amid a continued iPhone slide

Molly Alter wears a lot of hats. She’s a mocumentary filmmaker working on a project about an alternate reality where charades is big business. She’s a caesar salad connoisseur and…

How filming a cappella concerts and dance recitals led Northzone’s newest partner Molly Alter to a career in VC

Microsoft has a long and tangled history with OpenAI, having invested a reported $13 billion in the ChatGPT maker as part of a long-term partnership. As part of the deal,…

Microsoft now lists OpenAI as a competitor in AI and search

The San Jose-based startup raised $60 million in a round that values it lower than the $500 million valuation it garnered in its most recent round, according to multiple sources.

Sequoia-backed Knowde raises Series C at a valuation cut

X (formerly Twitter) can no longer be accessed in the Mac App Store, suggesting that it has been officially delisted.  Searches for both “Twitter” and “X” on Apple’s platform no…

Twitter disappears from Mac App Store

Google Thursday said that it is introducing new Gemini-powered features for Chrome’s desktop version, including Lens for desktop, tab compare for shopping assistance, and natural language integration for search history.…

Google brings Gemini-powered search history and Lens to Chrome desktop

When Xiaoyin Qu was growing up in China, she was obsessed with learning how to build paper airplanes that could do flips in the air. Her parents, though, didn’t have…

Heeyo built an AI chatbot to be a billion kids’ interactive tutor and friend

While the company was awarded a massive, $4.2 billion contract to accelerate Starliner development in 2014, it was structured as a “fixed-price” model.

Boeing bleeds another $125M on Starliner program, bringing total losses to $1.6B

Welcome back to TechCrunch Mobility — your central hub for news and insights on the future of transportation. Sign up here for free — just click TechCrunch Mobility! Summer road…

Anthony Levandowski bets on off-road autonomy, Nuro plots a comeback and Applied Intuition gets more investor love

Google’s new features include Gemini in BigQuery and Looker to help users with data engineering and analysis.

Google Cloud expands its database portfolio with new AI capabilities

Rad Power Bikes, the Seattle-based e-bike startup that has raised more than $300 million from investors, went through another round of layoffs in July, TechCrunch has exclusively learned. This is…

VC darling Rad Power Bikes hit with another round of layoffs

Five years ago, as robotaxis and self-driving truck startups were still raking in millions in venture capital, Anthony Levandowski turned to off-road autonomy. Now, that decision — which brought the…

Why Anthony Levandowski returned to his off-road autonomous vehicle roots with AV startup Pronto

Commercial space station company Vast is building a private microgravity research lab as part of its wider Haven-1 station plans. The module is set to launch no earlier than the…

Vast plans microgravity lab on its Haven-1 private space station

Google Cloud is giving Y Combinator startups access to a dedicated, subsidized cluster of Nvidia graphics processing units and Google tensor processing units to build AI models. It’s part of…

Google Cloud now has a dedicated cluster of Nvidia GPUs for Y Combinator startups

StackShare is one of the more popular platforms for developers to discuss, track, and share the tools they use to build applications.

Open source startup FOSSA is buying StackShare, a site used by 1.5M developers

Featured Article

Indian startups gut valuations ahead of IPO push

Ola Electric and FirstCry are set to test investor appetite with public listing, both pricing their shares below their previous valuation asks.

Indian startups gut valuations ahead of IPO push

The European Union’s risk-based regulation for applications of artificial intelligence has come into force starting from today.

The EU’s AI Act is now in force

The company also said it has received regulatory clearance to start Phase 2 clinical trials for a new drug in the U.S. later this year.

Healx, an AI-enabled drug discovery platform for rare diseases, raises $47M

The European Commission (EC) has given the go-ahead to HPE’s planned megabucks acquisition of Juniper Networks.

EU greenlights HPE’s $14B Juniper Networks acquisition

Meta, which develops one of the biggest foundational open source large language models, Llama, believes it will need significantly more computing power to train models in the future. Mark Zuckerberg…

Zuckerberg says Meta will need 10x more computing power to train Llama 4 than Llama 3

Axle Energy is a B2B, back-end infrastructure business focused on connecting flexible assets, such as electric vehicles and home batteries, to energy markets that aren’t otherwise available for consumers to…

Axle Energy’s sprint to decarbonize the grid lights up with $9M seed led by Accel