Tensormesh raises $4.5M to squeeze more inference out of AI server loads

Tensormesh raises $4.5M to squeeze more inference out of AI server loads

In a rapidly evolving landscape where AI infrastructure is paramount, the demand for maximizing inference capabilities from existing GPUs has never been greater. Capitalizing on this trend, Tensormesh has emerged from stealth mode, announcing a successful seed funding round of $4.5 million. The investment, spearheaded by Laude Ventures, also features contributions from notable angel investor Michael Franklin, a pioneer in database technology. The funding will be utilized to develop a commercial variant of the open-source LMCache utility, a project led by co-founder Yihua Cheng. When implemented effectively, LMCache has the potential to cut inference costs by up to 90%, solidifying its status as a vital resource in open-source applications and attracting integrations from industry giants such as Google and Nvidia. At the core of Tensormesh's offering is a key-value cache (KV cache), designed to enhance the processing of complex inputs by streamlining them into fundamental values. In conventional systems, the KV cache is typically discarded after each query, a practice that Tensormesh CEO Juchen Jiang identifies as a significant inefficiency. "It’s akin to having a highly skilled analyst who forgets everything after answering each question," Jiang explains. By retaining the KV cache, Tensormesh's technology allows for its reuse during subsequent queries, effectively optimizing inference power without additional server load. This approach is particularly beneficial for chat interfaces, which require continuous reference to an expanding conversation history, as well as for agentic systems that manage growing logs of actions and objectives. While AI firms could theoretically implement such improvements independently, the complexity involved often proves overwhelming. Tensormesh aims to address this gap by providing a ready-made solution, anticipating robust demand for a product that simplifies the process. "Maintaining the KV cache in a secondary storage system for efficient reuse, without compromising overall system performance, is a complex challenge," Jiang notes. "We've observed companies employing large teams for several months to develop similar systems, or they can opt for our product for a more efficient solution."

Sources : TechCrunch

Published On : Oct 23, 2025, 16:05

Aerospace
SpaceX's Latest Starlink Launch Marks Milestone Amid Booster Setbacks

On Friday, SpaceX marked a significant achievement by successfully launching its first batch of third-generation Starlin...

TechCrunch | Jul 24, 2026, 23:40
SpaceX's Latest Starlink Launch Marks Milestone Amid Booster Setbacks
AI
Anthropic Unveils Claude Opus 5: A Game-Changer in Cost-Effective AI

Anthropic has officially launched its latest AI model, Claude Opus 5, which the company claims is its most efficient and...

CNBC | Jul 24, 2026, 17:20
Anthropic Unveils Claude Opus 5: A Game-Changer in Cost-Effective AI
Cybersecurity
India's Crackdown on Open-Source Messaging App Raises Legal Concerns

The Indian government's recent initiative to enforce restrictions on Jack Dorsey's offline Bluetooth messaging applicati...

TechCrunch | Jul 24, 2026, 17:10
India's Crackdown on Open-Source Messaging App Raises Legal Concerns
AI
Anthropic Unveils Opus 5: A More Accessible and Powerful AI Model

Anthropic has officially launched its latest AI model, Opus 5, marking a significant addition to its lineup. While it is...

TechCrunch | Jul 24, 2026, 17:10
Anthropic Unveils Opus 5: A More Accessible and Powerful AI Model
Computing
Moody's Warns of Credit Risks as Tech Giants Ramp Up AI Investments

According to a recent report from Moody's Ratings, the surge in investment towards artificial intelligence infrastructur...

CNBC | Jul 24, 2026, 17:50
Moody's Warns of Credit Risks as Tech Giants Ramp Up AI Investments
View All News