Now Valued at $3 Billion, Gimlet Labs Raises $300 Million in Series B Led by Andreessen Horowitz for Industry’s First Multi-Silicon Inference Cloud for Agentic AI
SAN FRANCISCO, Sept. 04, 2026 (GLOBE NEWSWIRE) -- Gimlet Labs, the Applied AI research and product company, today announced that it has raised $300 million in Series B funding, bringing its current valuation to $3 billion and the total funding raised to $392 million. This round was led by Andreessen Horowitz with participation from major investor Sapphire Ventures, other new investors M12 and Arm and previous investors Menlo Ventures and Factory.
“AI demand is growing exponentially, while data centers and silicon can’t keep pace. The answer isn’t just more infrastructure - it’s a better architecture. Gimlet has built a new kind of inference cloud, heterogeneous by design, that matches each workload to the right silicon. By making GPUs and purpose-built accelerators work as one system, Gimlet delivers dramatically more throughput, more interactivity and more intelligence from every watt. We believe this is where inference infrastructure is headed,” said Raghu Raghuram, managing partner at Andreessen Horowitz and Gimlet Labs board member.
The explosion of agentic workloads has exposed a critical limitation in today’s AI infrastructure: homogeneous hardware alone cannot meet the speed and efficiency needs of agentic workloads. With inference now reaching quadrillions of tokens per month, and still growing, the one-size-fits-all approach leaves massive inefficiencies in performance and utilization, even as the industry gears up to spend an estimated $765 billion in AI CapEx this year, with 7.6 trillion in cumulative spending from 2026-2031, according to Goldman Sachs.*
In October 2025, Gimlet Labs emerged from stealth to solve this problem with the industry's first multi-silicon inference software designed for faster, more interactive agentic AI workloads. Gimlet disaggregates AI models to run each phase of inference on the most appropriate silicon, improving both inference latency and throughput, and is available through its own managed cloud or as a managed service in customer environments. Gimlet Labs works with leading AI chip companies, including NVIDIA, AMD, Intel, Arm, Cerebras and d-Matrix to support their chips in its multi-silicon architecture.
In March 2026, Gimlet Labs announced that it tripled its customer base and added one of the top three frontier labs as well as one of the top three hyperscalers as customers.
Since then, Gimlet Labs has secured billions of dollars in contracted revenue for Gimlet Cloud and is scaling to hundreds of megawatts in managed heterogeneous infrastructure. The new capital will be used to build out operations for its multi-silicon cloud as well as to continue to expand its team.
“We’ve reached a turning point where inference is the dominant AI workload and the demand for tokens is explosive. With Gimlet, our customers are able to serve massive volumes of tokens at very low latency, even as their AI workloads continue to grow across all dimensions. We’re able to deliver unprecedented performance because Gimlet software intelligently slices and orchestrates their workloads across different types of hardware from both mainstream and emerging chipmakers,” said Zain Asgar, co-founder and CEO of Gimlet Labs.
Gimlet is the only AI cloud built with multi-silicon inference software and heterogeneous infrastructure and combines:
- Multi-silicon datacenters that integrate GPUs, purpose-built AI accelerators including SRAM-based architectures and CPUs, in one coherent solution
- Inference software that disaggregates AI workloads, matching each phase of the workload to its optimized hardware architecture, achieving up to 10X gains in throughput and interactivity
- Intelligent orchestration of AI workloads across compute nodes
Gimlet’s approach enables dramatically more efficient use of limited hardware and power, resulting in a smaller overall footprint and cost.
About Gimlet Labs
Gimlet Labs’ mission is to drive breakthrough improvements in AI performance that result in massive increases of compute available for AI workloads. Gimlet Labs’ inference cloud is derived from its foundational research across the stack in order to enable the next generation of performant, scalable AI infrastructure. Its research combines theory and practice to push the boundaries of AI efficiency via techniques such as automated GPU kernel generation, workload orchestration and heterogeneous execution across diverse hardware. Gimlet is backed by Andreessen Horowitz and Menlo Ventures and headquartered in San Francisco. For more information, simply visit: https://gimletlabs.ai/.
* May 1, 2026 Goldman Sachs: Tracking Trillions: The Assumptions Shaping the Scale of the AI Build-Out
https://www.goldmansachs.com/insights/articles/tracking-trillions-the-assumptions-shaping-scale-of-the-ai-build-out
Media and Analyst Contact:
Amber Rowland
amber@therowlandagency.com
+1-650-814-4560
Legal Disclaimer:
EIN Presswire provides this news content "as is" without warranty of any kind. We do not accept any responsibility or liability for the accuracy, content, images, videos, licenses, completeness, legality, or reliability of the information contained in this article. If you have any complaints or copyright issues related to this article, kindly contact the author above.
