Positron AI Raises $230 Million Series B at Over $1 Billion Valuation to Scale Energy-Efficient AI Inference
4.2.2026 14:00:00 CET | Business Wire | Press Release
Positron AI, the leader in energy-efficient AI inference hardware, today announced an oversubscribed $230 million Series B financing at a post-money valuation exceeding $1 billion.
This press release features multimedia. View the full release here: https://www.businesswire.com/news/home/20260204250472/en/
Thomas Sohmers (L), CTO and cofounder, and Mitesh Agrawal (R), CEO of Positron AI (Credit: Kavita Agrawal)
The round was co-led by ARENA Private Wealth, Jump Trading, and Unless, and includes new and strategic investment from Qatar Investment Authority (QIA), Arm, and Helena. Existing investors Valor Equity Partners, Atreides Management, DFJ Growth, Resilience Reserve, Flume Ventures, and 1517 also participated. The financing validates Positron's mission to make AI inference dramatically cheaper and more energy-efficient at scale.
“We're grateful for this investor enthusiasm, which itself is a reflection of what the market is demanding,” said Mitesh Agrawal, CEO of Positron AI. “Energy availability has emerged as a key bottleneck for AI deployment. And our next-generation chip will deliver 5x more tokens per watt in our core workloads versus Nvidia’s upcoming Rubin GPU. Memory is the other giant bottleneck in inference, and our next generation Asimov custom silicon will ship with over 2304 GB of RAM per device next year, versus just 384 GB for Rubin. This will be a critical differentiator in workloads including video, trading, multi-trillion parameter models, and anything requiring an enormous context window. We also expect to beat Rubin in performance per dollar for specific memory-intensive workloads.”
Positron is building the infrastructure layer that makes AI usable at scale by lowering the cost and power required to run modern models. The company's shipping product, Atlas, is an inference system designed for rapid deployment and scaling. Atlas is also a fully American-fabricated and manufactured silicon and system, enabling fast production ramp and dependable supply for customers who need capacity quickly.
“Memory bandwidth and capacity are two of the key limiters for scaling AI inference workloads for next-generation models,” said Dylan Patel, founder and CEO of SemiAnalysis, an advisor and investor in Positron. SemiAnalysis is a leading research firm specializing in semiconductors and AI infrastructure that provides detailed insights into the full compute stack. “Positron is taking a unique approach to the memory scaling problem, and with its next-generation Asimov chip, can deliver more than an order of magnitude greater high-speed memory capacity per chip than incumbent or upstart silicon providers.”
Jump Trading Leads After Deploying Atlas
A key highlight of the round is Jump Trading's decision to co-lead after first becoming a customer.
"For the workloads we care about, the bottlenecks are increasingly memory and power—not theoretical compute,” said Alex Davies, Chief Technology Officer of Jump Trading. “In our testing, Positron Atlas delivered roughly 3x lower end-to-end latency than a comparable H100-based system on the inference workloads we evaluated, in an air-cooled, production-ready footprint with a supply chain we can plan around. The deeper we went, the more we agreed with Positron’s roadmap—Asimov and the Titan systems—as a memory-first platform built for future workloads. We invested because Positron combines traction today with a roadmap that can reshape the cost curve and capabilities for inference.”
“Jump Trading came to Positron as a customer,” said Agrawal. “As they saw our roadmap for Asimov, our custom silicon, and Titan, our next-generation system, they chose to step up as a co-lead investor. A customer becoming an investor is one of the strongest validations we can receive. It signals both technical conviction and real-world demand.”
Building Toward Asimov and Titan: A Memory-First Platform for Next-Generation Inference
Positron's next-generation custom silicon, Asimov, is designed around the reality that modern AI workloads are increasingly limited by memory bandwidth and capacity, not just compute flops. Asimov is designed to support 2 terabytes of memory per accelerator and 8 terabytes of memory per Titan system at similar realized memory bandwidth to NVIDIA's next-generation Rubin GPU. At rack scale, this translates to memory capacity of well over 100 terabytes.
“As AI inference scales, efficiency and system design matter more than raw benchmarks,” said Eddie Ramirez, Vice President of Go-to-Market, Cloud AI Business Unit, Arm. “Positron’s memory-centric approach, built on Arm technology, reflects how tightly coupled systems and a broad ecosystem come together to deliver scalable, performance-per-watt gains in next-generation AI infrastructure.”
This memory-first architecture unlocks high-value inference workloads, including long-context large language models, agentic workflows, and next-generation media and video models. Positron is on track to tape out its Asimov chip just 16 months after its June Series A financing gave it the resources to fully launch the design process, and the company intends to maintain this pace with future chips. “To us, development speed is an essential competitive advantage,” said Agrawal. “Competing with Nvidia means matching their shipping frequency, and we have designed our organization around that goal.”
“Positron is solving one of the most important bottlenecks in AI: delivering inference at scale within real-world power and cost constraints,” said Ari Schottenstein, Head of Alternatives at ARENA Private Wealth. “The combination of shipping traction today with Atlas, plus a credible path to Asimov, creates a rare opportunity to define a new category in AI infrastructure.”
Positron is building this platform with an ecosystem of industry leaders, including Arm, Supermicro and other key technology and supply-chain partners.
Momentum and Growth Trajectory
Positron expects strong revenue growth in 2026, positioning the company to become one of the fastest-growing silicon companies ever, achieving large-scale commercial traction in roughly 2.5 years from company launch. The company is working with multiple frontier customers across cloud, advanced computing, and performance-sensitive verticals, and continues to expand deployments and customer programs.
About Positron AI
Positron AI builds purpose-built hardware and software to make AI inference dramatically cheaper and more energy-efficient. Positron's shipping product, Atlas, is designed for rapid, scalable deployment, and the company's next-generation custom silicon, Asimov, targets tape-out toward the end of 2026 with production in early 2027. Positron's systems are built to serve long-context and next-generation AI workloads with leading economics. Learn more at positron.ai.
View source version on businesswire.com: https://www.businesswire.com/news/home/20260204250472/en/
Contacts
Media Contact:
Helen Cho
Bonfire Partners
press@bonfirepartners.io
(c) 2024 Business Wire, Inc., All rights reserved.
Business Wire, a Berkshire Hathaway company, is the global leader in multiplatform press release distribution.
Subscribe to releases from Business Wire
Subscribe to all the latest releases from Business Wire by registering your e-mail address below. You can unsubscribe at any time.
Latest releases from Business Wire
Grid Dynamics Appoints Industry Veteran Nikita Ivanov as Head of AI26.8.2026 22:05:00 CEST | Press Release
Grid Dynamics Holdings, Inc. (Nasdaq: GDYN) (“Grid Dynamics”), a premier AI transformation partner for the Fortune 1000, is advancing its AI strategy under the leadership of Nikita Ivanov, Vice President, Head of AI. Ivanov is responsible for leading the transition efforts of Grid Dynamics’ clients toward an AI-native software delivery model, and he oversees all elements of Grid Dynamics AI Native (GAIN) platforms. Ivanov brings more than two decades of experience founding and leading technology companies to his role as Vice President, Head of AI. He founded and led GridGain, a distributed computing company built for in-memory data processing. During his 15-year tenure at GridGain, Ivanov also founded the natural language processing framework Apache NLPCraft and, later, Humatron AI, an AI workforce platform. He then served as Field CTO for GenAI at DataStax. That same pattern of building foundational tools now runs through Grid Dynamics’ own AI-SDLC tools, Allium, Rosetta, and SpecFlow
Arc'teryx Introduces System 0™, a New Framework for the Future of Product Design and Circularity26.8.2026 15:00:00 CEST | Press Release
Arc'teryx Equipment, the global design company specializing in technical high-performance apparel and equipment, today introduced System 0™, a new framework that reimagines how products can be designed, manufactured, used, repaired, and ultimately regenerated. System 0 represents the next evolution for sustainability in the outdoor industry, as it embeds circular thinking across an entire product lifecycle while maintaining the uncompromising performance standards that define the Arc’teryx brand. This press release features multimedia. View the full release here: https://www.businesswire.com/news/home/20260826686539/en/ Arc’teryx athlete, Craig Murray, wearing the Sperro SV, a multisport mountain-grade hardshell built with circular design principles, available September 4, 2026. The first product being introduced as part of the System 0 platform is the Sperro SV jacket, a multisport mountain-grade hardshell engineered for durability, repairability, and disassembly at end of life – the
Xsolla Announces Five-Year Partnership With the Global Esports Federation26.8.2026 15:00:00 CEST | Press Release
Xsolla, a global commerce company, today announced a landmark five-year strategic partnership with the Global Esports Federation (GEF), the world-convening body for esports. As GEF's exclusive partner for commerce, payments, wallet, loyalty, and distribution, Xsolla will power the entire GEF ecosystem, including the flagship Global Esports Games, pro-series Global Esports Tour, global GEFcon, and 180+ Member Federations across five continents. This press release features multimedia. View the full release here: https://www.businesswire.com/news/home/20260826948844/en/ Graphic: Xsolla The announcement comes as GEF marks 100 days to go to the Los Angeles 2026 Global Esports Games (GEGLA26), adding further momentum to the road to Los Angeles and GEF’s flagship global esports celebration this December. The partnership launches at the GEGLA26, where Xsolla will begin serving as the unified commerce and payments layer for the GEF ecosystem. Across GEF properties, in more than 200 markets, Xso
NABEP Appoints Sara Chouraqui as General Counsel26.8.2026 15:00:00 CEST | Press Release
North American Blue Energy Partners (NABEP) has appointed Sara Chouraqui as General Counsel, with Elizabeth Collery and Victoria Jacobson each joining as Deputy General Counsel. This press release features multimedia. View the full release here: https://www.businesswire.com/news/home/20260826168543/en/ Sara Chouraqui joins NABEP as its General Counsel, bringing decades of legal expertise. These additions significantly strengthen NABEP’s senior leadership team and legal expertise as the company improves its operations, increases production and develops new strategic and commercial partnerships. As General Counsel, Chouraqui will lead NABEP’s global legal function and advise the company’s leadership on its legal, regulatory and corporate priorities, including major transactions and partnerships, governance, compliance and the management of legal and regulatory risk across the company’s operations. Her experience navigating complex cross-border legal and regulatory matters will be particu
Grafana Labs Crosses 10,000-Customer Milestone as AI Adoption Accelerates Growth Across the Platform26.8.2026 15:00:00 CEST | Press Release
Grafana Labs, the company behind the open observability cloud, today announced it has crossed 10,000 customers worldwide and surpassed $600 million in annual recurring revenue, as organizations standardize on open observability for the AI era. Teams need AI to help them understand their own systems faster, especially as complexity and overhead remain the top observability concern cited in Grafana Labs' 2026 Observability Survey – a need reflected by the adoption of Grafana Assistant, Grafana Cloud's context-aware AI agent for exploring and troubleshooting observability data, now used by more than 18,000 organizations. Teams also need observability purpose-built for AI systems themselves: When teams move AI into production, models drift, agents take unexpected actions, token costs spike without warning, and the latency and error signals that caught problems for the last decade don't explain what went wrong. The same survey found 57% of organizations are now implementing LLM observabilit
In our pressroom you can read all our latest releases, find our press contacts, images, documents and other relevant information about us.
Visit our pressroom