About
Positron builds purpose-built hardware and software for Transformer-model inference, including Atlas inference appliances and custom Asimov chips powering Titan systems. It sells to enterprises and research teams seeking faster, lower-power, lower-total-cost generative-AI inference, differentiated by an integrated architecture optimized specifically for Transformer workloads.
Market
Positron competes in AI infrastructure and accelerator silicon, specifically the rapidly expanding market for generative-AI and Transformer-model inference. It positions itself as an inference-first alternative to general-purpose GPUs, differentiating through claimed advantages in performance per watt, performance per dollar, memory capacity, air-cooled deployment, and reduced infrastructure requirements.
Positron targets enterprises, research teams, hyperscale cloud providers, data-center operators, and AI application companies in areas such as networking, gaming, content moderation, CDN, token-as-a-service, and financial trading. Its likely buyers are CTOs, infrastructure leaders, and CFOs seeking high-throughput inference with lower power consumption and total cost of ownership, particularly where liquid cooling or high rack density is impractical.
At a Glance
Problem
Positron addresses the cost, power, memory, and latency bottlenecks of running generative-AI models in production. As inference demand grows, serving a model—not training it—can become the dominant operating expense and a constraint on data-center capacity. Positron’s thesis is that transformer inference is increasingly limited by memory bandwidth and capacity rather than theoretical compute, making conventional GPU infrastructure inefficient for high-volume workloads.
The killer use cases are long-context language models, agentic workflows, enterprise copilots, video and media models, and latency-sensitive workloads such as trading. The economic objective is lower cost per token, lower power consumption, and better performance per dollar, allowing cloud providers and enterprises to deploy more inference within existing energy and hardware constraints.
Product / Service
Positron sells purpose-built inference infrastructure, centered on Atlas, a production-ready inference appliance that supports models with up to 500 billion parameters and is shipping now. Its software workflow is designed to avoid a complex new compiler stack: customers can upload trained models from the Hugging Face Transformers ecosystem, have Positron map them directly onto the hardware, and call them through an OpenAI-compatible API endpoint. This makes the offering an integrated hardware-and-software deployment platform rather than simply a chip.
Atlas is already positioned for production workloads, with the company citing substantially better performance per dollar and lower power consumption than NVIDIA’s H100. Positron’s next-generation Asimov accelerator, planned for 2027, takes a memory-first approach with up to 2.3 terabytes of memory per chip and a stated target of five times NVIDIA Rubin’s tokens per dollar and tokens per watt. The benefit is faster, more energy-efficient inference with lower infrastructure overhead, particularly for models whose working set and context window overwhelm conventional accelerator memory.
Market
Positron competes in AI infrastructure, specifically inference accelerators and purpose-built AI servers, against NVIDIA’s GPUs and specialized-chip companies such as Cerebras, FuriosaAI, EnCharge AI, Groq, and Lightmatter. Its differentiation is an inference-focused, memory-centric architecture and an appliance-based deployment model aimed at performance-per-dollar and performance-per-watt rather than maximum general-purpose training capacity.
The company is not pre-revenue in the practical sense: Atlas is shipping, is deployed in production environments, and Positron has publicly identified Parasail with SnapServe and Cloudflare as early customers; Jump Trading also became a customer before co-leading its Series B. Positron reported a $51.6 million Series A in July 2025 and a $230 million Series B in February 2026 at a valuation above $1 billion, while saying it had multiple frontier-customer programs and expected strong revenue growth in 2026. Public materials do not disclose a revenue figure, but the combination of shipments, production deployments, named customers, and follow-on financing indicates meaningful early commercial traction.
Founders & Leadership
Funding History
Not publicly disclosed
Valor Equity Partners, Atreides Management, DFJ Growth
ARENA Private Wealth, Jump Trading, Unless
Recent News
Positron AI was reported to be in talks to raise $750 million at an approximately $5 billion valuation. The proposed round would significantly exceed the company’s February 2026 Series B valuation.
i3D.net and Positron AI announced a partnership to combine Positron’s purpose-built inference hardware with i3D.net’s infrastructure. The initiative is launching first in Europe and focuses on sovereign AI inference.
The article examines Positron’s plans following its $230 million Series B, including its near-$1 billion post-money valuation and ambitions for rapid growth in AI inference hardware.
Arm identified Positron among the companies showing additional commercial momentum around its Arm AGI CPU ecosystem. The announcement places Positron alongside other ecosystem participants including Cerebras, Cloudflare, OpenAI, and SAP.
Coverage from Chiplets USA 2026 highlighted Positron’s $230 million financing, which pushed the AI chip company’s valuation beyond $1 billion as demand for inference hardware accelerates.
Jon Peddie Research covered Positron’s $230 million financing and its inference technology, noting that the company’s systems use Intel Agilex 7 M-Series FPGAs.
The funding round raised $230 million for Positron AI and valued the company at more than $1 billion. Qatar Investment Authority was among the strategic participants in the oversubscribed Series B.
Positron announced a $230 million Series B at a valuation above $1 billion. The funding is intended to support Atlas systems shipping today and the next-generation Asimov chip, targeted for tape-out in late 2026 and production in early 2027.
SiliconANGLE reported that Positron raised an oversubscribed $230 million round at a $1 billion valuation. The coverage described Atlas as the company’s current inference system and Asimov as its forthcoming custom silicon platform.
Positron’s Asimov product page describes custom accelerator silicon with up to 2.3TB of memory per chip, claimed fivefold advantages in tokens per dollar and watt versus NVIDIA Rubin, and an approximately 400W air-cooled design. The product is listed as coming in 2027.
Active Roles
17Business Model
Positron uses a hardware-led model, selling or deploying Atlas inference appliances and developing Titan systems and Asimov accelerator silicon for enterprise and research workloads. Its hardware is accompanied by model-management and OpenAI-compatible API software; public materials do not disclose a detailed price list or subscription structure.
Products
Customers
Tech Stack
Similar Companies
Competitors
Key Investors
ARENA Private Wealth, Jump Trading, Unless, Qatar Investment Authority, Arm, Helena, Valor Equity Partners, Atreides Management, DFJ Growth, Resilience Reserve, Flume Ventures, 1517 Fund