Companies

Positron

positron.ai

Positron builds purpose-built hardware and software accelerating generative-AI inference for better performance, efficiency, and cost.

HQReno, Nevada, United States
Founded2023
Employees87
Funding$300M+
Valuation$1B+
Revenue$5.4M estimated ARR
17 active roles
Profile 1mo agoJobs checked 18h ago
AI / MLAI InfrastructureInfrastructureSeries B$200M-$1B

About

Positron builds purpose-built hardware and software for Transformer-model inference, including Atlas inference appliances and custom Asimov chips powering Titan systems. It sells to enterprises and research teams seeking faster, lower-power, lower-total-cost generative-AI inference, differentiated by an integrated architecture optimized specifically for Transformer workloads.

Market

Positron competes in AI infrastructure and accelerator silicon, specifically the rapidly expanding market for generative-AI and Transformer-model inference. It positions itself as an inference-first alternative to general-purpose GPUs, differentiating through claimed advantages in performance per watt, performance per dollar, memory capacity, air-cooled deployment, and reduced infrastructure requirements.

Target Customers

Positron targets enterprises, research teams, hyperscale cloud providers, data-center operators, and AI application companies in areas such as networking, gaming, content moderation, CDN, token-as-a-service, and financial trading. Its likely buyers are CTOs, infrastructure leaders, and CFOs seeking high-throughput inference with lower power consumption and total cost of ownership, particularly where liquid cooling or high rack density is impractical.

At a Glance

Problem

Positron addresses the cost, power, memory, and latency bottlenecks of running generative-AI models in production. As inference demand grows, serving a model—not training it—can become the dominant operating expense and a constraint on data-center capacity. Positron’s thesis is that transformer inference is increasingly limited by memory bandwidth and capacity rather than theoretical compute, making conventional GPU infrastructure inefficient for high-volume workloads.

The killer use cases are long-context language models, agentic workflows, enterprise copilots, video and media models, and latency-sensitive workloads such as trading. The economic objective is lower cost per token, lower power consumption, and better performance per dollar, allowing cloud providers and enterprises to deploy more inference within existing energy and hardware constraints.

Product / Service

Positron sells purpose-built inference infrastructure, centered on Atlas, a production-ready inference appliance that supports models with up to 500 billion parameters and is shipping now. Its software workflow is designed to avoid a complex new compiler stack: customers can upload trained models from the Hugging Face Transformers ecosystem, have Positron map them directly onto the hardware, and call them through an OpenAI-compatible API endpoint. This makes the offering an integrated hardware-and-software deployment platform rather than simply a chip.

Atlas is already positioned for production workloads, with the company citing substantially better performance per dollar and lower power consumption than NVIDIA’s H100. Positron’s next-generation Asimov accelerator, planned for 2027, takes a memory-first approach with up to 2.3 terabytes of memory per chip and a stated target of five times NVIDIA Rubin’s tokens per dollar and tokens per watt. The benefit is faster, more energy-efficient inference with lower infrastructure overhead, particularly for models whose working set and context window overwhelm conventional accelerator memory.

Market

Positron competes in AI infrastructure, specifically inference accelerators and purpose-built AI servers, against NVIDIA’s GPUs and specialized-chip companies such as Cerebras, FuriosaAI, EnCharge AI, Groq, and Lightmatter. Its differentiation is an inference-focused, memory-centric architecture and an appliance-based deployment model aimed at performance-per-dollar and performance-per-watt rather than maximum general-purpose training capacity.

The company is not pre-revenue in the practical sense: Atlas is shipping, is deployed in production environments, and Positron has publicly identified Parasail with SnapServe and Cloudflare as early customers; Jump Trading also became a customer before co-leading its Series B. Positron reported a $51.6 million Series A in July 2025 and a $230 million Series B in February 2026 at a valuation above $1 billion, while saying it had multiple frontier-customer programs and expected strong revenue growth in 2026. Public materials do not disclose a revenue figure, but the combination of shipments, production deployments, named customers, and follow-on financing indicates meaningful early commercial traction.

Founders & Leadership

Thomas SohmersFounder
Founder and CTO
Edward KmettFounder
Founder and Chief Scientist
Barrett WoodsideFounder
Founder and VP of Product
Mitesh AgrawalChief Executive Officer
Bernie SardiniaHead of Operations & COO
Cameron McCaskillVP, Sales and Business Development

Funding History

2025-02
Seed$23.5M

Not publicly disclosed

2025-06
Series A$51.6M

Valor Equity Partners, Atreides Management, DFJ Growth

2026-01
Series B$234M (company announcement: $230M)

ARENA Private Wealth, Jump Trading, Unless

Recent News

2026-07-10funding
Positron to Raise $750 Million at a $5 Billion Valuation

Positron AI was reported to be in talks to raise $750 million at an approximately $5 billion valuation. The proposed round would significantly exceed the company’s February 2026 Series B valuation.

2026-06-24partnership
i3D.net & Positron AI Partner on EU Sovereign AI Inference

i3D.net and Positron AI announced a partnership to combine Positron’s purpose-built inference hardware with i3D.net’s infrastructure. The initiative is launching first in Europe and focuses on sovereign AI inference.

2026-04-16
Positron shoots for transformative 2026

The article examines Positron’s plans following its $230 million Series B, including its near-$1 billion post-money valuation and ambitions for rapid growth in AI inference hardware.

2026-03-24partnership
Arm expands compute platform to silicon products in historic company first

Arm identified Positron among the companies showing additional commercial momentum around its Arm AGI CPU ecosystem. The announcement places Positron alongside other ecosystem participants including Cerebras, Cloudflare, OpenAI, and SAP.

2026-03-09
Reno’s Positron Joins the AI Chip Unicorn Club

Coverage from Chiplets USA 2026 highlighted Positron’s $230 million financing, which pushed the AI chip company’s valuation beyond $1 billion as demand for inference hardware accelerates.

2026-02-06
Positron jumps up to the big-league investment circle

Jon Peddie Research covered Positron’s $230 million financing and its inference technology, noting that the company’s systems use Intel Agilex 7 M-Series FPGAs.

2026-02-05funding
QIA takes part in Positron $230M Series B funding round

The funding round raised $230 million for Positron AI and valued the company at more than $1 billion. Qatar Investment Authority was among the strategic participants in the oversubscribed Series B.

2026-02-04funding
Positron AI Raises $230 Million Series B at Over $1 Billion Valuation to Scale Energy-Efficient AI Inference

Positron announced a $230 million Series B at a valuation above $1 billion. The funding is intended to support Atlas systems shipping today and the next-generation Asimov chip, targeted for tape-out in late 2026 and production in early 2027.

2026-02-04
Positron AI raises $230M at over $1B valuation to build energy-efficient AI accelerator hardware

SiliconANGLE reported that Positron raised an oversubscribed $230 million round at a $1 billion valuation. The coverage described Atlas as the company’s current inference system and Asimov as its forthcoming custom silicon platform.

2026-02-04product
Asimov: Custom AI Accelerator Silicon

Positron’s Asimov product page describes custom accelerator silicon with up to 2.3TB of memory per chip, claimed fivefold advantages in tokens per dollar and watt versus NVIDIA Rubin, and an approximately 400W air-cooled design. The product is listed as coming in 2027.

Active Roles

17
Austin, TX/Engineering/2d ago
Hybrid (Austin, Texas, US)/Engineering/3d ago
Hybrid (Austin, Texas, US)/Engineering/3d ago
Austin, TX/Engineering/3d ago
Hybrid (Austin, Texas, US)/Engineering/3d ago
Austin, TX/Engineering/10d ago
Remote (United States)/Product/21d ago
Remote (United States)/Product/21d ago
Remote (United States)/Engineering/31d ago
Canada; Remote (United States)/Engineering/31d ago
Canada; Remote (United States)/Engineering/31d ago
Austin, TX/Engineering/31d ago
Remote (United States)/Engineering/31d ago
Canada; Remote (United States)/Engineering/31d ago
Remote (United States); Canada/Engineering/31d ago
Remote (United States)/Engineering/31d ago
Remote (United States); Canada/Engineering/31d ago

Business Model

Positron uses a hardware-led model, selling or deploying Atlas inference appliances and developing Titan systems and Asimov accelerator silicon for enterprise and research workloads. Its hardware is accompanied by model-management and OpenAI-compatible API software; public materials do not disclose a detailed price list or subscription structure.

Products

Atlas — shipping production-ready inference appliance supporting models up to 500B parametersTitan — planned 2027 inference system with 8TB+ memory powered by four Asimov chipsAsimov — planned 2027 purpose-built inference accelerator silicon with 2TB+ memory per chipThe Positronic Brain — Positron's broader AI-infrastructure vision

Customers

CloudflareParasail (via SnapServe)OracleJump TradingCrusoe (proof-of-concept deployment)

Tech Stack

Purpose-built AI inference accelerator siliconTransformer and large-language-model inferenceMemory-centric architecture with terabytes of accelerator memoryArm-based system technologyFPGA prototyping and custom ASIC developmentHugging Face Transformers model support with an OpenAI-compatible endpointAir-cooled server and rack systems

Competitors

NVIDIA
AMD
Cerebras
SambaNova
d-Matrix

Key Investors

ARENA Private Wealth, Jump Trading, Unless, Qatar Investment Authority, Arm, Helena, Valor Equity Partners, Atreides Management, DFJ Growth, Resilience Reserve, Flume Ventures, 1517 Fund