4 Sources
[1]
Positron AI says its Atlas accelerator beats Nvidia H200 on inference in just 33% of the power -- delivers 280 tokens per second per user with Llama 3.1 8B in 2000W envelope
To address concerns about power consumption of systems used for AI inference, hyperscale cloud service provider (CSP) Cloudflare is testing various AI accelerators that are not AI GPUs from AMD or Nvidia, reports the Wall Street Journal. Recently, the company began to test drive Positron AI's Atlas
[2]
Positron bets on energy-efficient AI chips to challenge Nvidia's dominance
Highly anticipated: A new front is emerging in the race to power the next generation of artificial intelligence, and at the center of it is a startup called Positron whose bold ambitions are gaining traction in the semiconductor industry. As companies scramble to rein in the soaring energy demands
[3]
Positron believes it has found the secret to take on Nvidia in AI inference chips -- here's how it could benefit enterprises
Want smarter insights in your inbox? Sign up for our weekly newsletters to get only what matters to enterprise AI, data, and security leaders. Subscribe Now As demand for large-scale AI deployment skyrockets, the lesser-known, private chip startup Positron is positioning itself as a direct
[4]
Positron AI Secures $51.6 Mn in Series A to Build its AI Inference Engine | AIM
Positron claims that Atlas offers 3.5 times better performance per dollar compared to NVIDIA's H100. Positron AI, known for developing American-made hardware and software for AI inference, has bagged a $51.6 million oversubscribed Series A funding round, raising its total capital for the year to
Share
Copy Link
Positron AI, a startup founded in 2023, is making waves in the AI hardware industry with its Atlas accelerator, claiming superior performance and energy efficiency compared to Nvidia's offerings for AI inference tasks.
Positron AI, a startup founded in 2023, is making waves in the AI hardware industry with its Atlas accelerator, claiming superior performance and energy efficiency compared to Nvidia's offerings for AI inference tasks
1
. The company has recently secured $51.6 million in Series A funding, bringing its total capital raised to over $75 million4
.
Source: Tom's Hardware
According to Positron AI, the Atlas accelerator can deliver around 280 tokens per second per user in Llama 3.1 8B with BF16 compute at 2000W, compared to approximately 180 tokens per second per user for an 8-way Nvidia DGX H200 server consuming 5900W
1
. The company claims that Atlas offers:3
Atlas is designed specifically for large-scale transformer models and packs eight Archer accelerators
1
. Key features include:4

Source: AIM
Positron AI is positioning itself as a direct challenger to Nvidia in the AI inference chip market. The company's focus on energy efficiency and cost-effectiveness has attracted attention from major cloud providers and enterprises
2
. Early adopters include:3
Related Stories
Positron AI is already working on its next-generation system, Titan, powered by the Asimov AI accelerator. Expected to launch in 2026, Titan aims to compete against inference systems based on Nvidia's Vera Rubin platforms
1
. Key features of the upcoming system include:1

Source: VentureBeat
The emergence of Positron AI and other startups in the AI chip space highlights the growing concern over the power consumption and cost of AI infrastructure. As AI models continue to grow in size and complexity, efficient inference solutions become increasingly critical
2
.However, Positron AI faces significant challenges in a market dominated by established players like Nvidia. The company will need to deliver on its performance and efficiency claims to gain widespread adoption and compete effectively in this rapidly evolving industry
3
.Summarized by
Navi
[3]
13 Feb 2025•Technology

04 Feb 2026•Startups

Yesterday•Startups

1
Science and Research

2
Technology
3
Technology