A Rival to Nvidia? Positron Atlas, the GPU That Saves Energy and Delivers Speed

A Rival to Nvidia? Positron Atlas, the GPU That Saves Energy and Delivers Speed

Ondřej Barták
Ondřej Barták
Entrepreneur and Programmer
4. 8. 2025
2 minutes reading · 19 views
A Rival to Nvidia? Positron Atlas, the GPU That Saves Energy and Delivers Speed

Competition for Nvidia? Positron Atlas, a GPU That Saves Energy and Delivers Speed

Imagine a world of artificial intelligence where performance does not come at the cost of enormous energy consumption. Positron AI has introduced the Atlas accelerator (GPU), which it claims outperforms the Nvidia H200 in inference while using only 33% of the energy. This article explores details from available sources, including performance comparisons and technical specifications, to provide you with a clear and engaging overview of this innovation.

Atlas vs. Nvidia H200

Positron AI's Atlas accelerator achieves approximately 280 tokens per second per user when running the Llama 3.1 8B model, all while consuming 2,000 W. In comparison, an 8x Nvidia DGX H200 system achieves around 180 to 182 tokens per second per user but requires up to 5,900 W. This means that Atlas consumes roughly one-third of the energy used by the Nvidia H200 while delivering similar or better results in transformer inference tasks. According to Positron AI's internal benchmarks and initial third-party tests, Atlas offers 3 to 4.5 times better performance per watt and 3 to 3.1 times better value per dollar. These figures indicate the potential to cut data center costs by up to half for comparable AI workloads.

Power consumption vs. tokens

Technical Details and Architecture

Atlas is designed specifically for inference, unlike the Nvidia H200, which is a more general-purpose GPU for AI. Its custom FPGA-based architecture achieves more than 93% memory bandwidth utilization, significantly higher than the 10 to 30% typical of GPUs. This approach enables higher throughput and lower latency for large language models. Atlas supports all transformer models from Hugging Face and offers an OpenAI-compatible API, making it easier to integrate into existing systems. However, it is not intended for AI training or other general-purpose computing tasks, where Nvidia still dominates.

Specifications

Availability and Practical Deployment

Atlas is already being shipped to enterprise and cloud customers, including Cloudflare, which deployed it at an early stage. This focus on inference delivers energy-saving benefits—up to 66 to 70% lower consumption for similar performance. It is important to note that these figures are largely based on Positron AI's internal testing and have not yet been widely verified by independent reviewers. Nevertheless, they suggest a promising shift toward more efficient AI computing.

This development from Positron AI could change how companies approach large AI models by combining speed, savings, and easy integration. If you are looking for ways to optimize your AI operations, Atlas is worth considering.

Advertisement

Content created with help from UpTier.

SEO and GEO on autopilot. UpTier’s multi-agent systems write and optimize content for search engines and AI answers.

Discover UpTier ↗

Category:AI
Did you enjoy this article?
Discover more interesting posts on our blog
Back to blog

Related posts

OpenAI gives Codex reusable cloud workspaces accessible from any deviceOpenAI gives Codex reusable cloud workspaces accessible from any device
Codex gains reusable cloud development environments, alongside voice controls in its CLI, code reviews in the ChatGPT desktop app and cloud-based security tools.
2 min read
2. 10. 2026
Amazon releases Strands Decider 2B for AI workflow decisionsAmazon releases Strands Decider 2B for AI workflow decisions
Strands Decider 2B selects from predefined options and returns a confidence score. The fully open-source model is available now and small enough to run locally.
2 min read
1. 10. 2026
OpenAI says it disrupted a campaign to extract hidden model reasoningOpenAI says it disrupted a campaign to extract hidden model reasoning
OpenAI reported a coordinated effort to extract protected model reasoning and said it closed an extraction pathway. It attributed the main cluster of activity to individuals associated with Moonshot AI, the developer of Kimi.
3 min read
1. 10. 2026
Přihlaste se k odběru našeho newsletteru
Zůstaňte informováni o nejnovějších příspěvcích, exkluzivních nabídkách, a aktualizacích.
CodedTrip

Operated by CodedTrip LLC, USA.

YouTube
TikTok