Index  ›  tech  ›  TechRadar

Israeli startup takes direct aim at Nvidia with Arm-powered AI server

TechRadar Published Aug 2, 2026 Reviewed Aug 2, 2026 ✓ Reviewed by citations.press editors
Israeli startup takes direct aim at Nvidia with Arm-powered AI server
Majestic Labs' architecture delivers over 50 times more fast memory than the Nvidia DGX B300 configuration.
more than 50 times · fast memory Majestic Labs, company
Majestic Labs' architecture delivers 1.7 times the interconnect bandwidth of the Nvidia DGX B300 configuration.
1.7 times · interconnect bandwidth Majestic Labs, company
A standard 40U rack can hold four Prometheus servers, drawing a total of 120 kW.
120 kW · rack power consumption Majestic Labs, company
An Nvidia DGX B300 system offers 2.3 TB of HBM3e memory and up to 4 TB of DDR5 system memory.
2.3 TB · HBM3e memory4 TB · DDR5 system memory Nvidia, product
Majestic Labs says a Prometheus server could cost between 10 and 50 times less than a GPU system of equivalent performance.
Majestic Labs, company
Majestic Labs raised $100 million in an A-round late in 2025.
100 million USD · funding Majestic Labs, company
Majestic Labs employs around 40 people across Tel Aviv and Los Angeles.
about 40 people · employees Majestic Labs, company
One Majestic rack holds the fast memory capacity of 25 Nvidia NVL72 Vera Rubin racks at a fraction of the power.
25 racks · fast memory capacity Majestic Labs, company
Majestic Labs claims up to 1000 times more memory per processor.
1000 times · memory per processor Majestic Labs, company
Each Prometheus server can house up to 12 AIUs and share between 8 TB and 128 TB of LPDDR6 memory.
12 AIUs · AIUs per server Majestic Labs, company
The memory pool is accessed through custom memory aggregation chiplets linked by copper cables up to one metre long.
1 metre · cable length Majestic Labs, company
A 128 TB configuration relying on 2 GB LPDDR6 dies would need roughly 64,000 dies, implying well over a hundred aggregation chiplets per server.
about 64000 dies · LPDDR6 dies Majestic Labs, company

Majestic Labs, a startup founded in 2023 by former Google and Meta engineers, has unveiled a server built to rival Nvidia's GPU and HBM combination.

The Tel Aviv-based company argues that pairing costly graphics processors with high-bandwidth memory has become a fundamentally memory-bound and dead-end approach for AI inference.

Its answer is the Prometheus server, which swaps GPUs for Ignite AI Processing Units combining Arm cores with RISC-V vector and tensor engines.

Each Prometheus server can house up to 12 AIUs, sharing between 8 TB and 128 TB of LPDDR6 memory across one contiguous, coherent pool.

That memory pool is accessed through custom memory aggregation chiplets linked by copper cables up to one metre long instead of memory attached directly to GPU packages.

A standard 40U rack can hold four such servers, drawing 120 kW total and cooled through cold-plate liquid systems rather than air.

By comparison, an Nvidia DGX B300 system with eight Blackwell GPUs offers 2.3 TB of HBM3e plus up to 4 TB of DDR5 system memory.

Majestic claims its architecture therefore delivers over 50 times more fast memory than that rival configuration, alongside 1.7 times its interconnect bandwidth.

One Majestic rack holds the fast memory capacity of 25 Nvidia NVL72 Vera Rubin racks at a fraction of the power,” Majestic Lab said.

“Organizations that could never justify hyperscaler infrastructure can now run any workload. In fact, there can be up to “1000× more memory per processor.”

Majestic Labs says the Prometheus server could cost between 10 and 50 times less than a GPU system of equivalent performance once it ships next year, while consuming less electricity per rack.

The server is designed to be OCP-compliant and will support PyTorch, vLLM and OpenAI's Triton frameworks, letting existing AI models run without modification.

Founded by CEO Ofer Shacham, President Sha Rabii and COO Masumi Reynders, the company employs around 40 people across Tel Aviv and Los Angeles and raised $100 million in an A-round late in 2025.

It claims to have already received significant orders from large enterprises, neoclouds and hyperscalers.

Yet several details remain unclear, including how many memory aggregation chiplets a single server actually requires.

If a 128 TB configuration relies on widely available 2 GB LPDDR6 dies, it would need roughly 64,000 of them, implying well over a hundred aggregation chiplets per server.

The numbers Majestic Labs presents are striking, but they remain the startup's own projections ahead of any independent benchmarking or shipped hardware.

Enterprise buyers considering a shift away from established GPU vendors may have to wait for independent validation to verify Majestic's claims.

Follow TechRadar on Google News and add us as a preferred source to get our expert news, reviews, and opinion in your feeds.

Efosa has been writing about technology for over 7 years, initially driven by curiosity but now fueled by a strong passion for the field. He holds both a Master's and a PhD in sciences, which provided him with a solid foundation in analytical thinking.

Please logout and then login again, you will then be prompted to enter your display name.

This article was originally published by TechRadar ↗. citations.press indexes the source-backed facts above and links to the original. Something wrong? Corrections policy · Report an error