Startup swaps costly AI GPUs for Arm cores and up to 128TB of 'cheap' LPDDR6 RAM instead of expensive HBM to smash through the memory wall
Date:
Sun, 02 Aug 2026 00:25:00 +0000
Description:
Majestic Labs unveiled an Arm-based AI server replacing GPUs with unified LPDDR6 memory, claiming lower costs and dramatically larger capacity.
FULL STORY ======================================================================Copy link Facebook X Whatsapp Reddit Pinterest Flipboard Threads Email Share this article 0 Join the conversation Follow us Add us as a preferred source on Google Newsletter Subscribe to our newsletter Startup replaces Nvidia GPUs with custom Arm-powered AI processing hardware Prometheus packs up to 128TB
of unified LPDDR6 memory onboard New server promises 1,000 more memory available per processor Majestic Labs, a startup founded in 2023 by former Google and Meta engineers, has unveiled a server built to rival Nvidia 's GPU and HBM combination.
The Tel Aviv-based company argues that pairing costly graphics processors
with high-bandwidth memory has become a fundamentally memory-bound and dead-end approach for AI inference. Its answer is the Prometheus server,
which swaps GPUs for Ignite AI Processing Units combining Arm cores with RISC-V vector and tensor engines. Latest Videos From TechRadar Watch full video here: A different way to scale memory Each Prometheus server can house up to 12 AIUs, sharing between 8 TB and 128 TB of LPDDR6 memory across one contiguous, coherent pool.
That memory pool is accessed through custom memory aggregation chiplets
linked by copper cables up to one metre long instead of memory attached directly to GPU packages. You may like This AI SSD tech makes 8 RTX 5090s perform like 46 GPUs in inference This tiny AMD PC just ran a massive 397B AI Model that required a server room full of GPUs a year ago Tiny company steals AMD's thunder and challenges Nvidia with old-tech PCIe AI accelerator
A standard 40U rack can hold four such servers, drawing 120 kW total and cooled through cold-plate liquid systems rather than air.
By comparison, an Nvidia DGX B300 system with eight Blackwell GPUs offers 2.3 TB of HBM3e plus up to 4 TB of DDR5 system memory. Are you a pro? Subscribe
to our newsletter Sign up to the TechRadar Pro newsletter to get all the top news, opinion, features and guidance your business needs to succeed! Contact me with news and offers from other Future brands Receive email from us on behalf of our trusted partners or sponsors By submitting your information you agree to the Terms & Conditions and Privacy Policy and are aged 16 or over.
Majestic claims its architecture therefore delivers over 50 times more fast memory than that rival configuration, alongside 1.7 times its interconnect bandwidth.
One Majestic rack holds the fast memory capacity of 25 Nvidia NVL72 Vera
Rubin racks at a fraction of the power, Majestic Lab said.
Organizations that could never justify hyperscaler infrastructure can now run any workload. In fact, there can be up to 1000 more memory per processor.
What to read next Key partner to Nvidia, ASML and TSMC brings next-gen RAM
and NAND replacements even closer AMD abandons HBM for inferior LPDDR5x as AI monster devours precious high-bandwidth memory Qualcomm targets Nvidia, AMD, Huawei with Dragonfly AI accelerator rack loaded with 43TB of LPDDR5x, future generations set to smash 7PB/s bandwidth Performance claims await broader testing Majestic Labs says the Prometheus server could cost between 10 and 50 times less than a GPU system of equivalent performance once it ships next year, while consuming less electricity per rack.
The server is designed to be OCP-compliant and will support PyTorch, vLLM and OpenAI's Triton frameworks, letting existing AI models run without modification.
Founded by CEO Ofer Shacham, President Sha Rabii and COO Masumi Reynders, the company employs around 40 people across Tel Aviv and Los Angeles and raised $100 million in an A-round late in 2025.
It claims to have already received significant orders from large enterprises, neoclouds and hyperscalers.
Yet several details remain unclear, including how many memory aggregation chiplets a single server actually requires.
If a 128 TB configuration relies on widely available 2 GB LPDDR6 dies, it would need roughly 64,000 of them, implying well over a hundred aggregation chiplets per server.
The numbers Majestic Labs presents are striking, but they remain the
startup's own projections ahead of any independent benchmarking or shipped hardware.
Enterprise buyers considering a shift away from established GPU vendors may have to wait for independent validation to verify Majestic's claims.
Via Blocks and Files Follow TechRadar on Google News and add us as a
preferred source to get our expert news, reviews, and opinion in your feeds.
======================================================================
Link to news story:
https://www.techradar.com/pro/startup-swaps-costly-ai-gpus-for-arm-cores-and-u p-to-128tb-of-cheap-lpddr6-ram-instead-of-expensive-hbm-to-smash-through-the-m emory-wall
--- Mystic BBS v1.12 A49 (Linux/64)
* Origin: tqwNet Technology News (1337:1/100)