Overview
News
Technologies
Salaries
Products
People
Growth
Financials
Overview
A plug-and-play computer for local AI inference. Run Qwen3.6, GLM-4.6 and more on-prem at the fastest tok/s in its price range. Open source.
News
Lucebox Geometric: DeepSeek V4 Flash 0731 reaches 32.7 tok/s on AMD Strix Halo
DeepSeek V4 Flash 0731 on 128 GB AMD Strix Halo: 82/92 quality, up to 32.7 tok/s decode, and 173 tok/s sparse prefill at 60K context.
Read more
Report
Lucebox (AMD Radeon AI PRO R9700 + Strix Halo) Beats NVIDIA DGX Spark by 3.63x on DeepSeek V4 Flash Decode Speed
51.1 tok/s on a Lucebox with AMD Radeon AI PRO R9700 + Strix Halo versus our 14.09 tok/s average on one NVIDIA DGX Spark for the full 284B DeepSeek V4 Flash model.
Read more
Report
DeepSeek V4 Flash: 284B model, up to 32 tok/s on AMD Ryzen AI MAX+ 395
AMD-Powered Lucebox runs the 284B DeepSeek V4 Flash model locally on AMD Ryzen AI MAX+ 395 with 128 GB unified memory, reaching up to 32 tok/s decode and roughly 250 tok/s sparse prefill.
Read more
Report
Report
lucebox/optimizations/kvflash at main Luce-Org/lucebox
LLM speculative inference server for consumer & heterogeneous hardware - lucebox/optimizations/kvflash at main Luce-Org/lucebox
Read more
Report
Pro access
Upgrade to see all 8 mentions
Upgrade to a paid plan to read every media mention of this company - funding news, awards, product launches and press releases from all the outlets writing about it.
Every media mention and press release
Funding news, awards and product launches
Fresh coverage from every outlet writing about the company
Upgrade now
Cancel anytime. Secure checkout. Instant activation.
