A 744-Billion-Parameter Model Now Runs at Home on 300 Watts
GLM-5.2 running across three DGX Sparks pulls roughly the power of three lightbulbs — a data point on how fast frontier-scale inference is escaping the cloud.
A striking benchmark surfaced from the local-inference community: @MiaAI_lab reported running a 744-billion-parameter model — GLM-5.2 — at home on roughly 300 watts total, spread across three DGX Sparks. For context, 300 watts is a gaming PC under load, or three incandescent bulbs. A three-quarters-of-a-trillion-parameter model at that power envelope would have been implausible a year ago.
Unlock the full briefing
Get every story in today's briefing, the full archive, and the daily AI intelligence brief.
All stories today
Full archive
Daily brief
Cancel anytime. Payments powered by Stripe.