A 744-Billion-Parameter Model Now Runs at Home on 300 Watts

GLM-5.2 running across three DGX Sparks pulls roughly the power of three lightbulbs — a data point on how fast frontier-scale inference is escaping the cloud.

A striking benchmark surfaced from the local-inference community: @MiaAI_lab reported running a 744-billion-parameter model — GLM-5.2 — at home on roughly 300 watts total, spread across three DGX Sparks. For context, 300 watts is a gaming PC under load, or three incandescent bulbs. A three-quarters-of-a-trillion-parameter model at that power envelope would have been implausible a year ago.

Unlock the full briefing

Get every story in today's briefing, the full archive, and the daily AI intelligence brief.

All stories today

Full archive

Daily brief

Cancel anytime. Payments powered by Stripe.