Skip to content

Topic

Local and on-device inference

Running capable models on consumer GPUs and workstations, including quantization and memory budgeting.

Current stories

build6 publishers

Meta's real announcement is the split: 30B on your GPU, everything else behind the API

Muse Glimmer ships as Apache 2.0 weights sized for a 24GB card. Muse Spark 1.2 stays on Muse Code and the Meta Model API. Plan capacity for two tiers, not one.

Perspective Coverage

6 publishers
Builder
Builder 38%
Operator
Operator 31%
Investor
Investor 31%

Reality

Evidence64
Adoption48
Hype gap+18
Incentives78
Confidence70