build1 distinct publisher
Four-bit weights leave 6 GB on a 24 GB card for KV cache and vision tensors
Meta's Muse Glimmer 30B and Alibaba's Qwen3.8-27B both landed in August under pure Apache 2.0 and both fit one 24 GB GPU, so the deployment question moves off licence terms and onto how you spend the memory that is left.
Publishers:dev.to
Reality
- Evidence16
- Adoption12
- Hype gap+52
- Incentives62