Unsloth fixed Studio in version 2026.6.9 after Pillar Security showed that reading a malicious model's config.json could run an attacker's Python code. Unsloth disputed parts of the finding and declined to publish an advisory, so no CVE was assigned.
Reality
- Evidence48
- Adoption
- Insufficient
- Hype gap+10
- Incentives55
- Confidence52
A dev.to guide to running local models on 8GB prices the KV cache between 15KB and 160KB per token depending on architecture. At 32K tokens held, that spread is the difference between 0.5GB and 5GB of a fixed budget.
Reality
- Evidence22
- Adoption
- Insufficient
- Hype gap+45
- Incentives62
- Confidence58
Alibaba's 27B model fits a 32GB card with 15GB to spare. Tom's Hardware still had to pick between llama.cpp's full 262K window at half-hour prefill and a supported vLLM deployment capped at 32K.
Reality
- Evidence60
- Adoption38
- Hype gap+15
- Incentives55
- Confidence58
Qwen3.8-Flash-Next puts 36 Gated DeltaNet layers and 12 sparse-attention layers on Hugging Face, which means the retrieval budget Qwen4 will inherit is something you can measure against your own traces now.
Perspective Coverage
5 publishers
- Builder
- Builder 52%
- Operator
- Operator 28%
- Investor
- Investor 20%
Reality
- Evidence58
- Adoption52
- Hype gap+32
- Incentives76
- Confidence71
A 15M-parameter model streams English text on a 2007 PSP at about one token per second. That is the extreme end of a sizing rule. The harder half of that rule is checking whether the file that fits is a format its own maintainer recommends.
Reality
- Evidence38
- Adoption31
- Hype gap+12
- Incentives58
- Confidence46
Z.ai says the base model did not change between GLM-5.2 and GLM-5.3, so the coding jump and the doubled exploitation score come out of the same post-training run. Security teams inherit the second half.
Perspective Coverage
8 publishers
- Builder
- Builder 45%
- Operator
- Operator 31%
- Investor
- Investor 24%
Reality
- Evidence55
- Adoption40
- Hype gap+18
- Incentives75
- Confidence70
The MIT-licensed 320B model card claims it beats GLM-5.2 at a tenth of the price and approaches Claude Opus 4.8 on coding, but it names no dollar rate, and the comparisons are largely the vendor's own.
Reality
- Evidence34
- Adoption18
- Hype gap+46
- Incentives82
- Confidence61
Google's Gemma milestone arrives with an engineer's caveat and no breakdown. Alibaba's rival claim of 3bn Qwen downloads is about 1.5 times what Hugging Face independently counted.
Reality
- Evidence34
- Adoption61
- Hype gap+38
- Incentives79
- Confidence52
Dynamic 3.0 ships Qwen3.8-27B GGUFs from 6.2GB up, with an unreproduced accuracy claim attached. The number that matters is the one that decides where the file fits.
Reality
- Evidence34
- Adoption45
- Hype gap+28
- Incentives74
- Confidence41
Hugging Face counts 28,531 community GGUF conversions of Alibaba's Qwen models against 54 from Alibaba itself. Procurement signs for the model; production loads the artifact.
Reality
- Evidence60
- Adoption71
- Hype gap+14
- Incentives55
- Confidence58
Muse Glimmer ships as Apache 2.0 weights sized for a 24GB card. Muse Spark 1.2 stays on Muse Code and the Meta Model API. Plan capacity for two tiers, not one.
Perspective Coverage
6 publishers
- Builder
- Builder 38%
- Operator
- Operator 31%
- Investor
- Investor 31%
Reality
- Evidence64
- Adoption48
- Hype gap+18
- Incentives78
- Confidence70
Two releases, two licences. Only the 27B is Apache 2.0, and because just 16 of its 64 layers keep a KV cache, long context costs a quarter of the usual memory.
Reality
- Evidence48
- Adoption34
- Hype gap+14
- Incentives70
- Confidence42