Skip to content

model

Gemma 4

Local model a commenter cites as the only other one to pass their private reasoning benchmark.

Known aliases

  • Gemma 4 26B-A4B
  • Gemma 4-31B
  • gemma4:e2b
  • gemma4:e4b
  • Gemma 4 family
  • Gemma 4 MoE

Relationships

No evidence-backed relationships are recorded.

Current stories

build5 publishers

Aleph Alpha's open Kolibri model routes each token through 3.46B of its 78.1B parameters

Aleph Alpha released Kolibri, an Apache 2.0 German-English model that activates 3.46B of its 78.1B parameters per token. Each token costs about as much compute as a small model, yet a team hosting it in Europe still has to fit every expert in memory.

Perspective Coverage

5 publishers
Builder
Builder 49%
Operator
Operator 36%
Investor
Investor 15%

Reality

Evidence70
Adoption
Insufficient
Hype gap+15
Incentives65
Confidence68
build1 publisher

Repacked 4-bit embeddings lift Gemma 4 decode up to 1.39x on a single L4

Repacking Gemma 4's QAT weights with 4-bit embedding tables made decode up to 1.39x faster on one SageMaker L4, according to a dev.to benchmark series. For teams serving Gemma 4 on vLLM, how the weights are stored becomes a setting to measure alongside model size.

Publishers:dev.to

Reality

Evidence55
Adoption
Insufficient
Hype gap+10
Incentives
Insufficient
Confidence50

Earlier coverage

  1. A self-healing scraper that must prove its repair against twelve records that cannot move

    Build · August 22, 2026 · 1 publisher

  2. A billion downloads, and nobody will say what a download is

    Product · August 21, 2026 · 1 publisher

  3. Ornith-1.5 moves the RL loop upstream, and the hard job becomes reward design

    Build · August 19, 2026 · 2 publishers

  4. He scored "tier one" for AI use. His actual pipeline has at least seven jobs in it

    Product · August 18, 2026 · 1 publisher

  5. A 27B Apache-2.0 model in 17GB makes local inference a wiring decision, not a demo

    Build · August 15, 2026 · 1 publisher