Skip to content

model

Gemma 4 26B

Open-weight Gemma-family model; a 26B-parameter mixture-of-experts variant with ~4B active parameters, often run locally in 4-bit form.

Known aliases

  • Gemma
  • Gemma 4 26B A4B
  • gemma-4-26B-A4B-it-AWQ-4bit

Relationships

No evidence-backed relationships are recorded.

Current stories

build1 publisherOne report

Speculative decoding in llama-server swaps real logprobs for 0.0 placeholders

llama-server b11430 reports logprob 0.0 for every speculatively decoded token, dragging one test's mean logprob from -0.48 to -0.0011. Nothing in the response or the server log flags the fill-ins, so evals and calibration built on those numbers go wrong quietly.

Publishers:dev.to

Reality

Evidence64
Adoption
Insufficient
Hype gap0
Incentives
Insufficient
Confidence58
build1 publisherOne report

A 4-bit Gemma 4 26B on one L4 trails TypeSafe's Jev by 2.1 points overall

A pre-registered run reads Gemma's label logits on one 24 GB GPU and scores it against the hosted API on the same 3,880 records, where it is level on yes/no questions and 4.5 points behind on multiple choice. One temperature fitted on 50 labels closes the calibration gap.

Publishers:dev.to

Reality

Evidence70
Adoption20
Hype gap−12
Incentives35
Confidence55