Skip to content

model

Qwen3.8-27B

Dense 27-billion-parameter vision-language checkpoint scheduled for release on August 14, 2026, with a 262,144-token native context window.

Known aliases

  • Qwen 27B
  • Qwen3.8
  • Qwen 3.8 27B
  • Qwen3.8 27B
  • qwen3.8:27b
  • Qwen3.8-27B-AWQ-INT4

Relationships

No evidence-backed relationships are recorded.

Current stories

build1 publisher

Size the model to the RAM you own before the 45-minute download

A 15M-parameter model streams English text on a 2007 PSP at about one token per second. That is the extreme end of a sizing rule. The harder half of that rule is checking whether the file that fits is a format its own maintainer recommends.

Publishers:dev.to

Reality

Evidence38
Adoption31
Hype gap+12
Incentives58
Confidence46
build1 publisher

Decode drags the entire model out of VRAM once per word

Prefill and decode sit on the same card and answer to different limits, which is why a GPU with more arithmetic and the same memory read rate leaves your time per output token exactly where it was.

Publishers:dev.to

Reality

Evidence44
Adoption
Insufficient
Hype gap+16
Incentives28
Confidence56

Earlier coverage

  1. Inco AI's DFlash 2: 21% longer accepted drafts for 1.3% latency and 18.5M parameters

    Build · August 19, 2026 · 1 publisher

  2. A 27B laptop model scores like a rented one, and thinks three times as hard to do it

    Product · August 19, 2026 · 1 publisher

  3. 128GB of DDR5 is $3,399: your 2026 memory budget is void

    Build · August 19, 2026 · 2 publishers

  4. Dual 3090s, no NVLink: the serving stack broke long before the model did

    Build · August 18, 2026 · 1 publisher

  5. A refusal-stripped 27B model now ships as a 17.9 GB llama.cpp pull

    Build · August 16, 2026 · 1 publisher

  6. Qwen3.8's 27B dense checkpoint is the one operators can actually host

    Build · August 14, 2026 · 1 publisher