Skip to content

model

GPT-5.6 Luna

Model configuration used as bounded Sol-Luna workers, invoked with a reasoning effort of Medium, High, XHigh or Max when the supervisor delegates.

Known aliases

  • gpt-5.6-luna
  • GPT-5.6-luna (medium)
  • in.openai.gpt-5.6-luna
  • Luna
  • openai.gpt-5.6-luna

Relationships

No evidence-backed relationships are recorded.

Current stories

build1 publisher

Epoch AI finds GPQA-grade answers getting 13 times cheaper every year

Epoch AI puts the price of answering a GPQA Diamond question at a fixed accuracy bar falling about 13 times a year, faster than compute under Moore's Law. Buyers get that discount only by moving to newer models, so the part of a stack that has to switch cheaply is the evaluation that qualifies each one.

Publishers:tomshardware.com

Reality

Evidence55
Adoption
Insufficient
Hype gap+30
Incentives
Insufficient
Confidence45
build1 publisher

Asking GPT-5.6 Luna to name an amphibian flags benchmark transcripts with black-box access

GPT-5.6 Luna says "frog" 70-95% of the time when asked for an amphibian after capability benchmarks, against 12-38% after real use, a LessWrong post reports. Anyone with black-box access can run the check, though its authors cannot yet say whether it detects evaluation awareness or lexical cues.

Publishers:lesswrong.com

Reality

Evidence45
Adoption
Insufficient
Hype gap+10
Incentives
Insufficient
Confidence40
build1 publisher

LangChain drops about 4,000 base input tokens from every default Deep Agents turn

The harness lost its hidden system prompt, 43% of its builtin tool descriptions and its todo list middleware. LangChain's own footnote says reward confidence intervals span zero for every model tested, so the evals settle the token saving more firmly than the quality.

Publishers:langchain.com

Reality

Evidence58
Adoption30
Hype gap+18
Incentives82
Confidence46

Earlier coverage

  1. Holding AI revenue flat now takes 69% more tokens than it did in March

    Invest · September 12, 2026 · 1 publisher

  2. 200 parallel sandboxes researched the Next.js backlog before maintainers closed 1,462 issues

    Build · September 11, 2026 · 1 publisher

  3. A 1.5x per-token price still bought a 25 percent cheaper correct answer in AWS's benchmark

    Build · September 11, 2026 · 1 publisher

  4. A 41% fall in token prices leaves labs needing 69% more volume to stand still

    Invest · September 9, 2026 · 1 publisher

  5. OpenAI's cost-per-task argument buys Luna room for ten failed tries before it loses on price

    Invest · September 8, 2026 · 1 publisher

  6. GLM-5.3-Flash benchmarks its tenth-of-the-price claim against its own predecessor

    Leadership · September 5, 2026 · 1 publisher

  7. Google releases third Gemini Flash model in six weeks

    Product · September 3, 2026 · 1 publisher

  8. Per-PTU throughput spans 25x across three models in the same GPT-5.6 family

    Build · August 30, 2026 · 1 publisher

  9. Judging the cheap model's output beats guessing which prompt is hard

    Build · August 30, 2026 · 1 publisher

  10. Changing one model-ID prefix pins GPT-5.6 inference to Mumbai and Hyderabad

    Build · August 27, 2026 · 1 publisher

  11. OpenAI's top model at $4/$20 is a three-month answer to a permanent build decision

    Build · August 25, 2026 · 1 publisher

  12. App factory or agent fleet manager: the fork is whose rate limit stops the work

    Build · August 24, 2026 · 1 publisher

  13. Codex's usage wall is being rebuilt as a cheaper tier, shipped binaries suggest

    Build · August 23, 2026 · 1 publisher

  14. Meta's coding agent has two prices: pay 18x more, or let it train on your repository

    Invest · August 22, 2026 · 1 publisher

  15. Callosum raises $100m for mixed-silicon scheduling, and the 2x accuracy claim is still its own

    Product · August 20, 2026 · 2 publishers

  16. Bedrock turns GPT-5.6 throughput into a routing choice, with residency as the price

    Build · August 20, 2026 · 1 publisher

  17. The 21-cent model bake-off that inverted when the judge got audited

    Build · August 20, 2026 · 1 publisher

  18. A 27B laptop model scores like a rented one, and thinks three times as hard to do it

    Product · August 19, 2026 · 1 publisher

  19. Four frontier models in four days, and the cheapest number in your agent plan has an expiry date

    Build · August 18, 2026 · 1 publisher

  20. OpenAI's Multi-Agent v2 turns tiered-model cost arbitrage into a supported architecture

    Invest · August 16, 2026 · 1 publisher

  21. US inference prices fell nearly a quarter in a month. Your unit economics are stale.

    Invest · August 16, 2026 · 1 publisher

  22. A coding orchestrator allowed to delegate chose zero workers, six times out of six

    Build · August 15, 2026 · 1 publisher