Skip to content

model

Gemini 3.1 Pro

Frontier model scoring 56.2 percent on PerceptionBench.

Known aliases

  • 3.1 Pro
  • Gemini 3.1
  • gemini-3.1-pro-preview
  • Gemini 3.1 Pro + Search

Relationships

No evidence-backed relationships are recorded.

Current stories

build1 publisher

Claude's new tokenizer and GPT-6's 272K price cliff break old LLM cost models

Claude 4.7 emits about 30 percent more tokens for the same text and GPT-6 bills roughly double above 272K input tokens, a dev.to digest reports. Budget checks built on old token counts now undercount, so prompt size needs a hard cap enforced in code.

Publishers:dev.to

Reality

Evidence35
Adoption
Insufficient
Hype gap+10
Incentives
Insufficient
Confidence35
build1 publisher

Deep Agents now swaps in apply_patch the moment you name a Codex model

LangChain's new harness profiles set prompts, tool implementations and tool names per model family. The company measures a 10 to 20 point gain on a tau2-bench subset it curated from tasks frontier models have not saturated.

Publishers:langchain.com

Reality

Evidence45
Adoption15
Hype gap+18
Incentives80
Confidence55