Skip to content

benchmark

CursorBench

Coding benchmark where Grok 4.6 scored 69.9% on v3.2 against Fable 5 Max at 70.5%.

Known aliases

  • CursorBench
  • CursorBench 3.2
  • CursorBench 4.0
  • CursorBench v3.2

Relationships

No evidence-backed relationships are recorded.

Current stories

invest3 publishers

xAI holds Grok's $2 token price for a model 40 Elo points behind Fable 5.1

Grok 4.7 arrived on Monday after five walked-back timelines, with 40% more parameters and Grok 4.6's list price intact. Cursor's own cost chart still puts its price per task above GPT-6 Astra and Claude Sonnet 5.

Perspective Coverage

3 publishers
Builder
Builder 42%
Operator
Operator 27%
Investor
Investor 31%

Reality

Evidence45
Adoption35
Hype gap+15
Incentives70
Confidence55
build4 publishers

Three frontier launches in a day, all pitched on price. Open weights set the ceiling.

Grok 4.6, Qwen3.8-Max and DeepSeek V4-Pro shipped inside about 24 hours, and two of the three came with downloadable weights. The benchmarks existed to justify a cheaper invoice.

Perspective Coverage

4 publishers
Builder
Builder 41%
Operator
Operator 31%
Investor
Investor 28%

Reality

Evidence68
Adoption52
Hype gap+22
Incentives74
Confidence63