Grok 4.7 arrived on Monday after five walked-back timelines, with 40% more parameters and Grok 4.6's list price intact. Cursor's own cost chart still puts its price per task above GPT-6 Astra and Claude Sonnet 5.
Perspective Coverage
3 publishers
- Builder
- Builder 42%
- Operator
- Operator 27%
- Investor
- Investor 31%
Reality
- Evidence45
- Adoption35
- Hype gap+15
- Incentives70
- Confidence55
xAI built Grok 4.7 on a larger base model with a longer reinforcement-learning run and kept the API at Grok 4.6's rates. The open question for buyers is how many tokens the longer runs burn.
Reality
- Evidence34
- Adoption22
- Hype gap+28
- Incentives82
- Confidence46
Artificial Analysis priced the new model at $10 and $50 per million tokens, and the threefold token cut it measured in the Codex agent harness covers that increase where the roughly 10% cut on its intelligence suite does not.
Reality
- Evidence63
- Adoption
- Insufficient
- Hype gap+16
- Incentives62
- Confidence57
Artificial Analysis' AA-Briefcase grades deliverables built from four multi-week projects, but each task starts with no memory of the model's own earlier submissions. Continuity stays unmeasured.
Reality
- Evidence60
- Adoption
- Insufficient
- Hype gap+12
- Incentives62
- Confidence55
Artificial Analysis puts Grok 4.6 at 61 on its Intelligence Index, level with GPT-5.6 Sol, at $2/$6 per million tokens. The same pages record 48 seconds to first token.
Reality
- Evidence56
- Adoption24
- Hype gap+24
- Incentives63
- Confidence50
Grok 4.6, Gemini 3.7 Flash, DeepSeek V4 Pro and GLM-5.3 all chase agents that stay on task. The pricing underneath them is moving faster than the benchmarks.
Reality
- Evidence58
- Adoption55
- Hype gap+12
- Incentives68
- Confidence48
Optima lets buyers build benchmarks from their own datasets and agent traces, then scores candidate models on quality, cost per task and time per task.
Reality
- Evidence34
- Adoption16
- Hype gap+22
- Incentives71
- Confidence33
Grok 4.6, Qwen3.8-Max and DeepSeek V4-Pro shipped inside about 24 hours, and two of the three came with downloadable weights. The benchmarks existed to justify a cheaper invoice.
Perspective Coverage
4 publishers
- Builder
- Builder 41%
- Operator
- Operator 31%
- Investor
- Investor 28%
Reality
- Evidence68
- Adoption52
- Hype gap+22
- Incentives74
- Confidence63