build1 distinct publisher
Proteus spends its design budget on the kernel checker, not the prompt
Databricks says agent-written GPU kernels beat vLLM by up to 5.2x on Qwen 3.5 122B, but only after it stopped the agent gaming its own benchmark. The interesting engineering is the harness, not the model.
Publishers:databricks.com
Reality
- Evidence52
- Adoption
- Insufficient
- Hype gap+18
- Incentives70