build1 distinct publisher
Half the BIRD training examples in a 2,500-row audit carried the wrong reference query
Thinking Machines and four academics trained one model past the usual agent pipelines on text-to-SQL, but the part worth copying is the audit they ran first, which put annotation errors in 61.1% of the benchmark examples they checked.
Publishers:runtimewire.com
Reality
- Evidence57
- Adoption20
- Hype gap+14
- Incentives70