Skip to content

benchmark

IFBench

Benchmark measuring how precisely language models follow explicit instructions and constraints.

Current clusters