Skip to content

Topic

Limits of model evaluation

Research that probes where machine learning systems fail on tasks they appear to solve, and how unsupervised groupings and benchmark scores can mislead.

Current clusters