Robust IF
Progress Over Time
Interactive timeline showing model performance evolution on Robust IF
Robust IF Leaderboard
| Context | Cost | License | ||||
|---|---|---|---|---|---|---|
| 1 | Microsoft | — | — | — |
What is Robust IF?
Robust IF evaluates instruction-following robustness on diverse, hard prompts, measuring whether a model reliably adheres to constraints across challenging single-turn and multi-turn scenarios.
Robust IF is a text benchmark evaluating models on instruction following and reasoning tasks. LLM Stats tracks 1 models on this benchmark, scored on a 0–1 scale. The current average is 0.6, with the leader at 0.6.
Compare leaders on the best AI for instruction following and best AI for reasoning leaderboards.
Current leaders
MAI-Code-1-Flash from Microsoft currently leads the Robust IF leaderboard with a score of 0.612 across 1 evaluated AI models.
FAQ
Common questions about the Robust IF benchmark and leaderboard.