SWE-fficiency
Progress Over Time
Interactive timeline showing model performance evolution on SWE-fficiency
SWE-fficiency Leaderboard
| Context | Cost | License | ||||
|---|---|---|---|---|---|---|
| 1 | MiniMax | — | 1.0M | $0.30 / $1.20 |
What is SWE-fficiency?
SWE-fficiency is an open-source benchmark and workflow that evaluates language models on optimizing the runtime efficiency of real-world software engineering tasks, measuring how well agents can improve code performance autonomously.
SWE-fficiency is a text benchmark evaluating models on agents and code tasks. LLM Stats tracks 1 models on this benchmark, scored on a 0–1 scale. The current average is 0.3, with the leader at 0.3.
Compare leaders on the best AI for agents and best AI for code leaderboards.
Current leaders
MiniMax M3 from MiniMax currently leads the SWE-fficiency leaderboard with a score of 0.348 across 1 evaluated AI models.
FAQ
Common questions about the SWE-fficiency benchmark and leaderboard.