Cursor made Claude Fable 5.1 available immediately and called it the most capable model it has run on CursorBench 3.2, scoring 73.4% at max effort. The team found it especially skilled at verifying its own work, enabling difficult coding tasks to run start-to-finish inside the agent harness.
Key Takeaways
- βFable 5.1 is the top model Cursor has run on CursorBench 3.2 at 73.4% max effort
- βStronger self-verification lets the agent finish hard coding tasks end to end
- βThe model is live in Cursor the same day as the Anthropic launch