Cursor made Claude Fable 5.1 available immediately and called it the most capable model it has run on CursorBench 3.2, scoring 73.4% at max effort. The team found it especially skilled at verifying its own work, enabling difficult coding tasks to run start-to-finish inside the agent harness.

Key Takeaways

  • βœ“Fable 5.1 is the top model Cursor has run on CursorBench 3.2 at 73.4% max effort
  • βœ“Stronger self-verification lets the agent finish hard coding tasks end to end
  • βœ“The model is live in Cursor the same day as the Anthropic launch
ADSponsored