Cursor said Claude Fable 5.1 is live and is the strongest model it has run on CursorBench 3.2, scoring 73.4% at max effort. The team highlighted its skill at verifying its own work on hard, end-to-end coding tasks. GitHub Copilot also made the model generally available the same day.

Key Takeaways

  • βœ“73.4% at max effort on CursorBench 3.2, the best result Cursor has recorded
  • βœ“Strong at self-verification, enabling difficult coding jobs to be taken from start to finish
  • βœ“GitHub Copilot GA in the app, CLI, and VS Code for long-running coding, codebase research, and agentic workflows
ADSponsored