CursorBench 4.0: picking a Claude model by cost and tokens
CursorBench has been updated to version 4.0, and Sonnet 5.5, Opus 5.5 and Fable 5.1 are now on the leaderboard.
I compared them with the previous generation (Sonnet 5 and Opus 5), looking at cost and tokens per task.
[Read More]