From the HackerLinks archive

CursorBench

Cursor's coding-model evaluation suite for comparing practical agent performance.

At a glance:
First seen:2026-09-06
Last seen:2026-09-06
Times seen:1
Website:cursor.com

The short version

Cursor's coding-model evaluation suite for comparing practical agent performance.

Why it caught our attention

A commenter said its rankings consistently matched their subjective model evaluations.

Where it surfaced on Hacker News

2026-09-06

Their bench has always been one that most-fit my mental model of how good each of these models are.

jjcm · recommendation · evaluated

Direct HN comment

Artificial Analysis Intelligence Index v4.2

Also surfaced in this discussion

Artificial Analysis Omniscience Index

Artificial Analysis Intelligence Index v4.2

2026-09-06