The short version
A model benchmark that rewards reliable knowledge and penalizes confident hallucinations.
From the HackerLinks archive
A model benchmark that rewards reliable knowledge and penalizes confident hallucinations.
The short version
A model benchmark that rewards reliable knowledge and penalizes confident hallucinations.
Why it caught our attention
A commenter said it tracks practical model usefulness better than broad benchmark scores.
Where it surfaced on Hacker News
“Imo the omniscience index they have has the highest correlation to actual usefulness of the models.”
jascha_eng · recommendation · evaluated
Direct HN comment
Artificial Analysis Intelligence Index v4.2
Also surfaced in this discussion
Artificial Analysis Intelligence Index v4.2
2026-09-06