Goalposts

Hacker News has set AI a lot of challenges over the years. Which ones has it met?

2025 Januaryhatthew

About whether the Humanity's Last Exam benchmark is a meaningful final test for AI.

Given the name, I expected it to be more like "write a 500 page novel that a publisher accepts", "solve an open math problem", "improve united airlines' flight schedule", "develop a novel commercially-viable pharmaceutical", "control this humanoid robot to cook a fried egg in this random person's kitchen", "decisively pass the turing test where the judge is an expert in AI". Academic trivia is cool but is nowhere near the "last exam" necessary for AI.

An AI writes a publisher-accepted novel, solves an open math problem, or decisively passes a Turing test judged by an AI expert.

Has this happened?

Yes 76 (55%)Not sure 27 (19%)No 36 (26%)

Votes cast 1–2 October 2026: 100,590 votes from 9,694 people.