Goalposts

Hacker News has set AI a lot of challenges over the years. Which ones has it met?

2025 Junejanalsncm

About an o3 Pro review claiming models are now too good for simple tests.

Trying out o3 Pro made me realize that models today are so good in isolation, we’re running out of simple tests.

Are Towers of Hanoi not a simple test? Or chess? A recursive algorithm that runs on my phone can outclass enormous models that cost billions to train.

A reasoning model should be able to reason about things. I am glad models are better and more useful than before but for an author to say they can’t even evaluate o3 makes me question their credibility.

https://machinelearning.apple.com/research/illusion-of-think...

AGI means the system can reason through any problem logically, even if it’s less efficient than other methods.

A reasoning AI reliably solves simple tests like Towers of Hanoi or chess that a small recursive program handles easily.

Has this happened?

Yes 72 (62%)Not sure 29 (25%)No 16 (14%)

Votes cast 1–2 October 2026: 100,590 votes from 9,694 people.