About which tasks LLMs still fail more often than average humans do.
Has this happened?
Yes 62 (48%)Not sure 29 (23%)No 37 (29%)
Votes cast 1–2 October 2026: 100,590 votes from 9,694 people.
Hacker News has set AI a lot of challenges over the years. Which ones has it met?
About which tasks LLMs still fail more often than average humans do.
Votes cast 1–2 October 2026: 100,590 votes from 9,694 people.
So far I've only seen examples based on tokenization ("what's the third letter in word?"), and the fact that answers are generated sequentially ("write a sentence that includes the count of words in itself"). But those feel unsatisfactory because those are not cognitive limitations.
It could solve those, the reason it can't is that it doesn't understand its own limitations, ie it has no self awareness.
Humans plan things based around their own limitations, if an AI model can't even do that how can it be considered an AGI?