About a better AGI benchmark: telling an AI to go make money.
Has this happened?
Yes 8 (6%)Not sure 29 (22%)No 97 (72%)
Votes cast 1–2 October 2026: 100,590 votes from 9,694 people.
Hacker News has set AI a lot of challenges over the years. Which ones has it met?
About a better AGI benchmark: telling an AI to go make money.
Votes cast 1–2 October 2026: 100,590 votes from 9,694 people.
It’s a good question. I think a better benchmark than the current options is “go make $X dollars this quarter.” Right now the models fail this miserably. Claude can’t even run a vending machine inside Anthropic HQ. So there is still some kind of strategic activity that comes naturally to humans that LLMs struggle with. I know the big conundrum is “scaling solves this in the next N years” but my bet is that N > ~20 in this case.