Goalposts

Hacker News has set AI a lot of challenges over the years. Which ones has it met?

About whether LLMs really reason when they solve math word problems.

They do not "solve" word puzzles. They output text that appears to be the best response to the prompt, based on their training data. If the puzzle is solvable by doing this, then they get the answer right. If the puzzle is not solvable doing that, they are unlikely to get the answer right.

If I ask you to multiply two (largeish) numbers together, you will be able to do so, using an algorithm/process that you can apply to the multiplication of any two numbers, whether anyone has ever told you about those numbers before or not.

LLM's cannot do this. Give them a math problem that doesn't exist in their training set and they cannot solve it. This has been demonstrated many times.

An LLM multiplies large numbers it has never seen by applying a general algorithm, like a human does.

Has this happened?

Yes 79 (67%)Not sure 22 (19%)No 17 (14%)

Votes cast 1–2 October 2026: 100,590 votes from 9,694 people.