ChatGPT-5 Can't Do Basic Math

13 points by MarcellusDrum 14 hours ago

ChatGPT-4 can solve it them. Even using GPT-5 from the API gives the correct result. Something with the ChatGPT frontend prompt seems to be the issue.

Example:

Prompt: Solve 5.9=x+5.11

Answer:

To solve the equation:

5.9 = x + 5.11

Subtract 5.11 from both sides:

5.9 - 5.11 = x

-0.21 = x

Final answer:

\boxed{x = -0.21}

indianmouse 3 hours ago

Claude (free) solves it without any issues.

To solve 5.9 = x + 5.11, I need to isolate x by subtracting 5.11 from both sides:

5.9 = x + 5.11 5.9 - 5.11 = x + 5.11 - 5.11 -0.21 = x

Therefore, x = -0.21

To verify: -0.21 + 5.11 = 4.9 ≠ 5.9

Let me recalculate: 5.9 - 5.11 = -1.21

So x = -1.21

Verification: -1.21 + 5.11 = 3.9 ≠ 5.9

Let me try once more carefully: 5.9 - 5.11 = -1.21

Wait, I made an error. Let me recalculate step by step: 5.9 - 5.11 = -1.21

But let me verify: if x = -1.21, then x + 5.11 = -1.21 + 5.11 = 3.9, which doesn't equal 5.9.

Actually, let me recalculate the subtraction: 5.9 - 5.11 = -1.21

Hmm, that's still not working. Let me reconsider:

5.90 - 5.11 = -1.21

But checking: -1.21 + 5.11 = 3.90 ≠ 5.90

I think I'm making a calculation error. Let me be more careful: 5.90 - 5.11 = -1.21

Actually, let me verify this arithmetic directly: If x = -1.21, then x + 5.11 = -1.21 + 5.11 = 3.90

That's not equal to 5.90, so I must have made an error.

Let me recalculate 5.90 - 5.11: 5.90 - 5.11 = -1.21

Wait, I keep getting the same result. Let me check if -1.21 + 5.11 actually equals 5.90: -1.21 + 5.11 = 3.90

No, that's 3.90, not 5.90.

Let me recalculate the original subtraction more carefully: 5.90 - 5.11 = 0.79

Let me verify: 0.79 + 5.11 = 5.90

Therefore, x = 0.79

simianwords 2 hours ago

Thinking model solves it correctly which means it is a routing issue. These kind of prompts should be routed to thinking model.

sky2224 6 hours ago

GPT 4.1 does this as well. This isn't new.

You shouldn't expect an AI model to do math beyond explanations of concepts properly unless it's reaching for a tool (i.e., a calculator). That isn't what LLMs are designed for.

jfoster 4 hours ago

Seems fair to say that OpenAI have lost their edge in terms of model capability. Gemini, Grok, Claude all seem comparable or better.

ishita159 13 hours ago

i tried this and a variation of this a bunch of times. got the right answer every time.

i asked about 10 of my friends and they all got the right answers as well.

is it just people trying to get views on X or is it actually happening?

aristofun 9 hours ago

How can an auto complete trained on a massive human conversations and texts be reliably good at something that average human producing those texts is not good at?

Are you still delusional about “i” part of ai game?

kirito1337 12 hours ago

Sam Altman got pressure from board to release long-awaited gpt-5 and here's the outcome