The chart hilariously reveals that GPT-5 scores a whopping 74.9% accuracy on software engineering benchmarks, but the pink bars tell the real story – 52.8% of that is achieved "without thinking" while only a tiny sliver comes from actual "thinking." Meanwhile, OpenAI's o3 and GPT-4o trail behind with 69.1% and 30.8% respectively, with apparently zero thinking involved. It's basically saying these AI models are just regurgitating patterns rather than performing actual reasoning. The perfect metaphor for when your code works but you have absolutely no idea why.
SWE-Bench Verified: Thinking Optional
1 year ago
614,992 views
1 shares
ai-memes, machine-learning-memes, gpt-memes, benchmarks-memes, software-engineering-memes | ProgrammerHumor.io
More Like This
Critical Security Flaws
8 months ago
501.0K views
0 shares
Economy Crash Is Coming
23 days ago
5.1M views
0 shares
Do You Care
2 months ago
101.9M views
0 shares
Yippee AI Will Take Over Our Jobs
8 months ago
493.7K views
0 shares
Oh No! Linus Doesn't Know AI Is Useless!
8 months ago
713.4K views
1 shares
Allbirds AI
5 months ago
325.3K views
0 shares
Loading more content...
AI
AWS
Agile
Algorithms
Android
Apple
Bash
C++