Post #2102981
2026-02-23 23:26 UTC
And that score is matched by GPT-5. Humans are running out of "tricky" puzzles to retreat to.
Replies (4)
-
@First_Thunder@lemmy.zip 2026-02-24 00:21
What this shows though is that there isn’t actual reasoning behind it. Any improvements from here will likely be because this is a popular problem, and results will be brute forced with a bunch of data, instead of any meaningful change in how they “think” about logic
-
@CileTheSane@lemmy.ca 2026-02-24 18:59
> Humans are running out of "tricky" puzzles to retreat to. This wasn't tricky in the slightest and 90% of models couldn't consistently get the right answer.
-
@XLE@piefed.social 2026-02-24 13:22
You don't need to do the dehumanizing pro-AI dance on behalf of the tech CEOs, Facedeer
-
@realitista@lemmus.org 2026-02-24 00:26
You're getting downvoted but it's true. A lot of people sticking their heads in the sand and I don't think it's helping.