@thesmokingman@programming.dev
Post #2022544
2026-05-04 07:06 UTC
Sorry, I assumed you would have actually read [the DELEGATE-52 study](https://arxiv.org/abs/2604.15597) linked instead of just the abstract. For “a model optimized for replacing StackOverflow” that is “better at writing papers than most students” LLMs sure did pretty bad at those tasks over multiple rounds.
Replies (1)
-
@turdas@suppo.fi 2026-05-04 07:16
As the chart on page 7 of the paper shows, LLMs are good at exactly the kind of tasks you'd expect (producing and manipulating language), and bad at exactly the kind of tasks you'd expect (doing almost anything else). All this paper shows is that (1) they aren't AGI, and (2) as a consequence of not being AGI they aren't good unsupervised. Why do you lie like this?