Post #3734442
2026-07-11 07:19 UTC
I studied transformer architecture models and have played around with them (unfortunately) enough to understand how they work. Under the surface the model produces what look like XML tags to designate which tokens are thinking tokens and which are “normal” output. That is literally the only hard difference between the two output modes. The reinforcement learning might tune the thinking to be more like “what a human would expect to see in a thinking block” but it’s still the same RNG madlib process generating everything underneath and any attempt to ascribe intelligence to this process should be met with ~lethal force~ incredulous cynicism.
Replies (1)
-
@BioMan@awful.systems 2026-07-12 02:38
God I remember having to explain to dozens of people that ‘reasoning’ models just exude a lot of text ‘talking to themselves’ and then summarize it. They were all just “It CANT be that silly” and many outright would not believe me, because that was not ‘reasoning’