Post #2198902
2026-05-07 16:22 UTC
ten years ago there was some interesting work on building deep classification models that not only made a classification but also made a confidence estimate (and not just taking the softmax score as a pseudo confidence). I wonder why something like that isn’t used for LLMs, maybe an auxiliary lighter weight LLM that’s been trained explicitly as a bullshit detector. A bit like LLM as a judge but more principled. (1/2)
Replies (1)
-
@dx@social.ridetrans.it 2026-05-07 16:22
I’m sure it’s been done/tried, since this isn’t a groundbreaking idea, and there must be some reason it isn’t used or isn’t effective. (2/2)