Trey Causey

Writing, advising, and seeking a next role in AI. Based in Seattle.

1 min

Source cdn.openai.com

Why Language Models Hallucinate

“[L]anguage models hallucinate because the training and evaluation procedures reward guessing over acknowledging uncertainty”

"Why Language Models Hallucinate", released by Kalai, Nachum, Vempala, and Zhang (mostly of OpenAI) last week.

It has both a potentially surprising conclusion and a definitely (to me) surprising recommendation for solutions — surprising to me because I made a similar point a couple of months ago!

The tl;dr for the conclusion is that language models hallucinate because of the way they are trained and evaluated. No, not because of next-token prediction necessarily, but because of the incentives to provide an answer. Language models are “optimized to be good test-takers” where saying “I don’t know” is not rewarded.

The authors demonstrate that, largely independent of architecture, even in situations where only correct factual information is present in the training data models will still be incentivized to guess incorrectly.

The proposed solution is (naturally) socio-technical by changing how models are scored and disincentivizing guessing over expressing uncertainty.

See: Solving LLM Hallucinations is (mostly) a UX Problem”

Share on XShare on LinkedInMarkdown for an LLM

Send your thoughts

Name and email are optional.

At least 10 characters.

Was this helpful or interesting?

All notes

ESC
Type to search...