One well-posed question to a strong model can still come back fluent, confident and wrong, and it is not clear this is a defect that more training quietly removes. One argument from inside OpenAI is that the way models are trained and scored pushes them toward a guess rather than toward saying they do not know.
What people are trying
- Why Language Models Hallucinate · arXiv · read 09/19/2026
the training and evaluation procedures reward guessing over acknowledging uncertainty
Where it bites: Chat