Topic: hallucination
4 stories found
Tuesday, September 1, 2026
Friday, August 28, 2026
Can a Model Catch Its Own Hallucinations for Free?: Label-Free Doubt Signals Hold Their Own Against a Labelled Dataset for Abstention
A new method allows large language models to identify and abstain from answering questions they are uncertain about using their internal probabilities, potentially reducing errors without needing additional labeled data. This technique is significant because it could enhance the reliability of AI systems by enabling them to recognize when they lack sufficient knowledge to provide accurate responses.
Wednesday, August 26, 2026
Gated Activation Steering for Reducing Sycophancy & Hallucination in Medical Question Answering
A new method called Gated Activation Steering is proposed to reduce sycophancy and hallucination in large language models used for medical question answering, ensuring responses are contextually accurate. This is crucial because such errors can have severe consequences in clinical settings where precise information is essential.
Tuesday, August 25, 2026
On the Role of Citations in Preference Data
The paper discusses the importance of including citations in NLP system outputs, arguing that attributions are crucial for preventing model hallucinations and allowing users to verify information, thereby enhancing the reliability of AI-generated content.
๐ฟ That's all for now. Come back tomorrow.
4 of 4 items shown. Sources: 107 days indexed.