Topic: hallucination

4 stories found

Friday, August 28, 2026

research40

Can a Model Catch Its Own Hallucinations for Free?: Label-Free Doubt Signals Hold Their Own Against a Labelled Dataset for Abstention

A new method allows large language models to identify and abstain from answering questions they are uncertain about using their internal probabilities, potentially reducing errors without needing additional labeled data. This technique is significant because it could enhance the reliability of AI systems by enabling them to recognize when they lack sufficient knowledge to provide accurate responses.

arxiv.orgโ†—

Wednesday, August 26, 2026

research35

Gated Activation Steering for Reducing Sycophancy & Hallucination in Medical Question Answering

A new method called Gated Activation Steering is proposed to reduce sycophancy and hallucination in large language models used for medical question answering, ensuring responses are contextually accurate. This is crucial because such errors can have severe consequences in clinical settings where precise information is essential.

arxiv.orgโ†—

Tuesday, August 25, 2026

research40

On the Role of Citations in Preference Data

The paper discusses the importance of including citations in NLP system outputs, arguing that attributions are crucial for preventing model hallucinations and allowing users to verify information, thereby enhancing the reliability of AI-generated content.

arxiv.orgโ†—

๐ŸŒฟ That's all for now. Come back tomorrow.

4 of 4 items shown. Sources: 107 days indexed.