arXiv AI By Fabrice Lamarche, Gaya Mehenni, Neshat Elhami Fard, Odette Rios-Ibacache, Li Ming Wang, John Kildea, Amal Zouaq

MedHal: a Synthetic Dataset for Medical Hallucination Detection

Read the original on arXiv AI →

MedHal is a large-scale synthetic dataset created to detect hallucinations in medical AI-generated text. It includes diverse medical sources and tasks that cover both intrinsic and extrinsic hallucinations, providing a substantial volume of samples for training. The authors demonstrate that models trained on MedHal outperform general-purpose hallucination detectors, highlighting its usefulness for medical AI development.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv AI.

arXiv AI
Sep 15

Hallucination in Multimodal Foundation Models: A Survey on Causes, Corrections, and Evaluations

The article surveys hallucination issues in Large Vision‑Language Models (LVLMs), a type of multimodal foundation model that blends visual data with large language models. It categorizes hallucination causes into model architecture and data quality, presents a taxonomy of mitigation strategies, and critically evaluates existing evaluation benchmarks from both discriminative and generative viewpoints. The survey also outlines open challenges and future research directions to improve LVLM reliability and trustworthiness.

By Yinghao Guo, Wei Lan, Wenyi Chen, Qingfeng Chen, Shichao Zhang, Shirui Pan, Huiyu Zhou, Yi Pan