MHB: Medical Hallucination Benchmark for Large Language Models in Complex Clinical Tasks

Content not yet available

This lecture has no active video or poster.

AAAI 2026

•

January 25, 2026

•

Singapore, Singapore

Please log in to leave a comment

Next from AAAI 2026

Multi-Agent VLMs Guided Self-Training with PNU Loss for Low-Resource Offensive Content Detection
poster

Multi-Agent VLMs Guided Self-Training with PNU Loss for Low-Resource Offensive Content Detection

AAAI 2026

+6
Deyi Ji and 8 other authors

25 January 2026

Similar lecture

CliMedBench: A Large-Scale Chinese Benchmark for Evaluating Medical Large Language Models in Clinical Scenarios
poster

CliMedBench: A Large-Scale Chinese Benchmark for Evaluating Medical Large Language Models in Clinical Scenarios

EMNLP 2024

Gerard de Melo
Liang He
+5
Zetian Ouyang and 7 other authors

13 November 2024