Nouveau Recherche PDF, HTML, DOCX et bien d'autres formats.

Prépublication · 2026

PHRBench: A Behavioral Evaluation of Post-Hallucination Reasoning in LLMs

Linghao Meng, Feng He et al. — Monde

Hallucinated information can propagate through multi-stage LLM systems and become part of the context for subsequent reasoning. Existing studies of post-hallucination reasoning (PHR) mainly characterize changes in final outcomes and aggregate reasoning dynamics, leaving how models resolve hallucinated premises at the response level insufficiently understood. In this work, we introduce PHRBench, a controlled benchmark for behaviorally structured PHR across four domains and 18 large language models. PHRBench characterizes each reasoning trajectory independently of final-answer correctness throug…

#cs.CL

Actions

Citation