Zhe Yu, Mohan Li, Lei Yu, Ka-Ho Chow, Chengwei Qin, Xingyu Wu, Wenpeng Xing, Shuguang Xiong, Meng Han · arXiv (Cornell University) 2026 · 2026
DOI: 10.48550/arxiv.2610.05472
Counts differ because each database indexes a different set of publications. We treat OpenAlex as the canonical count; Google Scholar is not shown (no API, and crawling it violates its ToS).
Reasoning errors can propagate into later decisions and memory. This survey synthesizes 312 papers and first-party reports on text-based reasoning hallucinations around three questions: what evidence is observable, what study designs establish, and which corrective actions the evidence supports. UIPCA records unsupported premises (U), invalid inferences (I), dependent reuse (P), visible answer-trace consistency (C), and action-policy failures (A). Across 58 reviewed sources, no comparison establishes that a specified intervention improves reasoning while reducing factual reliability under matched conditions. The synthesis connects diagnosis to verification, repair, selective release, and persistent-state control across memory, tools, and training feedback.
No comments yet — start the discussion below.