| 1 | Domain-Specific Hallucination Detection in Large Language Models | Varun Teja Chundru, Debasmita Biswas | cs.CL | 2026-09-10 |
| 2 | OmniHallu: Unified Hallucination Detection for Cross-Modal Comprehension and Generation in Multimodal Large Language Models | Jianjiang Yang, Peihang Li, Shanqing Xu +3 | cs.CL | 2026-09-10 |
| 3 | Two-Token Features and Small-Large Ensembles for VLM Hallucination Detection | Eli Schwartz | cs.CL | 2026-09-09 |
| #4 | Improving Cross-Lingual Token Representations by Adding a Pinch of SALT | Guillem Ramírez | cs.CL | 2026-09-09 |
| #5 | Can We Trust Video Hallucination Detectors? VidHalLoc for Evaluating the Evaluators | Xinyu Chen, Adnan Mahmood, Mark Dras | cs.CV | 2026-09-09 |
| #6 | CT-SAFR: Safe and Interpretable Chain-of-Thought Reasoning for Autonomous Robots: A Multi-Layered Verification Framework for Trustworthy AI-Driven Robotic Decision Making | Cagri Temel | cs.RO | 2026-09-09 |
| #7 | Evidence-Aligned Entity Verification for Hallucination Detection in Retrieval-Augmented Generation | Runsong Jia, Zhen Fang, Mengjia Wu +2 | cs.AI | 2026-09-08 |
| #8 | Leveraging Low-Level Symbolic Competences for Unsupervised Grounding in Hallucination Detection | Renato Vukovic, Hsien-chin Lin, Carel van Niekerk +5 | cs.CL | 2026-09-04 |
| #9 | Beyond Majority Vote: Multi-Perspective Adjudication for Medical Hallucination Detection | Joe Cecil, Marjorie Freedman | cs.CL | 2026-09-03 |
| #10 | From Tokens to Semantics: Leveraging Complementary Signals for Hallucination Detection in Black-Box LLMs | Urja Pawar, Rajitha Ramanayake, Owen O'Neill +4 | cs.CL | 2026-09-02 |
| #11 | Detecting Object Hallucinations in Large Vision-Language Models via Cross-Modal Attention Drifts and Mask-Based Verification | Xuanbing Wen, Boxu Chen, Le Yang +4 | cs.CV | 2026-09-02 |
| #12 | Enoki: Efficient Multi-Level Hallucination Detection | Elisei Rykov, Timur Ionov, Nikolay Ivanov +5 | cs.CL | 2026-09-01 |
| #13 | VisER: Visual Evidence and Reliance for Object Hallucination Detection in LVLMs | Afsaneh Hasanebrahimi, Hanxun Huang, Christopher Leckie +1 | cs.CV | 2026-08-31 |
| #14 | ImageEval 2026: Culturally Grounded Arabic Multimodal Evaluation | Samir Abdaljalil, Hunzalah Hassan Bhatti, Ahlam Bashiti +11 | cs.CL | 2026-08-31 |
| #15 | Prediction of Prediction (PoP): Inter-Layer Activation Fusion for Single-Pass Hallucination Detection in Large Language Models | Himal Badu | cs.CL | 2026-08-27 |
| #16 | Overview of SHROOM-Visions 2026: A Shared Task on Hallucination Detection in Large Vision-Language Models | Raúl Vázquez, Aman Sinha, Chuyuan Li +12 | cs.CL | 2026-08-26 |
| #17 | Lost in Speech: Trilingual Spoken Hallucination Detection Across Audio and Transcripts | Meruyert Aristombayeva, Jason S. Lucas, Chaewan Chun +1 | cs.CL | 2026-08-25 |
| #18 | When Do Supervised UQ Ensembles Improve LLM Hallucination Detection? A Robustness Study | Mohit Singh Chauhan, Vipin Gyanchandani, Dylan Bouchard | cs.LG | 2026-08-25 |
| #19 | Credal Large Language Models for Semantic Commitment under Uncertainty | Shireen Kudukkil Manchingal, Sofiia Nikolenko, Fabio Cuzzolin | cs.CL | 2026-08-24 |
| #20 | Enforcing LLM Safety through DMD-based Classification of Prompt-Response Embedding Dynamics | Mohamed Akrout, Olivera Kotevska, Dan Wilson | cs.AI | 2026-08-20 |
| #21 | Do Large Language Models Play Six Degrees of Separation? Measuring Topological Compression in Long-Context Manifolds | Md. Faiyaz Abdullah Sayeedi | cs.CL | 2026-08-18 |
| #22 | BEAR-Bench: A Bilingual Enterprise and Academic Reasoning Benchmark for Multimodal Models | Liubov Chubarova, Alexandra Kuleshova, Daniil Volkov +2 | cs.CL | 2026-08-18 |
| #23 | Mixture-of-Expert Blocks Contain Strong Hallucination Detection Signals | Joao Fonseca, Rodrigo Rodrigues, Paolo Romano | cs.AI | 2026-08-18 |
| #24 | HalluTracer: Hallucination Detection via Depth-Averaging Truth Signals | Zhihao Guo, Zonghan Wu, Huan Huo +6 | cs.CL | 2026-08-17 |
| #25 | Hallucination Span Detection with Input-Side Evidence Alignment | Miyu Yamada, Yuki Arase | cs.CL | 2026-08-16 |