| 1 | Corrupt Plans, Clean Traces: Evading Chain-of-Thought Monitoring with Plan Injection | Keertana Chidambaram, Andrew Ilyas, Vasilis Syrgkanis | cs.AI | 2026-09-14 |
| 2 | EvoOntology: A Self-Evolving Ontology Layer for Data Agents | Meiduo Chong, Shaolei Zhang, Ju Fan +1 | cs.AI | 2026-09-14 |
| 3 | Look Before You Leap: Factual Decoding with Internal Attribution Signals | Hayeong Ryu, JungMin Yun, Byeonggeuk Lim +2 | cs.CL | 2026-09-14 |
| #4 | EEG-Xplain: Decoding Neural Black-Boxes of EEG Foundation Models | Hansong Ma, Junxiao Wang | cs.AI | 2026-09-14 |
| #5 | RESKILL: Explicit Failure Attribution and Structured Repair for Interactive Language Agents | Mengyi Deng, Xin Li, Duyi Pan +4 | cs.CL | 2026-09-14 |
| #6 | Authorship attribution and aesthetic evaluation of AI poetry: a case study with Haiku | Livia Oddi, Simone Scardapane, Toru Sugimoto +1 | cs.CL | 2026-09-14 |
| #7 | ChatGPT Images 2.5 in the Wild: A Launch-Period Dataset and Detector Evaluation | Dennis Ng, Xingyu Shen, Ankit Raj +6 | cs.CV | 2026-09-14 |
| #8 | Data Attribution at Scale via Influence Matrix Estimation | Yuxi Chen, Hamza Golubovic, Han Tong +2 | stat.ML | 2026-09-14 |
| #9 | IMPACT-VLA: Interaction-aware Multimodal Propagation Attribution via Counterfactual Trajectories for Vision-Language-Action Policies | Jinwoong Kim, Sangjin Park | cs.RO | 2026-09-14 |
| #10 | Shapley Value Estimation for Multi-Site Data with Blockwise-Missing Features | Siqi Li, Wangxuan Fan, Yiming Li +2 | stat.ML | 2026-09-14 |
| #11 | ModularRSI: Modular and Generalizable Recursive Harness Self-Improvement | Siwei Wu, Jincheng Ren, Yizhi Li +11 | cs.CL | 2026-09-14 |
| #12 | Calibrating Interpretability Instruments Before Trusting Their Verdicts | Orion Reblitz-Richardson | cs.LG | 2026-09-13 |
| #13 | From Visual Attribution to Clinical Reasoning: Explainable Parkinson's Disease Screening from Hand-Drawn Patterns | Aritra Dey, Utsav Kumar Nareti, Chandranath Adak +3 | cs.CV | 2026-09-13 |
| #14 | ATTRICITE: Training an Open 4B Model for Citation Recovery toward Faithful Attribution | Yee Man Choi, Xuehang Guo, Songcheng Cai +3 | cs.DL | 2026-09-13 |
| #15 | The Attribution-Compression Frontier in Retrieval-Augmented Generation | Deepanshu Mody | cs.CL | 2026-09-13 |
| #16 | T-SMART: Mechanism-Level Attribution for Tool-Augmented Time-Series Question Answering | Ivan Delgado, Himansi Gupta, Bishal Khatri +4 | cs.LG | 2026-09-12 |
| #17 | Exact Finite Attention Responses From RoPE Derivatives | Julie Huang, Maggie Chlon, Gregory Gutin +1 | stat.ML | 2026-09-12 |
| #18 | Map Users and Mapmakers: The Scope of Cognitive Attribution from Acquired Representations | Yiling Wu | cs.AI | 2026-09-12 |
| #19 | PolicyMem: Geometric Policy Memory for LLM Governance | Yuanchen Bei, Zhengzhang Chen, Yanjun Zhao +3 | cs.CL | 2026-09-12 |
| #20 | GeoSkill:Experience-Driven Hierarchical Skill Learning with Collaborative Revision forGeospatialAgents | Han Luo, Xian Xu, Yinhe Liu +1 | cs.AI | 2026-09-12 |
| #21 | Magenta: Closing the Loop Between Mathematical Reasoning and Lean Verification | Joshua Ong Jun Leang, Haonan Li, Zheng Zhao +6 | cs.AI | 2026-09-10 |
| #22 | 2AM: Grounding Agent-Side Memory as Guidance for Steerable Action Models in Long-Horizon Manipulation | Yutong Hu, Fengjiao Chen, Xuezhi Cao +1 | cs.RO | 2026-09-10 |
| #23 | X-RACE: XAI-assisted Recurrent neural network Attribution for Channel Estimation | Abdul Karim Gizzini, Yahia Medjahdi | eess.SP | 2026-09-10 |
| #24 | A Multi-View and Confusion-Guided Ensemble Framework for Robust Synthetic Image Attribution | Zuomin Qu | cs.CV | 2026-09-10 |
| #25 | The information geometry of large language models is shared, learned, and controllable | Dario Picozzi | cs.LG | 2026-09-10 |
| #26 | TRACE: Training Reasoning Agents for Causal Exploration with Synthesized Rewards | Rui Sun, Zhan Shi, Bing He | cs.AI | 2026-09-09 |
| #27 | AgentAudit: An Open, Extensible Framework for Full-Lifecycle Trust Evaluation of AI Agents | Shrey Nag, Sachita, Abhishek Kumar Singh +2 | cs.AI | 2026-09-09 |
| #28 | ReCite: Agentic Reasoning for Faithful Citation | Yuyang Huang, Bobo Li, Jiajia Song +4 | cs.CL | 2026-09-08 |
| #29 | Voice or Stereotype? Disentangling Acoustic and Content-Based Gender in Speech-to-Speech Models | Xiaoqun Liu, Tanu Mitra, Harshit Rajgarhia +1 | cs.SD | 2026-09-08 |
| #30 | MeClear: Cooperative Game-Theoretic Attribution and Risk-Aware Memory Clearance for Long-Horizon LLM Agents | Boyu Yang, Jiazheng Sun, Zilong Lu +3 | cs.AI | 2026-09-08 |
| #31 | The BatchNorm Illusion: Diagnosing Normalization Artifacts in Machine Unlearning Evaluation | Aaryaman Kalani, Murari Mandal, Dhruv Kumar +2 | cs.LG | 2026-09-08 |
| #32 | Charts Are Beyond Pixels: Probing for Layer-Wise Chart Understanding and Editing | Xiaochuan Zhong, Yifan Hou, Chenxi Pang +1 | cs.CV | 2026-09-08 |
| #33 | Tracing Stereotypes from Representation to Output in Multilingual LLMs | Ariun-Erdene Tumurchuluun, Yusser Al Ghussin, Pinzhen Chen +2 | cs.CL | 2026-09-08 |
| #34 | Vision: Data-Centric Anchoring for Robust and Interpretable Agentic AI | Arun Vignesh Malarkkan, Xinyuan Wang, Yanjie Fu | cs.AI | 2026-09-08 |
| #35 | Clean Accuracy Does Not Guarantee Provenance Robustness: A Prospective Codec-Stress Evaluation of Audio Attribution | Gang Shi | cs.SD | 2026-09-07 |
| #36 | Attributing Cohen's d: Training Data Attribution for Disease-Related Effects in Normative Age Biomarkers | Jakob Snel, Marc-Andre Schulz | cs.LG | 2026-09-07 |
| #37 | Noēsis: Deterministic-First Retrieval with Two-Tier Context Hydration for Factuality-Critical Queries on Small Local Models | Nicola Cogotti | cs.IR | 2026-09-07 |
| #38 | Human-like moral judgments conceal divergent motive attributions in large language models | Xiaoyan Wu, Jean-Claude Dreher | cs.AI | 2026-09-07 |
| #39 | From Synthetic Priors to Model Behavior: Structural Coverage in Tabular Foundation Models | He Zhao, Ryan Thompson, Daniel M. Steinberg +3 | cs.LG | 2026-09-07 |
| #40 | AuthBench: A Large-Scale Multilingual Benchmark for Authorship Representation across Genres and Lengths | MaoXun Huang, Zhenxing Zhang, Claire Cardie | cs.CL | 2026-09-06 |
| #41 | We Built a Mirror and Mistook It for a Mind: Causal Liability and the Fallacy of AI Consciousness | Afshin Khadangi | cs.AI | 2026-09-06 |
| #42 | Causal Attribution for Agentic Decisions: Estimators, Coupling, and a Traceability Specification | Ajay Pravin Mahale | cs.AI | 2026-09-06 |
| #43 | From Interpretability Methods to Interpretable Models | Julien Colin, Nuria Oliver, Thomas Serre | cs.CV | 2026-09-04 |
| #44 | Forgetting Without Restarting: Execution-State Unlearning for Stateful LLM Agents | Chao Yao, Yangbo Wei, Zhen Huang +5 | cs.CR | 2026-09-04 |
| #45 | Cost-Aware Hierarchical Multi-Agent Ransomware Detection and Family Attribution | Mubashar Iqbal, Asifullah Khan | cs.CR | 2026-09-04 |
| #46 | How Faithful Is Attribution for Sales Forecasting? A Counterfactual Study | Glib Kechyn | cs.LG | 2026-09-04 |
| #47 | An Attention-Guided Global and Local Fusion Framework for Lesion-Focused Image Classification | Mst Shafia Tasnima, Md Samaun Elaheea, Tanjim Taharat Aurpab +1 | cs.CV | 2026-09-04 |
| #48 | DCFA: Dual-view Causal-inspired Attribution for Failure Reasoning in LLM-based Multi-agent Systems | Zehao Wang, Lanjun Wang, Shilong Jin +2 | cs.AI | 2026-09-04 |
| #49 | A Computationally Feasible Framework for Causal Probabilistic Explanation | Rafal Urbaniak, Sam Witty, Daniel Waxman +7 | cs.AI | 2026-09-03 |
| #50 | DRACO: Fine-Grained Credit Assignment with Dynamic Rubrics for Long-Horizon Agent Training | Shubham Gandhi, Saurabh Goyal, Kiran Kate +1 | cs.AI | 2026-09-03 |