| 1 | Epistemic Warrant for LLM Recommendations: Characterizing the Basis for Reliance When Ground Truth Is Unavailable | Shai Vardi, João Sedoc | cs.AI | 2026-09-03 |
| 2 | Translation as a Decision Space: A Multi-Agent Perspective on Low-Resource Dialect Generation | Hasan Alkhder, Mohammad Abboush, Igor Tchappi +2 | cs.CL | 2026-09-03 |
| 3 | The Head Complexity of Boolean Functions in Single-Layer Attention | Rajmohan Rajaraman, Ravi Sundaram, Amanuel Tesfaye | cs.CC | 2026-09-03 |
| #4 | The Dually Flat Geometry of Planning as Inference | Nikola Milosevic, Asaki Kataoka, Nicolas Hinrichs +2 | cs.AI | 2026-09-03 |
| #5 | WorldReward: Reward Modeling for Camera-Conditioned World Models | Yibin Wang, Zehan Wang, Junshu Tang +13 | cs.CV | 2026-09-03 |
| #6 | Headroom-Drift Replay: A Primitive for Principled Replay Control in GRPO | Hyun Bin Park, Du-Seong Chang | cs.LG | 2026-09-03 |
| #7 | Speak for Me: Giving LLMs the Situational Awareness to Participate in a Meeting | Muneeb Khan, Frederic Kirstein, Terry Ruas +1 | cs.AI | 2026-09-03 |
| #8 | Value-Preserving Architectures for Agentic AI Systems | Alessandro Pesare, Tommaso Dolci, Katja Hose +1 | cs.AI | 2026-09-03 |
| #9 | Beyond Shallow Alignment: How Post-Training Methods Determine Refusal Circuits And Steering Robustness | Hoang Cuong Nguyen, Mark Dras, Usman Naseem | cs.CL | 2026-09-03 |
| #10 | Adapting to Evolving Requirements: Agentic AI for Retail Supply Chain Operations | Lei Zheng, Liping Yang, Zihao Li +3 | cs.AI | 2026-09-03 |
| #11 | Pushing the (Decision) Boundaries: Dynamically Calibrating Differentially Private Noise to Explainability in Federated Learning | Michael Khavkin, Kichang Lee, Jaeho Jin +2 | cs.LG | 2026-09-03 |
| #12 | EF1-Constrained Nash Social Welfare with Identical Additive Valuations: Complexity, Guarantees, and Experiments | Zih-Sian Yang, Yi-Hao Chen, Yu-Te Kuan +3 | cs.GT | 2026-09-03 |
| #13 | Select, Compress, Reinvest: A Controlled Study of Visual-Token Allocation in Long-Video MLLMs | Prakhar Khatri | cs.CV | 2026-09-03 |
| #14 | Evaluating Criterion-Conditioned Behaviour of Large Language Models in Content Moderation | Danting Zhang, Bei Peng, Robert Loftin | cs.CL | 2026-09-03 |
| #15 | DNative-Twin: Decision Graphs and Digital Twins for Reconstructable Agentic Decisions | Junjie Pang, Zhenzhen Xie, Haoke Han +3 | cs.AI | 2026-09-03 |
| #16 | Rethinking World Models for Safety-Critical Embodied Systems | Kailang Ma, Heye Huang, Inhi Kim +1 | cs.AI | 2026-09-03 |
| #17 | OBER+: Continuity-Aware Reporting and Traceable Continuous Improvement in Outcome-Based Education | Elakkiya Rajasekar | cs.LG | 2026-09-03 |
| #18 | Proactive Service Agents: A Unified Decision Framework, Methods, and Evaluation | Yan Tang, Tingyu Cao, Yuanbo Tang +2 | cs.AI | 2026-09-03 |
| #19 | Can LLMs Extract Architectural Design Decisions from Source Code Commits? - A Preliminary Exploratory Study | Amey Karan, Rudra Dhar, Mohamed Soliman +1 | cs.SE | 2026-09-03 |
| #20 | The Impact of Synthetic Data Augmentation on Discourse-Pragmatic Function Classification | Sara Sorahi, Kevin Tang, Reza Kazemian | cs.CL | 2026-09-03 |
| #21 | Tree-Structured Vector Quantization For Efficient And Progressive Image Compression | Xinkun Wang, Tianyi Xu, Qingyu Luo +4 | cs.CV | 2026-09-03 |
| #22 | A computable representation of the physical laboratory enables verifiable workflows | Xiaobo Li, Luyao Ge, Xiaohui Li +8 | cs.AI | 2026-09-03 |
| #23 | ToolDF: Tool-Integrated Reasoning for Mixed-Authenticity Audio Deepfake Detection | Taewoo Kim, Young Han Lee, Nam In Park +1 | eess.AS | 2026-09-03 |
| #24 | Drive-HWM: Hierarchical World Models for Dynamic-Latent Guided Autonomous Driving | Zhaoxin Fan, Tianbao Zhang, Wenjun Wu +5 | cs.CV | 2026-09-03 |
| #25 | Building Pretraining Data for World Models: An Unreal Engine-Based Pipeline for Action-Conditioned Video Generation | Haoyu Wang, Songchun Zhang, Haoran Li +3 | cs.CV | 2026-09-03 |
| #26 | EPIC: Explicit Posterior Item Conditioning for Semantic ID Diffusion Recommendation | Tuan-Binh Tran, Thanh Tam Nguyen, Quoc Viet Hung Nguyen +3 | cs.IR | 2026-09-03 |
| #27 | When Retrieval Helps: Selective Retrieval for Single-Turn Mental-Health QA | Hyunseo Oh, Chong-Kwon Kim, Yoonhyuk Choi | cs.CL | 2026-09-03 |
| #28 | Preprocessing Failure and Adversarial Detection in Depthwise-Separable Edge Vision Systems | Jannatul Masruk Mukta, Rifa Sanjida, Adrita Rahman Tory +2 | cs.CV | 2026-09-03 |
| #29 | Plan Pointers and Record-Directive Form in Budgeted Verification of Inherited Agent Memory | Kazuki Nakayashiki | cs.IR | 2026-09-03 |
| #30 | Accountable AI with Grounded, Faithful, Consistent, Actionable Rationales: A Case Study in Clinical Trial Matching with VERDICT | Zikai Zhou, Yufei Jin, Yilin Xu +3 | cs.CL | 2026-09-03 |
| #31 | DE-Venus: A Data-Efficient RLVR Framework for Large Language Models | Shenzhi Yang, Guangcheng Zhu, Kai Tang +11 | cs.LG | 2026-09-03 |
| #32 | PACE: Towards Surfacing Hidden Conflicts in User Requests | Yoojin Kim, Jihyoung Jang, Hyounghun Kim | cs.CL | 2026-09-03 |
| #33 | Speculative Macro Commit for Faster Tool-Using Agents | Zeyu Liu, Souvik Kundu, Peter A. Beerel | cs.AI | 2026-09-03 |
| #34 | Counterfactual Fairness Audits of Multi-Step Clinical LLM Agents Require a Measured Per-Action Instability Floor | Rohith Reddy Bellibaltu, Manpreet Singh, Deepak Parashar +1 | cs.CL | 2026-09-02 |
| #35 | The Analyst in the Prompt: Role, Retrieval, and Memory Biases in LLM Financial Analysis | Ahmed Asaad, Amr Mohamed, Yang Zhang +1 | cs.CL | 2026-09-02 |
| #36 | VoxReason: Listener-Free Evaluation of Source-Grounded Speech Planning Before Synthesis | Mengzhe Geng | cs.SD | 2026-09-02 |
| #37 | VeriPhy: Agentic Physical Reasoning for World Model Evaluation and Refinement | Wenzhuo Xu, Yuchen Zhu, Chongjian Ge +8 | cs.CV | 2026-09-02 |
| #38 | Feasible but Not Safe: Constraint Violations and Report-Channel Attacks in Learned Cell-Free ISAC Association | Mehdi Zafari, Iman Mohammadi, A. Lee Swindlehurst | cs.NI | 2026-09-02 |
| #39 | LeanStream: A Speculate-and-Refine Streaming Framework for Efficient on-Device LLM Inference | Renyuan Liu, Yuyang Leng, Kaiyan Liu +7 | cs.LG | 2026-09-02 |
| #40 | Population-Calibrated Graph Screening at 835-Million-Address Scale, with Label-Free Transfer to New Chains | Yury Korolev | cs.CR | 2026-09-02 |
| #41 | ObserverBench: Testing Mechanistic Estimates for Intervention and Control | Vijay Erramilli | cs.LG | 2026-09-02 |
| #42 | SolarWM: Open Data and Scalable Training for Long-Horizon Video World Models | Junchao Huang, Guian Fang, Shengju Qian +15 | cs.CV | 2026-09-02 |
| #43 | Discriminative World Models for Web Agents | Kelvin Li, Dhruv Pendharkar, Anish Pahilajani +6 | cs.AI | 2026-09-02 |
| #44 | Towards Trustworthy Autonomous Robots: An Explainable AI-Based Decision Framework | Cagri Temel | cs.RO | 2026-09-02 |
| #45 | Large Language Models (LLMs) for Telecom Root Cause Analysis (RCA): A Structured Reasoning Framework for Evidence-Grounded Diagnosis | Hao Zhou, Mandar Kulkarni, Hao Chen +3 | cs.AI | 2026-09-02 |
| #46 | Dutch Books for Language Models | Isaiah Andrews, Suproteem Sarkar | econ.GN | 2026-09-02 |
| #47 | Untangling the Mechanisms of Misleading Context in Medical Question Answering | Robin Linzmayer, Noémie Elhadad | cs.CL | 2026-09-02 |
| #48 | CORAL: An LLM-Native Harness for Production Recommender Systems | Muhammad Rafay Azhar, Yuhang Zhou, Gilbert Jiang +7 | cs.CL | 2026-09-02 |
| #49 | Toward Collective-Centric Evaluation of Preference Inference for Participatory Democracy | Pierre-Antoine Lequeu, Salim Hafid, Paul Lerner +6 | cs.SI | 2026-09-02 |
| #50 | Generating Medical Image Counterfactuals using Causal Explanations | David A. Kelly, Tom Yaacov, Nathan Blake +2 | cs.CV | 2026-09-02 |