| 1 | Seeing Before Synthesizing: VLM-Guided Transition Event Discovery for Weakly-Supervised Dense Video Captioning | Ye-Chan Kim, Seunghee Choi, SeungJu Cha +4 | cs.CV | 2026-09-03 |
| 2 | A Computationally Feasible Framework for Causal Probabilistic Explanation | Rafal Urbaniak, Sam Witty, Daniel Waxman +7 | cs.AI | 2026-09-03 |
| 3 | Epistemic Warrant for LLM Recommendations: Characterizing the Basis for Reliance When Ground Truth Is Unavailable | Shai Vardi, João Sedoc | cs.AI | 2026-09-03 |
| #4 | The Dice Roll Method: A Standardized Protocol for Repeated-Query Auditing of Large Language Model Brand Recommendations | Dmitrij Żatuchin | cs.IR | 2026-09-03 |
| #5 | InSituMeasure: Probing Situated Measurement Grounding in Industrial Scenes with Multimodal Large Language Models | Chao Shen, Xinyuan Li, Yunfan Zhou +4 | cs.AI | 2026-09-03 |
| #6 | FiMI Banking: A Sovereign Model for Indian Retail Banking | NPCI AI Research Team, Aman Kumar, Asit Desai +15 | cs.AI | 2026-09-03 |
| #7 | Bioinfoysis Technical Report | Qingyang Shao, Xin Zhang, Zhouyang Yuan +24 | cs.AI | 2026-09-03 |
| #8 | Urban Boundaries, Social Barriers: A Benchmark and Vision-Centric Framework for Mapping Gated Communities and Equity Implications | Minwei Zhao, Weiming Zhang, Jiawang Du +4 | cs.CV | 2026-09-03 |
| #9 | SimSkill: A Lifelong Learning AI Agent for Autonomous Mastery of Traffic Simulation | Qi Liu, Qinzheng Wang, Yiming Bie | cs.AI | 2026-09-03 |
| #10 | KnowVis: Knowledge-Centric Visual Summarization for Video Lectures | Yi Xu, Yifan Hou, Xiaoyu Zhang | cs.CV | 2026-09-03 |
| #11 | MetaStructAtlas: A Grounded 3D Vision-Language Dataset and Benchmark for Functional and Structural Reasoning in Whole-Body PET/CT | Chenguang Zheng, Le Xue, Yichi Zhang +8 | cs.CV | 2026-09-03 |
| #12 | Enhancing Financial Question Answering: A Novel Benchmark Dataset of Banks' financial statements | Arianna Miola, Bruno Spaccavento, Lorenzo Silotto +2 | cs.CL | 2026-09-03 |
| #13 | The Attention Triangle in Audio-Video Models | Sagi Polaczek, Noa Kraicer, Gal Metzer +4 | cs.AI | 2026-09-03 |
| #14 | Text2Thermal: Physics-Aware Thermal Image Synthesis from Textual Priors | Tayeba Qazi, Brejesh Lall, Prerana Mukherjee | cs.CV | 2026-09-03 |
| #15 | Drive-HWM: Hierarchical World Models for Dynamic-Latent Guided Autonomous Driving | Zhaoxin Fan, Tianbao Zhang, Wenjun Wu +5 | cs.CV | 2026-09-03 |
| #16 | Toward Physically Grounded JEPA World Models for Goal-Conditioned Robotic Planning | Muyuan Liu, Yue Huang, Zheng Liang +1 | cs.RO | 2026-09-03 |
| #17 | GPS-Bench: A Governance Policy Benchmark for Automating Policy Analysis | Linh Le, Melanie Bui, My Chiffon Nguyen +2 | cs.AI | 2026-09-03 |
| #18 | CulturalMenuBench: Probing the Knowledge-Application Gap in Multimodal Culinary Reasoning | Bo Zeng, Linfeng Gao, Peiqin Lin +9 | cs.AI | 2026-09-03 |
| #19 | Lost in Reordering: Structural Sensitivity of Multilingual LLMs under Semantics-Preserving Perturbations | Karthika Nhayakkat, Rajat Verma, Maharaj Brahma +4 | cs.CL | 2026-09-03 |
| #20 | LongCounsel-8: A Benchmark Suite for Longitudinal Depression Tracking from Multi-Session Counseling Dialogues | Jiayi Li, Zhaomin Wu, Bingsheng He | cs.LG | 2026-09-03 |
| #21 | Making Every Tool Call Count: Necessary Tool-Evidence Path Rewards for Agentic Vision-Language Models | Xingming Long, Yu Liu, Zhiwei Yang +7 | cs.AI | 2026-09-03 |
| #22 | When Retrieval Helps: Selective Retrieval for Single-Turn Mental-Health QA | Hyunseo Oh, Chong-Kwon Kim, Yoonhyuk Choi | cs.CL | 2026-09-03 |
| #23 | Decoupled Analysis-Judging: An Automated Creativity Evaluator Using LLMs in Complex Multi-step Creativity Tasks | Xiangyu Wang, Jin Wu, Xiaoyu Li +2 | cs.CL | 2026-09-03 |
| #24 | Mudragen: Geometrically Supervised Generation of Interacting Two-Hand Mudras for Preserving Indian Classical Dance Heritage | Jagadish Kashinath Kamble, Jayanta Mukhopadhyay, Debaditya Roy +1 | cs.CV | 2026-09-03 |
| #25 | Chiaroscuro for Emotions: A Contrastive Emotion Benchmark Grounded in Appraisal Theory | Divyesh Bommana, Mohammad Saim, Tianyu Jiang | cs.CL | 2026-09-03 |
| #26 | FrameBench:A Language Understanding Benchmark Based on Frame Semantics | Chihiro Yano, Ryohei Sasano | cs.CL | 2026-09-03 |
| #27 | Accountable AI with Grounded, Faithful, Consistent, Actionable Rationales: A Case Study in Clinical Trial Matching with VERDICT | Zikai Zhou, Yufei Jin, Yilin Xu +3 | cs.CL | 2026-09-03 |
| #28 | FPCO-Dialog: A Multi-Turn False-Premise Benchmark for Correction and Cooperation in Vision-Language Models | Jiayuan Ma, Yuqi Lu, Weiyang Guo +5 | cs.CL | 2026-09-03 |
| #29 | PACE: Towards Surfacing Hidden Conflicts in User Requests | Yoojin Kim, Jihyoung Jang, Hyounghun Kim | cs.CL | 2026-09-03 |
| #30 | FlowBalance: Verifier-Grounded Self-Improvement from On-Policy Reasoning Experience | Zixun Huang, Kishan Panaganti, Haitao Mi +1 | cs.LG | 2026-09-03 |
| #31 | SWIM: Student Writing Simulation via Proficiency-Conditioned Generation | Heejin Do, Jakub Kontak, Mrinmaya Sachan | cs.CL | 2026-09-02 |
| #32 | VoxReason: Listener-Free Evaluation of Source-Grounded Speech Planning Before Synthesis | Mengzhe Geng | cs.SD | 2026-09-02 |
| #33 | WireSeg-32K: A Physics-Grounded Synthetic Dataset for Wire Instance Segmentation | Zilin Dai, Lehong Wang, Yi Yang +1 | cs.CV | 2026-09-02 |
| #34 | Verify Before You Distill: Prompt-Level Teacher Gating for On-Policy Distillation | Zhiwei Zhang, Zechen Sun, Fei Zhao +6 | cs.LG | 2026-09-02 |
| #35 | Thinking in Pictures: A Systematic Benchmark for Reasoning-driven Image Generation | Yutong Liu, Nan Huang, Xu Cao +1 | cs.CV | 2026-09-02 |
| #36 | Towards Trustworthy Autonomous Robots: An Explainable AI-Based Decision Framework | Cagri Temel | cs.RO | 2026-09-02 |
| #37 | Large Language Models (LLMs) for Telecom Root Cause Analysis (RCA): A Structured Reasoning Framework for Evidence-Grounded Diagnosis | Hao Zhou, Mandar Kulkarni, Hao Chen +3 | cs.AI | 2026-09-02 |
| #38 | DiscoSign: Discourse-Aware Text to Sign Language Gloss Translation | Vasileios Baltatzis, Mert Inan, Connor Gillis +4 | cs.CL | 2026-09-02 |
| #39 | Measurement-Driven Sub-Network Selection for On-Premise Retrieval-Augmented Factory Agents | Vasileios Rizeakos, Georgios Paisios, Alexandros Machairas +2 | cs.AI | 2026-09-02 |
| #40 | Bilevel Coordinated Reflection: A Game-Theoretic Approach to Multi-Agent LLM Systems | Yihang Chen, Yuxiang Chen, Yuxuan Huang +3 | cs.AI | 2026-09-02 |
| #41 | Modern Transformers Are Implicit Hybrids: From Functional Differentiation to Principled Hybrid Architecture Design | Runlin Shi, Bojian Yin, Guoqi Li | cs.LG | 2026-09-02 |
| #42 | Query Rewriting for Complex Object Segmentation in 4D Gaussian Representations | Thanh-Khoi Nguyen, Thien-Phuc Tran, Minh-Triet Tran | cs.CV | 2026-09-02 |
| #43 | WinoQueer-NL: Assessing Bias in Dutch Language Models toward LGBTQ+ Identities | Jiska Beuk, Gerasimos Spanakis | cs.CL | 2026-09-02 |
| #44 | Orthogonal Ensembles and Tested Explanations for Performer-Independent Body-Motion Emotion Recognition | Naoto Nishida, Yoshio Ishiguro | cs.CV | 2026-09-02 |
| #45 | Addressing Trust in AI Systems through Education: A Didactic Perspective | Pierre Haritz, Hendrik Krone, Thomas Liebig | cs.CY | 2026-09-02 |
| #46 | Scalable Kronecker-Fisher Approximation: Efficient Hessian Analysis for Billion-Parameter Language Models Compression | Viacheslav Yusupov, Daria Cherniuk, Evgeny Frolov | cs.LG | 2026-09-02 |
| #47 | Towards One-for-All Robustness Across a Continuum of Threat Levels | Zhichao Hou, Xiaorui Liu | cs.LG | 2026-09-02 |
| #48 | Before the Script, Set the Stage: How Worldview Simulation Amplifies Psychologically Grounded Persuasion in Multi-Turn Jailbreaking | Siyu Chen, Haoran Wang, Xiaojian Li +3 | cs.CL | 2026-09-02 |
| #49 | PaperCompiler: Faithful Paper-to-Code Generation via Repository-Level Specification Compilation | Yunhao Liu, Hong Phuc Pham, Jaehong Yoon | cs.CL | 2026-09-02 |
| #50 | PhoenixNest-Video: Evidence-Grounded Multimodal Agent Framework for Automated Video Interview Assessment | Fan Yuxuan, Huang Miaojun, Zhang Haimei +2 | cs.AI | 2026-09-02 |