| 1 | Temporal Self-Distillation: Learning Visual State Tracking in Videos Without Supervision | Shravan Venkatraman, Wenshuai Zhao, Mohammad Hassan Vali +1 | cs.CV | 2026-09-03 |
| 2 | Knowledge Acquisition During Pre-training? Large Language Models Learn Better With Auxiliary Views | Joseph Lee, Yidi Huang, Dokyoon Kim +2 | cs.CL | 2026-09-03 |
| 3 | Zero-Shot Novel Depth Synthesis Using 3D Foundation Models Scene Representations | Denis M. Akola, David F. Fouhey | cs.CV | 2026-09-03 |
| #4 | A Case Study on Emergent Cheating and Whistleblowing in Autonomous Research Swarms | Davide Paglieri, Logan Cross, Tim Genewein +3 | cs.AI | 2026-09-03 |
| #5 | Para-Pipe: Exploiting Hierarchical Operator Parallelism of ML Computational Graphs on SoCs | Yujie Zhang, Huiying Lan, Ehsan Aghapour +5 | cs.DC | 2026-09-03 |
| #6 | SENTINEL-RL: Offloading Topological Reasoning from LLM Agents in the Security Operations Center | Uday Vallabhaneni, Cassie L. Cagwin, David J. Wild | cs.CR | 2026-09-03 |
| #7 | Persistent Identity Preservation in Generative Image Models: A Benchmark and Evaluation System | Mengwei Ren, Xuaner Zhang, Zhihao Xia | cs.CV | 2026-09-03 |
| #8 | InSituMeasure: Probing Situated Measurement Grounding in Industrial Scenes with Multimodal Large Language Models | Chao Shen, Xinyuan Li, Yunfan Zhou +4 | cs.AI | 2026-09-03 |
| #9 | The Blind Spot in 2D Infants' Pose Estimation:Robust Learning from Noisy Annotations | Emanuele Cardinale, Marco Proietti, Alessandro Cacciatore +3 | cs.CV | 2026-09-03 |
| #10 | IchthyoNoma: Nomenclature and Context Sensitivity of Zero-Shot Biological Vision--Language Models for Bangladeshi Freshwater Fish Recognition | Nazim-E-Alam, Tarek Rahman, Md Kishor Morol | cs.CV | 2026-09-03 |
| #11 | Investigating the Ability of Large Language Models to Analyze Recipes for Diabetes | Revathy Venkataramanan, Aditya Luthra, Venkatesan Nadimuthu +1 | cs.CL | 2026-09-03 |
| #12 | FiMI Banking: A Sovereign Model for Indian Retail Banking | NPCI AI Research Team, Aman Kumar, Asit Desai +15 | cs.AI | 2026-09-03 |
| #13 | Beyond Endpoint Scores: Time- and Capacity-Conditioned Evaluation of Continual Knowledge Updating | Heejin Choi | cs.LG | 2026-09-03 |
| #14 | Semantic Bayesian World Models | Tommaso Soru | cs.AI | 2026-09-03 |
| #15 | A Reverse Sign Language Dictionary: Open-Vocabulary Sign Recognition from Continuous Signing via Video Captioning and Description Retrieval | Santiago Poveda-Gutiérrez, Hideki Nakayama, Mayumi Bono | cs.CV | 2026-09-03 |
| #16 | SimSkill: A Lifelong Learning AI Agent for Autonomous Mastery of Traffic Simulation | Qi Liu, Qinzheng Wang, Yiming Bie | cs.AI | 2026-09-03 |
| #17 | KnowVis: Knowledge-Centric Visual Summarization for Video Lectures | Yi Xu, Yifan Hou, Xiaoyu Zhang | cs.CV | 2026-09-03 |
| #18 | Genetic Algorithms for Tractable Bayesian Network Fusion via Pre-Fusion Edge Pruning | Pablo Torrijos, José A. Gámez, José M. Puerta +1 | cs.NE | 2026-09-03 |
| #19 | Can LLMs Extract Architectural Design Decisions from Source Code Commits? - A Preliminary Exploratory Study | Amey Karan, Rudra Dhar, Mohamed Soliman +1 | cs.SE | 2026-09-03 |
| #20 | What Do CAE Simulation Agents Really Need Beyond a Generic Harness? | Jiasheng Shi, Tianhan Zhang | cs.CE | 2026-09-03 |
| #21 | Symmetries and Causality: Causal Effect Identification Beyond IID Data | Martin Rabel, Jakob Runge | math.ST | 2026-09-03 |
| #22 | Semantic-Aware Subgraph State Space Model for WSI Classification in Histopathology | Feixing Chen, Hao Lu, Lin Luo +1 | cs.CV | 2026-09-03 |
| #23 | Extracting Forgotten Prompts from Targeted Unlearned Models | Au Ashley Hoi-Ting, Meghdad Kurmanji, William F. Shen +2 | cs.LG | 2026-09-03 |
| #24 | A computable representation of the physical laboratory enables verifiable workflows | Xiaobo Li, Luyao Ge, Xiaohui Li +8 | cs.AI | 2026-09-03 |
| #25 | KC-Bench: A Dynamic Interactive Benchmark for Evaluating Knowledge Conflicts in LLM Agents | Yaxing Lyu, Shengjie Zhou, Binbin Toh +2 | cs.AI | 2026-09-03 |
| #26 | The Attention Triangle in Audio-Video Models | Sagi Polaczek, Noa Kraicer, Gal Metzer +4 | cs.AI | 2026-09-03 |
| #27 | NeoRed: A Knowledge-Logic-Alignment Multimodal Large Language Model for Neonatal Respiratory Disease Diagnosis | Yinan Liu, Hongtai Xia, Haoran Xu +3 | cs.AI | 2026-09-03 |
| #28 | CulturalMenuBench: Probing the Knowledge-Application Gap in Multimodal Culinary Reasoning | Bo Zeng, Linfeng Gao, Peiqin Lin +9 | cs.AI | 2026-09-03 |
| #29 | PPO-STGNN: A Proximal Policy Optimization Approach with Spatio-Temporal Graph Neural Networks for DAG Task Scheduling in Cloud-Edge-End Computing | Yangshuo Qi, Chenwei Wang, Zihan Shen +1 | cs.AI | 2026-09-03 |
| #30 | Building and Evaluating Fixed-Voice Thai TTS from Synthetic Speech | Kunat Pipatanakul, Potsawee Manakul, Warit Sirichotedumrong +3 | cs.CL | 2026-09-03 |
| #31 | Making Every Tool Call Count: Necessary Tool-Evidence Path Rewards for Agentic Vision-Language Models | Xingming Long, Yu Liu, Zhiwei Yang +7 | cs.AI | 2026-09-03 |
| #32 | Pattern Over-Generalization of Knowledge Graph Embedding | Junsik Kim, Kangil Kim | cs.CL | 2026-09-03 |
| #33 | AutoGraphForge: Towards Automated Graph Theory Discovery | Ján Pastorek | cs.AI | 2026-09-03 |
| #34 | Preprocessing Failure and Adversarial Detection in Depthwise-Separable Edge Vision Systems | Jannatul Masruk Mukta, Rifa Sanjida, Adrita Rahman Tory +2 | cs.CV | 2026-09-03 |
| #35 | Preserving Knowledge across Space and Time for Continual Video Deepfake Detection | Taehoon Kim, Jongwook Choi, Heejae Jo +2 | cs.CV | 2026-09-03 |
| #36 | Guide, Not Bind: Why Defeasible Priors Fail in Augmented Lagrangian Causal Discovery | Sairam Sundararaman, Sara Girdhar, Manit Narasimha Murthy +2 | cs.LG | 2026-09-03 |
| #37 | The Civilization Framework: Sovereign-Anchored Communication Between Personal Multi-Agent Systems | Guangjun Liu | cs.MA | 2026-09-03 |
| #38 | Exploring the Potential of Contrastive Language-Image Pre-training for Multi-Source Remote Sensing Data | Xiangyang Miao, Kelu Yao, Yekai Huang +7 | cs.CV | 2026-09-03 |
| #39 | When Depth Hurts: Reliability-Aware Geometry Distillation for Depth-Free RGB-D Salient Object Detection | Xuehao Wang, Jiaxin Hua, Runmei Li +4 | cs.CV | 2026-09-03 |
| #40 | FrameBench:A Language Understanding Benchmark Based on Frame Semantics | Chihiro Yano, Ryohei Sasano | cs.CL | 2026-09-03 |
| #41 | Grassmann--Plücker Parametrization of Convolutional Filter Subspaces: Regularity and Closed Embeddings | Hongyu Yuan, Huaiqing Zuo | math.AG | 2026-09-03 |
| #42 | ALRA: Adaptive Local Relational Alignment for Logit-Based Pre-training Distillation of Autoregressive Language Models | Quang Hoang Trung, Quang Huu Hieu, Nguyen Van Hoang Phuc +1 | stat.ML | 2026-09-03 |
| #43 | From Zero to Hero: An Open LLM Ecosystem for Armenian | Erik Arakelyan, Khatun Avetisyan, Meri Davtyan +5 | cs.LG | 2026-09-03 |
| #44 | PACE: Towards Surfacing Hidden Conflicts in User Requests | Yoojin Kim, Jihyoung Jang, Hyounghun Kim | cs.CL | 2026-09-03 |
| #45 | Language-encoded network topology enables large language models to reason about complex networks | Ucchwas Talukder Utsha, Sakib Mostafa, James Zou +1 | cs.LG | 2026-09-03 |
| #46 | ProgResViT: Progressive Resolution and Width for Adaptive Vision Transformers | Ali Hojjat, Janek Haberer, Olaf Landsiedel | cs.CV | 2026-09-02 |
| #47 | LLMs Learn Better In-Context from Rules than from Examples | Xiang Fu, Seungmin Cho, Yukyung Lee +1 | cs.CL | 2026-09-02 |
| #48 | MemoryLACE: Memory Lifecycle-Aware Consolidation and Evidence Retrieval | Meriem Yacoubi, Pia Schmidt, Nenad Petrovic +3 | cs.CL | 2026-09-02 |
| #49 | Portable Causal Fairness Across Synthetic Data Generator Families | Steven Golob, Sikha Pentyala, Martine De Cock | cs.LG | 2026-09-02 |
| #50 | No country for old linguists: LLM-brain alignment underdetermines neural computation | Elliot Murphy | cs.CL | 2026-09-02 |