| 1 | An Empirical Study of Harness Design for Coding Agents | Run-Ze Fan, Zihao Zhang, Simin Ma +6 | cs.AI | 2026-09-17 |
| 2 | Large Language Models as Falsifiers for Cyber-Physical Systems | Ali ArjomandBigdeli, Jiawei Zhou, Stanley Bak | eess.SY | 2026-09-17 |
| 3 | RISC-V and machine learning: a survey | Shriman Keshri, Apparna Singh, Chinmaya Kumar Palo +2 | cs.LG | 2026-09-17 |
| #4 | Multi-center Medical Data Mining with FL-Net - A One-stop Shop for Federated Learning | Simon Süwer, Julian Klemm, Elisa Acitelli +38 | cs.LG | 2026-09-17 |
| #5 | Chronicle: Cut-Point Replay for Regression Testing of LLM Agents | Tisha Chawla, Susheem Koul | cs.CL | 2026-09-17 |
| #6 | SoL-Pi: Recursively Scaling Auto-Research Loops for Efficient Agent Harness | Haozhe Liu, Tian Ye, Sensen Gao +11 | cs.AI | 2026-09-17 |
| #7 | The Bias of Nonlinear Two-Time-scale Stochastic Approximation under Constant Step-Sizes | Djamel Rassem Lamouri, Dorian Baudry, Nicolas Gast | cs.LG | 2026-09-17 |
| #8 | STR-Agent: An LLM-Driven Agent for QoS-Aware Routing in LEO Satellite Networks | Bowen Lu, Mugen Peng, Yaohua Sun +3 | cs.NI | 2026-09-17 |
| #9 | MTVA-Bench: Evaluating the Language Model Inside Cascaded Voice Agents | Pritish Mishra, Ishaan Kumar, Akshat Mandoli +1 | cs.AI | 2026-09-17 |
| #10 | A Scalable Trust Discovery Architecture for the Internet of Agents | Song Zhang, Jiankang Yao, Hongtao Li +4 | cs.CR | 2026-09-17 |
| #11 | UnifiedPlayers: Enhance Tool-Integrated Reasoning in Agentic Reinforcement Learning | Wenjie Liao, Liangjie Zhao, Zehong Cao | cs.AI | 2026-09-17 |
| #12 | MATCH: Model-Aware Tool Learning with Curriculum Scheduling and Hierarchically Gated Rewards | Shihao Liu, Hao Yin, Lijun Liu +3 | cs.LG | 2026-09-17 |
| #13 | A Proposal for an Agentic AI Architecture to Support Multi-Domain Decision-Making in the Brazilian Armed Forces | Gioliano de Oliveira Braga, Sidnei Barbieri, Ágney Lopes Roth Ferraz +2 | cs.AI | 2026-09-17 |
| #14 | Robust Workflow Generation via Adversarial Learning for Audio Deepfake Detection | Xiang Li, Pin-Yu Chen, Wenqi Wei | cs.SD | 2026-09-17 |
| #15 | Not All AI Agents Are Equal: Characterizing Resource and Performance Dynamics | Wonmi Choi, Minuk Park, Zhixiong Niu +3 | cs.AI | 2026-09-17 |
| #16 | CitySTAR: Structured and Topology-Aware Reasoning for Open-Vocabulary Urban 3D Grounding | Shuai Zhang, Hongye Hou, Qinghe Liu +5 | cs.CV | 2026-09-17 |
| #17 | Self-Replicating Neural Cellular Automata: Quantifying Emergent Phenotypic and Genotypic Diversity in an OpenEnded Substrate | Sanyam Jain, Felix Simon Reimers, Stefano Nichele | q-bio.PE | 2026-09-17 |
| #18 | SlugTrails: An Egocentric Benchmark for Floor Plan Localization in Large Buildings | Yunqian Cheng, Roberto Manduchi | cs.CV | 2026-09-17 |
| #19 | DeliveryGym: An RL Environment for Long-Horizon Embodied Agent Planning with Adaptive Curriculum | Haoqiang Kang, Yiming Zhang, Yiyang Guo +5 | cs.LG | 2026-09-17 |
| #20 | Integrating knowledge from case reports: a medical ontology based multimodal information system with structured summary | Shuyu Guo, Lan Huang, Yichen Liu +2 | cs.AI | 2026-09-17 |
| #21 | ALIBI: Adversarial Legitimacy Injection in Binary Input against LLM Malware Analyzers | Hyeongjun Choi, Wonyoung Jung, Haehoon Seo +1 | cs.CR | 2026-09-17 |
| #22 | SeetaPsych v1.0: An Open-source Computer Vision Toolkit for Behavior-based Psychological Measurement | Jiabei Zeng, Chiqin Li, Kaizhou Li +7 | cs.CV | 2026-09-17 |
| #23 | VideoResearcher: Self-Improving Tool Design for Long-Video Understanding | Dingqiang Ye, Dongdi Zhao, Kaishen Wang +11 | cs.CV | 2026-09-17 |
| #24 | ScientistTwo: Pioneering the Human Knowledge Frontier with Autonomous AI | Jaehyun Nam, Jinsung Yoon, Yanzhou Pan +4 | cs.AI | 2026-09-17 |
| #25 | From Intent to Action: Benchmarking LLM Safety in Vehicle Voice Command Authorization | Diba Afroze, Xingli Zhang, Yazhou Tu +1 | cs.AI | 2026-09-17 |
| #26 | Red-Teaming Auto Mode: Improving Blocking Classifiers Against Malign Coding Agents | Alex Remedios, Simon Storf, Fabien Roger +1 | cs.CR | 2026-09-17 |
| #27 | Compositional Reasoning in Language Models under Reinforcement Learning Post-Training | Yu He, Yingxi Li, Yifei Wang +1 | cs.AI | 2026-09-16 |
| #28 | Sharpness-Aware Minimization (SAM) Improves Classification Accuracy of Bacterial Raman Spectral Data Enabling Portable Diagnostics | Kaitlin Zareno, Jarett Dewbury, Siamak K. Sorooshyari +2 | cs.LG | 2026-09-16 |
| #29 | Closed-World Resolution Against Tool Hallucination in LLM Agents | Laxmipriya Ganesh Iyer | cs.AI | 2026-09-16 |
| #30 | Can Vision-Language Models Judge Olympic Diving? From Reasoning to Scores in Zero-Shot Action Quality Assessment | Henry O. Velesaca, David Freire-Obregon, Luigi Miranda +1 | cs.CV | 2026-09-16 |
| #31 | Kinematics-Grounded Agentic AI for Robotic Additive Manufacturing Process Planning | Jingzhan Ge, Ruimin Chen, Azadeh Haghighi +2 | cs.RO | 2026-09-16 |
| #32 | A frontend-backend architecture for tool calls in full-duplex speech models | Ke Hu, Slyne Deng, Chen Chen +10 | cs.CL | 2026-09-16 |
| #33 | In-Context Robot Learning with VLM Agents | Dongzhou Cheng, Taoran Yi, Ye Fang +12 | cs.CV | 2026-09-16 |
| #34 | Characterizing Web Search by Conversational LLM Agents: From Search Decisions and Strategies to Results and Responses | Mahsa Amani, Seungeon Lee, Abhisek Dash +9 | cs.AI | 2026-09-16 |
| #35 | ScienceIDE: Turning World's Scientific Codebase into Agent Learnable Environments | Hejia Geng, Zesen Huang, Haoyang Li +42 | cs.CL | 2026-09-16 |
| #36 | Randomized SVD Approximations for Spectral Co-Clustering of Word-Document Matrices | Fateme Mazdarani, Carlos Toxtli | cs.LG | 2026-09-16 |
| #37 | Evidence-Grounded Agentic Formulation Development in an Autonomous Laboratory | Michael M. Craig, Riley J. Hickman, Yingshan Ma +3 | cs.LG | 2026-09-16 |
| #38 | Prepared Or Unprepared? Evaluating Healthcare Workforce Readiness for Clinical Adoption of Artificial Intelligence in Nigeria | Abbas M. Rabiu, Abdulrazaq A. Zubair, Um-mulkhairi Ibrahim +6 | cs.CY | 2026-09-16 |
| #39 | ASLEval: Measuring Privacy Exposure Displacement in LLM Agent Sessions | Guosen Wu, Huizhen Huang, Guoxiong Long +2 | cs.CR | 2026-09-16 |
| #40 | The AR Fairness Metamodel: A Structured Framework for Fairness Measures | Julian Alfredo Mendez, Timotheus Kampik | cs.CY | 2026-09-16 |
| #41 | Ask the Tool, Don't Guess: Agent Tool Calls Hold Their Progress, and the Serving System Should Read It | Yipeng Liu, Yingqiang Zhang, Feifei Li +1 | cs.DC | 2026-09-16 |
| #42 | ReFigBench: Benchmarking Scientific Figure Reconstruction as Editable PowerPoint Artifacts | Liyang Fan, Chi Wei, Yitai Li +8 | cs.CL | 2026-09-16 |
| #43 | WaveTLM: Reliable Time-Series Language Modeling through Task Compilation | Jiahui Chen, Bingke Zhu, Hongyu Pan +1 | cs.LG | 2026-09-16 |
| #44 | DISTA-Net++: Rethinking Infrared Small Target Unmixing Beyond Sub-Pixel Separation | Mengze Xu, Zhu Liu, Weidong Sheng +4 | cs.CV | 2026-09-16 |
| #45 | PAPC: Platform Mediation for Privacy-Propagation Externalities in AI-Mediated Workflows | Tao Huang, Guosen Wu, Chen Hou +1 | cs.CR | 2026-09-16 |
| #46 | Clueing up LLMs with Tool-Augmented Deductive Reasoning | Rebecca Ansell, Autumn Toney-Wails | cs.AI | 2026-09-16 |
| #47 | The Uneven Impact of Generative AI on Student Learning: Examining the Roles of Reliance, Evaluation Literacy, and Course Policy in AI-related Courses | Lydia Manikonda, Mei Si, Sirajam Munira +2 | cs.AI | 2026-09-16 |
| #48 | Selection Is Retrieval, Abstention Is Not: On-Device Tool Routing over 70 Korean-English Actions | Janghoon Lee | cs.CL | 2026-09-16 |
| #49 | Recursive Reasoning or Statistical Extrapolation? In-Context Learning in Multi-Agent Interdependent Decision-Making | Yu Liu, Wenwen Li, Yifan Dou +1 | cs.AI | 2026-09-16 |
| #50 | Provable Guarantees for Spectral Structured Prediction | Violet Zheng, Jean Honorio | cs.LG | 2026-09-16 |