| 1 | Workspace Models: Lightweight Robotic Memory via Saliency-Driven Supervision | Nitish Dashora, Douglas Chen, Idan Shenfeld +3 | cs.RO | 2026-09-17 |
| 2 | GeoAAC: Geometry-Based Adaptive Action Chunking from Denoising Trajectories in VLA Policies | Xin Chen, Sen Chen, Yujuan Ding +5 | cs.RO | 2026-09-17 |
| 3 | Agile-WAM: An Agile Tactile World Action Model for Contact-Rich Robot Control | Hanchu Zhou, Brendan Lynch, Raman Goyal +4 | cs.RO | 2026-09-17 |
| #4 | OPTED: On-Policy Fine-Tuning for End-to-End Driving using a Render-Free Teacher | Damiano Da Col, Maximilian Igl, Peter Karkus +5 | cs.RO | 2026-09-17 |
| #5 | MILER: Semantic Mid-Level Representation for Sim-to-Real Reinforcement Learning in Unstructured Autonomous Driving | Thomas Steinecker, Denis Trescher, Alexander Bienemann +2 | cs.RO | 2026-09-17 |
| #6 | Learning Foresight without Explicit Trajectories for 3D Diffusion Policies | Zhongbo Zhang, Zaibin Zhang, Yifan Wang +3 | cs.RO | 2026-09-17 |
| #7 | UniPolicy: Unified Objective-Specific Policies for Generative Search Advertising | Kun Yao, Yuhang Zhou, Yichi Zhang +6 | cs.CL | 2026-09-17 |
| #8 | INSPECT: Learning Robot View Selection from Assistant Use | Di Wen, Kailun Yang, Wenhao Guo +7 | cs.RO | 2026-09-17 |
| #9 | Accelerating Visual Policy Learning with Sampling-Based Model Predictive Control | Yilang Liu, Haoxiang You, Qian Wang +2 | cs.RO | 2026-09-17 |
| #10 | Mitigating Retaliatory Algorithmic Collusion in Repeated Games | Karthik Sivachandran, Rohan Paleja | cs.LG | 2026-09-17 |
| #11 | Model-based Bootstrap for Offline Policy Evaluation in Tabular Reinforcement Learning | Weiwei Wang, Yuqiang Li, Xianyi Wu +1 | stat.ML | 2026-09-17 |
| #12 | MaskHarness-WAM: Instance-Grounded Harnessing for Long-Horizon Robot Manipulation | Zitai Huang, Taiyi Su, Jian Zhu +6 | cs.RO | 2026-09-17 |
| #13 | Learning and Transferring Closed-Loop Robot Software | So Kuroki, Yujin Tang | cs.RO | 2026-09-17 |
| #14 | Improving Cross-embodiment Transfer in Latent Action Models with Action-Similarity Supervision | Maxime Alvarez, Renzo Caballero, Tatsuya Matsushima +2 | cs.RO | 2026-09-17 |
| #15 | UniExo: Unified Multi-Skill Policies for Musculoskeletal Locomotion and Co-Adaptive Exoskeleton Control | Yifei Yuan, Jakob Wolf, Ghaith Androwis +1 | cs.RO | 2026-09-17 |
| #16 | Beyond Patch Removal: Persistent Adversarial Effects in Vision-Language-Action Policies | Enhao Wu, Fusen Guo, Yuxin Cao +3 | cs.CV | 2026-09-17 |
| #17 | A Policy Profile for Croissant: Refusal as a Property of the Dataset | Alexander Chernov | cs.LG | 2026-09-17 |
| #18 | Reach or Solve? Attributing Agentic RL Gains with Checkpoint Handoffs | Xuan Liu, Jingbin Qian | cs.AI | 2026-09-17 |
| #19 | Agentic AI Networking for Heterogeneous Unmanned Aerial Systems in Low-Altitude Wireless Networks | Nguyen Duc Minh Quang, Chang Liu, Shuangyang Li +1 | cs.AI | 2026-09-17 |
| #20 | GLAMDRING: Gait Learning And Morphology co-Design via Reinforcement LearnING of CPGs | Amogh Joshi, Kaushik Roy | cs.RO | 2026-09-16 |
| #21 | Predict Before You Deploy: Offline Prediction of Quantization-Induced Task Degradation for World Action Models | Jiuyi Xu, Jinjia Guo, Meida Chen +2 | cs.RO | 2026-09-16 |
| #22 | Stable Policy Learning | Harvey Barnhard, Giacomo Opocher, Rahul Singh | econ.EM | 2026-09-16 |
| #23 | Improving Offline Goal-Conditioned Reinforcement Learning via Selective Reward Stimulation | Jing Zhang | cs.LG | 2026-09-16 |
| #24 | From Rollout to Reset: A Graph-Based Harness for Autonomous Long-Horizon Manipulation Evaluation | Jing Jiang, Yue Yang, Xinkai Jiang +3 | cs.RO | 2026-09-16 |
| #25 | Efficient Nash Equilibrium Computation for Cybersecurity Games | Michael Lanier, David Farmer, Yevgeniy Vorobeychik | cs.GT | 2026-09-16 |
| #26 | In-Context Robot Learning with VLM Agents | Dongzhou Cheng, Taoran Yi, Ye Fang +12 | cs.CV | 2026-09-16 |
| #27 | Dreaming the Sound of Contact: Leveraging Video and Audio Generation for Zero-Shot Force-Aware Manipulation and Data Generation | Guanhua Ji, Tianyu Li, Dayoon Suh +3 | cs.RO | 2026-09-16 |
| #28 | rMuscle: Robotic Muscle Memory for Efficient Vision-Language-Action Model Inference | Kaijun Zhou, Zhiyang Li, Le Chen +1 | cs.RO | 2026-09-16 |
| #29 | Compositional Policy Violations: When Step-Level Compliance Fails In Agentic AI Workflows | Ashwini Kurady, Sri Sai Charith Grandhi, Rajesh Gupta +1 | cs.AI | 2026-09-16 |
| #30 | A Convergence Framework for Deep $V$-Learning: Error Propagation and Sharp Action-Gap Bounds | Yury Kolomeytsev | cs.LG | 2026-09-16 |
| #31 | CERA-MoA: Co-Evolving Routing Mechanisms with Continually Learning LLM Agents | Jiaxuan Jiang, Liyuan He, Zhixuan Fang | cs.AI | 2026-09-16 |
| #32 | FRAUDSkill: Structured Frozen-Weight Skill Optimization for Audio Anti-Fraud Detection | Chengxian Hu, Zhiming Ma, Mingjun Pan +9 | cs.SD | 2026-09-16 |
| #33 | The Uneven Impact of Generative AI on Student Learning: Examining the Roles of Reliance, Evaluation Literacy, and Course Policy in AI-related Courses | Lydia Manikonda, Mei Si, Sirajam Munira +2 | cs.AI | 2026-09-16 |
| #34 | VLA-ULAP: Interleaving Cloud VLA Calls with Ultra-Lightweight Local Action Prediction at the Edge | Deyu Cao, Ryuji Oi, Kosuke Matsushima +4 | cs.RO | 2026-09-16 |
| #35 | A Geometric Theory of Decision Boundaries in Structured Markov Decision Processes | Fredy Pokou | cs.LG | 2026-09-16 |
| #36 | WetRobo: A Reproducible Robot Kit for Coding Agents in Biological Laboratories | Yuna Oikawa, Kei Endo, Takanori Uzawa +5 | cs.AI | 2026-09-16 |
| #37 | Position: It is Time to Virtualize Foundation Models with a Self-evolving Operating System Layer | Suparna Bhattacharya, Tarun Kumar, Cong Xu +7 | cs.AI | 2026-09-16 |
| #38 | Autonomy in Check: Governor-Mediated Adaptive Security at the Edge | Ijaz Ahmad, Ijaz Ahmad, Flavio Esposito +1 | cs.CR | 2026-09-16 |
| #39 | Visual Compliance via Executable Safety Rule Entailment | Jisoo Kim, TaeYoon Kwack, Jinwoo Jang +1 | cs.AI | 2026-09-16 |
| #40 | APGEM: Adaptive Policy-Guided Error Mitigation for Quantum Reinforcement Learning on a Real-World CVRP Case Study | Shabir Ahmad Sofi, Bisma Majid, Mir Mohammad Yousuf | cs.LG | 2026-09-16 |
| #41 | Reinforcement Learning for Real-Time Vision-Language-Action Policies | Perry Dong, Kuo-Han Hung, Dorsa Sadigh +1 | cs.RO | 2026-09-16 |
| #42 | A Comprehensive Review of Generative Physical Artificial Intelligence | Satyam Gaba, Krutiksinh Rana, Siva Sai +2 | cs.RO | 2026-09-16 |
| #43 | Contiguity, Not Importance: Budgeted Repair of Stale KV Caches After Document Edits | Mingyang Mao, Wyatt Mackey, Xiaomin Lin | cs.AI | 2026-09-16 |
| #44 | When to Call an LLM: A Confidence-Gated Hybrid for Cost-Effective Emotion Recognition in Conversational AI | Sai Babu Udayagiri, Arjun Chouhan, Ravisekhar Kanagala +1 | cs.AI | 2026-09-16 |
| #45 | Do Frontier Models Seek Safety Evidence Before Acting? | Omer Tafveez | cs.AI | 2026-09-15 |
| #46 | Adaptive hybrid coupling with operator inference, the overlapping Schwarz alternating method and reinforcement learning | Trishit Mondal, Irina Tezaur, Anthony Gruber | cs.LG | 2026-09-15 |
| #47 | Learning Multi-Humanoid Pickup and Transport via Decentralized Object-Centric Control | Bikram Pandit, Mohitvishnu S. Gadde, Aayam Kumar Shrestha +1 | cs.RO | 2026-09-15 |
| #48 | FairCompressAgent: An Agentic Framework for Fairness-Aware Model Compression for FPGA Deployment | Yuanbo Guo, Yiyu Shi | cs.AI | 2026-09-15 |
| #49 | CALOS: Control-Affine Lyapunov On-manifold Safety Layer for Safe Deep Reinforcement Learning for Quadrotors | Fabrizio Cesareo, Sebastiano Mengozzi, Nicola Mimmo +1 | cs.RO | 2026-09-15 |
| #50 | REVERSAL-BENCH: A Reversibility Axis and Reset Oracle for Measuring the Reset-Free RL Cliff | Riyaaz Shaik, Chandru Venkataraman | cs.LG | 2026-09-15 |