| 1 | Coding Agents with an Obstacle-Aware Harness for Safe Robot Manipulation | Bingxin Xu, Yuzhang Shang, Zhen Dong +1 | cs.RO | 2026-09-17 |
| 2 | Workspace Models: Lightweight Robotic Memory via Saliency-Driven Supervision | Nitish Dashora, Douglas Chen, Idan Shenfeld +3 | cs.RO | 2026-09-17 |
| 3 | GeoAAC: Geometry-Based Adaptive Action Chunking from Denoising Trajectories in VLA Policies | Xin Chen, Sen Chen, Yujuan Ding +5 | cs.RO | 2026-09-17 |
| #4 | Agile-WAM: An Agile Tactile World Action Model for Contact-Rich Robot Control | Hanchu Zhou, Brendan Lynch, Raman Goyal +4 | cs.RO | 2026-09-17 |
| #5 | Learning Foresight without Explicit Trajectories for 3D Diffusion Policies | Zhongbo Zhang, Zaibin Zhang, Yifan Wang +3 | cs.RO | 2026-09-17 |
| #6 | HIL-UMI: Bringing Human-in-the-Loop Post-Training of Vision-Language-Action Models to Universal Manipulation Interface | Zimu Han, Yiming Zeng, Jiyao Zhang +9 | cs.RO | 2026-09-17 |
| #7 | DexTouch-WM: Learning Action-Conditioned Tactile World Models from Human Touch for Dexterous Robot Manipulation | Yan Qin, Yue Chen, Wenwei Lin +8 | cs.RO | 2026-09-17 |
| #8 | Accelerating Visual Policy Learning with Sampling-Based Model Predictive Control | Yilang Liu, Haoxiang You, Qian Wang +2 | cs.RO | 2026-09-17 |
| #9 | SenseFuse: Label-Free Fusion of Image and Shape Encoders for Open-Vocabulary 3D Instance Segmentation | Euiseok Han, Tri Ton, Hwanhee Kim +2 | cs.CV | 2026-09-17 |
| #10 | TouchSight: Bare-Handed Tactile Prediction from Egocentric Video via Generative Visual Augmentation | Danyan Zhou, Jinxuan Lu, Jiawei Lin +3 | cs.CV | 2026-09-17 |
| #11 | AnyviewMeter: Adapting Robotic Reward Models with Camera Geometry and Multi-View Attention | Yuang Tu, Runjia Tan, Yujie Yan +2 | cs.CV | 2026-09-17 |
| #12 | MAGMA-GEN: Validated Recovery Supervision from Ambiguous Failures via Counterfactual Re-Execution | Loan Bernat, Matthieu Grard, Ariane Herbulot +1 | cs.AI | 2026-09-17 |
| #13 | MaskHarness-WAM: Instance-Grounded Harnessing for Long-Horizon Robot Manipulation | Zitai Huang, Taiyi Su, Jian Zhu +6 | cs.RO | 2026-09-17 |
| #14 | Delphi Scanner: efficient and interpretable static malware detection via API sequence modeling | Bijied Brahimi, Vincent Cohadon, Gabriel Glazman +3 | cs.CR | 2026-09-17 |
| #15 | Uni-LaDiR: Latent Diffusion Unifies Multimodal Reasoning | Haoqiang Kang, Yizhe Zhang, Nikki Lijing Kuang +2 | cs.LG | 2026-09-17 |
| #16 | TacSushi: Tactile-Grounded World-Action Modeling for Dexterous Sushi Manipulation | Haodi Hu, Kaen Kogashi, Toshiaki Koike-Akino | cs.RO | 2026-09-17 |
| #17 | ParticleSplat: Self-supervised Object-centric Latent Particle Splatting | Lyuxing He, Daniel Guo, Elizabeth Terveen +3 | cs.CV | 2026-09-16 |
| #18 | From Rollout to Reset: A Graph-Based Harness for Autonomous Long-Horizon Manipulation Evaluation | Jing Jiang, Yue Yang, Xinkai Jiang +3 | cs.RO | 2026-09-16 |
| #19 | PointZero: 3D Point Track Completion for Learning Transferable 3D Dynamics | Bardienus P. Duisterhof, Kaifeng Zhang, Adam Hung +5 | cs.CV | 2026-09-16 |
| #20 | Dreaming the Sound of Contact: Leveraging Video and Audio Generation for Zero-Shot Force-Aware Manipulation and Data Generation | Guanhua Ji, Tianyu Li, Dayoon Suh +3 | cs.RO | 2026-09-16 |
| #21 | Track, Articulate, Act: Generating Articulation from Casual Human Videos | Jiaming Zhang, Homanga Bharadhwaj | cs.CV | 2026-09-16 |
| #22 | rMuscle: Robotic Muscle Memory for Efficient Vision-Language-Action Model Inference | Kaijun Zhou, Zhiyang Li, Le Chen +1 | cs.RO | 2026-09-16 |
| #23 | HAP: A Hand-Driven Active Perception Framework for Egocentric Head Motion Prediction | Yunji Feng, Junyi Ma, Guanzhong Sun +2 | cs.CV | 2026-09-16 |
| #24 | ActiveScale: Scaling Active Perception for Robots across Model, Data, and Hardware | Shuai Zhou, Kaisheng Pang, Wenxuan Song +3 | cs.RO | 2026-09-16 |
| #25 | Market Signal Injection: Adversarial Context Manipulation of LLM Pricing Agents | Dohun Lee, Hyunwoo Park | cs.AI | 2026-09-16 |
| #26 | Acting in Meters: Learning Metric Interactions for Precise Robotic Manipulation | Lijie Wang, Zheng Lu, Yiming Wang +12 | cs.RO | 2026-09-16 |
| #27 | Reinforcement Learning for Real-Time Vision-Language-Action Policies | Perry Dong, Kuo-Han Hung, Dorsa Sadigh +1 | cs.RO | 2026-09-16 |
| #28 | Energy-Regularized Imitation Learning for Force- and Work-Aware Robotic Manipulation | Toshiki Otani, Hiromu Taketsugu, Norimichi Ukita | cs.RO | 2026-09-16 |
| #29 | Symbolic Temporal Supervision of LLM Agents Using Contracts | Yifeng Xiao, Pierluigi Nuzzo | cs.AI | 2026-09-16 |
| #30 | RoboVAD: A Large Cross-Domain Evaluation Benchmark for Anomaly Detection in Robotic Arm Manipulation Videos | Alexandru-Bogdan Dura, Sebastian Balmus, Radu Tudor Ionescu | cs.CV | 2026-09-15 |
| #31 | Learning Multi-Humanoid Pickup and Transport via Decentralized Object-Centric Control | Bikram Pandit, Mohitvishnu S. Gadde, Aayam Kumar Shrestha +1 | cs.RO | 2026-09-15 |
| #32 | REVERSAL-BENCH: A Reversibility Axis and Reset Oracle for Measuring the Reset-Free RL Cliff | Riyaaz Shaik, Chandru Venkataraman | cs.LG | 2026-09-15 |
| #33 | PhysStream: Streaming Physics-Grounded Video Generation with Structured Scene Memory and Fine-Grained Motion Control | Chuhao Chen, Peter Wonka, Chaoyang Wang +4 | cs.CV | 2026-09-15 |
| #34 | GeoLAM: Learning Geometry-Grounded Latent Actions from Unlabeled Human Videos | Yifan Xie, Hekun Tian, Jinkun Liu +3 | cs.CV | 2026-09-15 |
| #35 | TEMPO: Learning Temporal Context for Dynamic Robot Manipulation | Zhenyang Feng, Jimin Heo, Erik B. Sudderth +1 | cs.RO | 2026-09-15 |
| #36 | The Latent That Never Was: A Forensic Re-run of the CVAE Ablation in Action Chunking Transformers | Bo Kang | cs.RO | 2026-09-15 |
| #37 | World Models for Embodied Intelligence: From Plausible to Controllable to Actionable | Nanjie Yao, Hao Wang, Chong Cheng +10 | cs.RO | 2026-09-15 |
| #38 | MEgoVista: Multi-view Ego-aware Motion Estimation for Metric 4D Hands and Head in the Wild | Jiangong Xiao, Zhihao Zhang, Yifei Dong +7 | cs.CV | 2026-09-15 |
| #39 | Weave: Learning Whole-Body Dexterous Loco-Manipulation from Human-Object Interactions | Liu Cao, Xingze Wu, Jingzhi Cui +4 | cs.RO | 2026-09-15 |
| #40 | ProxiDex: Learning Dynamics-Guided Proximity Policy for Dexterous Manipulation | Yushan Bai, Boyu Zheng, Zhiyang Mao +4 | cs.RO | 2026-09-15 |
| #41 | Interpreting and Steering LLM Agents for Social Simulations | Jiayue Gaveal Fan, Arul Murugan, Shreyas Krishnan +1 | cs.LG | 2026-09-14 |
| #42 | Reasoning with Image Generation | Nishad Singhi, Hector Garcia Rodriguez, Aditya Arora +2 | cs.CV | 2026-09-14 |
| #43 | Autonomous Droplet Navigation via Model-Based Reinforcement Learning | Rajneesh Anand, Mayuresh V. Kothare | cs.LG | 2026-09-14 |
| #44 | Occupancy Network-Guided Autonomous Robotic Partial Nephrectomy | Ethan Kilmer, Pit Henrich, Jiawei Ge +13 | cs.RO | 2026-09-14 |
| #45 | Disentangling Representation Evolution in Transformers through Directional Decomposition | Shwai He, Haichao Zhang, Shen Yan | cs.CL | 2026-09-14 |
| #46 | SlipSense: Multimodal Tactile Learning for Low-Latency and Generalized Slip Detection | Tong Jian, Aditya Thurvas Senthil Kumar, Xinyi Li +7 | cs.RO | 2026-09-14 |
| #47 | Bench2Dex: Benchmarking Visuo-Tactile Bimanual Dexterous Manipulation Across Dexterous Hands | Zhenjie Yang, Yideng Zhang, Dongjie Zhang +21 | cs.RO | 2026-09-14 |
| #48 | Universal Defenses for Tool-Integrated LLM Agents Against Adversarial Attacks | Xiaoyan Li, Yunli Wang | cs.CR | 2026-09-14 |
| #49 | Artificial entrepreneurial cognition: Locating and causally steering an opportunity recognition dial inside large language models (LLMs) | Christian Fisch, Angela Altmeier, Martin Obschonka +2 | cs.CL | 2026-09-14 |
| #50 | CWM: Controllable White-Box Meta-Prompting for Adaptive Retrieval-Augmented Generation and Reasoning Ability | Keuntae Kim, Eunhye Jeong, Yong Suk Choi | cs.AI | 2026-09-14 |