| 1 | PointZero: 3D Point Track Completion for Learning Transferable 3D Dynamics | Bardienus P. Duisterhof, Kaifeng Zhang, Adam Hung +5 | cs.CV | 2026-09-16 |
| 2 | In-Context Robot Learning with VLM Agents | Dongzhou Cheng, Taoran Yi, Ye Fang +12 | cs.CV | 2026-09-16 |
| 3 | Dreaming the Sound of Contact: Leveraging Video and Audio Generation for Zero-Shot Force-Aware Manipulation and Data Generation | Guanhua Ji, Tianyu Li, Dayoon Suh +3 | cs.RO | 2026-09-16 |
| #4 | Track, Articulate, Act: Generating Articulation from Casual Human Videos | Jiaming Zhang, Homanga Bharadhwaj | cs.CV | 2026-09-16 |
| #5 | rMuscle: Robotic Muscle Memory for Efficient Vision-Language-Action Model Inference | Kaijun Zhou, Zhiyang Li, Le Chen +1 | cs.RO | 2026-09-16 |
| #6 | PhysVGGT: Feed-Forward Dense Physical Property Estimation from A Single Image | Sneha Paul, Guile Wu, Bingbing Liu +1 | cs.CV | 2026-09-16 |
| #7 | ActiveScale: Scaling Active Perception for Robots across Model, Data, and Hardware | Shuai Zhou, Kaisheng Pang, Wenxuan Song +3 | cs.RO | 2026-09-16 |
| #8 | CSWAM: Better Causal Semantic Representations for Out-of-Distribution Generalization in World Action Models | Tianbin Liu, Jian Zhu, Taiyi Su +4 | cs.CV | 2026-09-16 |
| #9 | WetRobo: A Reproducible Robot Kit for Coding Agents in Biological Laboratories | Yuna Oikawa, Kei Endo, Takanori Uzawa +5 | cs.AI | 2026-09-16 |
| #10 | StrucPhysVideo: Learning Physical Dynamics from Structured Captions and Robot Actions | Awomo-WM Team, :, Enhui Ma +7 | cs.CV | 2026-09-16 |
| #11 | RecMorph: Topology-Guided Spatial Recurrence for Generalized Morphology Control | Quanrui Rao, Yong Liu, Xueming Xiao +4 | cs.RO | 2026-09-16 |
| #12 | ${M}^2$Tok: Multi-head Multi-codebook Discrete Action Tokenization for Vision-Language-Action Models | Chunpu Xu, Zhixuan Liang, Yuhao Zhang +6 | cs.RO | 2026-09-16 |
| #13 | Acting in Meters: Learning Metric Interactions for Precise Robotic Manipulation | Lijie Wang, Zheng Lu, Yiming Wang +12 | cs.RO | 2026-09-16 |
| #14 | Reinforcement Learning for Real-Time Vision-Language-Action Policies | Perry Dong, Kuo-Han Hung, Dorsa Sadigh +1 | cs.RO | 2026-09-16 |
| #15 | Characterizing Replay Retention Under Dynamics Shift in Model-Based Reinforcement Learning | Everest Yang, Skye Thompson, George D. Konidaris | cs.RO | 2026-09-16 |
| #16 | Energy-Regularized Imitation Learning for Force- and Work-Aware Robotic Manipulation | Toshiki Otani, Hiromu Taketsugu, Norimichi Ukita | cs.RO | 2026-09-16 |
| #17 | PRISM: Predictive Representation of Interaction Style and Motion for Social Robot Navigation | Bo-Han Chen, Hiromu Taketsugu, Norimichi Ukita | cs.CV | 2026-09-16 |
| #18 | A Comprehensive Review of Generative Physical Artificial Intelligence | Satyam Gaba, Krutiksinh Rana, Siva Sai +2 | cs.RO | 2026-09-16 |
| #19 | Beyond Pixel Similarity: Task-Aware Evaluation of GAN-Based Synthetic Sonar Data for Robotic Perception | Hannan Ejaz Keen, Muhammad Moazam Fraz, Karsten Berns | cs.RO | 2026-09-16 |
| #20 | Teaching AI, Robotics, & Community: A Hubs-Based K-12 Education Framework for Reaching Rural Schools | Maxwell J. Jacobson, Gustavo Rodriguez-Rivera, Petros Drineas +1 | cs.AI | 2026-09-16 |
| #21 | Missing Bridges: Composition-Aware Active Imitation Learning | Maxwell J. Jacobson, Ahmed H Qureshi, Yexiang Xue | cs.AI | 2026-09-16 |
| #22 | RoboVAD: A Large Cross-Domain Evaluation Benchmark for Anomaly Detection in Robotic Arm Manipulation Videos | Alexandru-Bogdan Dura, Sebastian Balmus, Radu Tudor Ionescu | cs.CV | 2026-09-15 |
| #23 | Learning Multi-Humanoid Pickup and Transport via Decentralized Object-Centric Control | Bikram Pandit, Mohitvishnu S. Gadde, Aayam Kumar Shrestha +1 | cs.RO | 2026-09-15 |
| #24 | CALIPER: Metric-Grounded Model-Free Recognition of Visually Similar Industrial Parts | Alankrit Gupta, Chenxi Tao, Seung-Kyum Choi | cs.CV | 2026-09-15 |
| #25 | HINT-Plan: Human Intention-Aware Robot Task Planning in Context-Rich Environments using Vision Language Models | Yuchen Liu, Luigi Palmieri, Lujun Li +3 | cs.RO | 2026-09-15 |
| #26 | SlotDiT: Object-Centric Representations for Diffusion Transformers | Gjergj Plepi, Sven Behnke | cs.CV | 2026-09-15 |
| #27 | PanoGS-SLAM: Panoramic 3D Gaussian Splatting SLAM | Yongqi Mao, Hao Shi, Yufan Zhang +3 | cs.CV | 2026-09-15 |
| #28 | FluxVLA Engine: A One-Stop VLA Engineering Platform for Embodied Intelligence | Yinhao Li, Weixin Mao, Zihan Lan +21 | cs.RO | 2026-09-15 |
| #29 | MyoFlow: Anchor-Tied Rectified Flow for HD-sEMG Gesture Recognition Across Sessions and Subjects | Chenhao Wu, Dingjie Peng, Satoshi Funabashi +5 | cs.LG | 2026-09-15 |
| #30 | EventEgoHands++: Event-based Egocentric 3D Hand Mesh Reconstruction with Real Dataset | Ryosei Hara, Wataru Ikeda, Masashi Hatano +1 | cs.CV | 2026-09-15 |
| #31 | Continual Learning for Traversability Prediction with Uncertainty-Aware Adaptation | Hojin Lee, Yunho Lee, Daniel A Duecker +1 | cs.RO | 2026-09-15 |
| #32 | Intrinsic Robot Rewarding: Reusing VLA Representations for Autonomous Evaluation and Policy Improvement | Tobias Schaffer, Mohab Elkhayat, Daniela Nicklas +2 | cs.RO | 2026-09-15 |
| #33 | GeoLAM: Learning Geometry-Grounded Latent Actions from Unlabeled Human Videos | Yifan Xie, Hekun Tian, Jinkun Liu +3 | cs.CV | 2026-09-15 |
| #34 | Bridging Learned Visual Perception and Symbolic Belief-Space Planning | Guy Azran, Michael Navat, Sarah Keren | cs.AI | 2026-09-15 |
| #35 | TEMPO: Learning Temporal Context for Dynamic Robot Manipulation | Zhenyang Feng, Jimin Heo, Erik B. Sudderth +1 | cs.RO | 2026-09-15 |
| #36 | CoAdapt: An LLM-based Framework for Adaptive Collaborative Perception in IIoT Robotic Swarms | Houssam Hajj Hassan, Antonia Maria Masucci, Lynda Zitoune +1 | cs.AI | 2026-09-15 |
| #37 | The Latent That Never Was: A Forensic Re-run of the CVAE Ablation in Action Chunking Transformers | Bo Kang | cs.RO | 2026-09-15 |
| #38 | Visual Cue Guided Video Planning for Generalizable Robot Navigation | Hojin Lee, Sizhe Lester Li, Maximilian Hilger +4 | cs.RO | 2026-09-15 |
| #39 | Differentiable Mesh State Estimation via Factor Graph Inference for Deformable Object Reconstruction | Lidia Al-Zogbi, Fangjie Li, Samuel Tobin +13 | cs.RO | 2026-09-15 |
| #40 | Weave: Learning Whole-Body Dexterous Loco-Manipulation from Human-Object Interactions | Liu Cao, Xingze Wu, Jingzhi Cui +4 | cs.RO | 2026-09-15 |
| #41 | Can Knowledge Transfer Parameters Be Learned? LePoKet for Efficient Robotic Vision | Yanick C. Tchenko, Felix Mohr, Hicham Hadj-Abdelkader +1 | cs.CV | 2026-09-15 |
| #42 | The Neverwhere Visual Parkour Benchmark Suite | Ziyu Chen, Henghui Bao, Haoran Chang +12 | cs.RO | 2026-09-14 |
| #43 | Auto-HSI: Personalized human control of a robot swarm on demand by using LLMs for online automatic code generation | Alessandro Nazzari, Nathan Cerisara, Dorian Tonnis +5 | cs.RO | 2026-09-14 |
| #44 | Occupancy Network-Guided Autonomous Robotic Partial Nephrectomy | Ethan Kilmer, Pit Henrich, Jiawei Ge +13 | cs.RO | 2026-09-14 |
| #45 | Safe Error Correction for Language Models: Frozen-Base Adjustment with Capability Preservation | Gautam Kishore | cs.AI | 2026-09-14 |
| #46 | Discovery Foundation Models: Toward Open-Ended Discovery Intelligence | Ling Yang, Zhenfei Yin, Yingcheng Wu | cs.CL | 2026-09-14 |
| #47 | SlipSense: Multimodal Tactile Learning for Low-Latency and Generalized Slip Detection | Tong Jian, Aditya Thurvas Senthil Kumar, Xinyi Li +7 | cs.RO | 2026-09-14 |
| #48 | Bench2Dex: Benchmarking Visuo-Tactile Bimanual Dexterous Manipulation Across Dexterous Hands | Zhenjie Yang, Yideng Zhang, Dongjie Zhang +21 | cs.RO | 2026-09-14 |
| #49 | Robust and Efficient Communication for Multi-Agent Learning | Rafael Pina, Varuna De Silva, Corentin Artaud | cs.LG | 2026-09-14 |
| #50 | Reconstructing Is Not Acting: Action-Centric Latent Dynamics Modeling | Dingjie Fu, Dianxing Shi, Yangyang Xu +1 | cs.CV | 2026-09-14 |