| 1 | A Low-Cost, Open Platform for End-to-End Autonomous Driving on a Miniature Ackermann Vehicle | Gustavo Claudio Karl Couto, Eric Aislan Antonelo, Gabriel George Zipperer | cs.LG | 2026-09-03 |
| 2 | Continuous Actions from Discrete Minds: Latent-Aligned Planning for End-to-End Autonomous Driving | Ruoyu Yao, Yusen Xie, Qingzhao Liu +5 | cs.CV | 2026-09-03 |
| 3 | Understanding Autonomous Driving Datasets by Describing Differences between Image Subsets in Natural Language | Julian Truetsch, Felix Hauser, Christoph Stiller +1 | cs.CV | 2026-09-03 |
| #4 | SV-WAM: An Efficient Surround-View World-Action Model for End-to-End Autonomous Driving | Jinyang Wang, Shiwei Li, Junjian Wang +12 | cs.CV | 2026-09-03 |
| #5 | Drive-HWM: Hierarchical World Models for Dynamic-Latent Guided Autonomous Driving | Zhaoxin Fan, Tianbao Zhang, Wenjun Wu +5 | cs.CV | 2026-09-03 |
| #6 | ShallowStream: Index Shallow then Answer Deep for Streaming Video Understanding | Jitai Hao, Ke Yang, Qiang Huang +1 | cs.CV | 2026-09-02 |
| #7 | VIPS: Vehicle-Infrastructure Cooperative Planning Benchmark via Pseudo-Simulation | Hoonhee Cho, Jae-Young Kang, Giwon Lee +3 | cs.CV | 2026-09-02 |
| #8 | Towards Zero-Shot Transfer Across Embodiments For Driving VLAs | Caio Azevedo, Stefano Sabatini, Sascha Hornauer +1 | cs.CV | 2026-09-02 |
| #9 | CrashDiffuser: VLM-Guided Collision Intent Reasoning for Fine-Grained Safety-Critical Traffic Scenario Generation | Shucheng Zhang, Yuang Zhang, Bingzhang Wang +3 | cs.RO | 2026-09-02 |
| #10 | DiffuSearch: How Hybrid Trajectory Planning Benefits from Aligned Objectives in Diffusion and Action Space | Steffen Hagedorn, Aron Distelzweig, Alexandru P. Condurache | cs.RO | 2026-09-02 |
| #11 | TAPVid-MV: A Benchmark for Tracking Any Point in 3D Across Multiple Views | Skanda Koppula, Frano Rajic, Abdullah Faiz Ur Rahman +9 | cs.CV | 2026-09-01 |
| #12 | Monocular Depth Estimation from a Single Image: Progress and Opportunities | Muxin Liu, Xiaoyang Lyu, Yang-Tian Sun +4 | cs.CV | 2026-09-01 |
| #13 | DNC-IMM: Early Lane-Change Intention Recognition via Neural Calibration Based on Driving Context Information | Woong-Chan Byun, Seung-Hyun Kong | cs.RO | 2026-09-01 |
| #14 | Vision-Language-Guided Pseudo-Labels for Unsupervised Domain Adaptation in Semantic Segmentation for Waste Sorting | Udo Schlegel, Shubhangi, Gabriel Dax +3 | cs.CV | 2026-09-01 |
| #15 | CoLT-Drive: Counterfactual Long-Tail Benchmarking and Knowledge-Preserving Adaptation for Driving Affordance Prediction | Zhengxu Tang, Guofeng Cui, Ziyu Gong +8 | cs.CV | 2026-08-31 |
| #16 | Beyond Textual Chain-of-Thought: A Survey on Action-Grounded Reasoning in Autonomous Driving | Zhengxu Tang, Xiaozhou Zhang, Guofeng Cui +10 | cs.CV | 2026-08-31 |
| #17 | Qwen-Drive-1.0: An Initial Step towards a Vision-Language Foundation Model for Autonomous Driving | Xin Zhou, Zongchuang Zhao, Zhibo Yang +13 | cs.CV | 2026-08-31 |
| #18 | Driving on Memory | Christian Löwens, Thorben Funke, Alexandru Paul Condurache | cs.CV | 2026-08-31 |
| #19 | Physical Adversarial Examples for Person Detectors in Thermal Images Based on 3D Modeling | Xiaopei Zhu, Siyuan Huang, Zhanhao Hu +3 | cs.CV | 2026-08-31 |
| #20 | What Emerges and What Breaks in Self-Play Driving | Laur Sisask, Ardi Tampuu, Tambet Matiisen | cs.LG | 2026-08-31 |
| #21 | Real-Time Scene-Adaptive Tone Mapping for High-Dynamic Range Object Detection | Gongzhe Li, Linwei Qiu, Peibei Cao +3 | cs.CV | 2026-08-31 |
| #22 | Towards Continual Test-Time Adaptation of Vision-Language Models in Open-Vocabulary Semantic Segmentation | Chandler Timm C. Doloriel, Yunbei Zhang, Sarthak Kumar Maharana +5 | cs.CV | 2026-08-30 |
| #23 | How Far Can 5,500 Hours of Driving Take You? A Scaling Law Analysis of Video Diffusion Models | Victor Besnier, Anh-Quan Cao, Elias Ramzi +5 | cs.CV | 2026-08-28 |
| #24 | Safety by Design: Realized-Cost Constraints for Contextual Bandits with Continuous Actions | Spyros Dragazis, Aldo Pacchiano | cs.LG | 2026-08-27 |
| #25 | DPA-I2P: Depth-Guided Projective Alignment for Image-to-Point-Cloud Registration in Autonomous Driving | Wenxin Zhang, Hang Li, Zhiwei Xu +3 | cs.CV | 2026-08-27 |
| #26 | Gating Before Commitment: Anticipating Intent Divergence to Prevent Post-Interaction Decision Failures in Autonomous Driving | Cong Xu, Ravi Sankar | cs.RO | 2026-08-26 |
| #27 | Projection-Aware End-to-End Learned Video Compression for 360-Degree Video | Niloofar Maani | cs.CV | 2026-08-26 |
| #28 | SkyDrive: Learning to Drive in a New City from Aerial Traffic Monitoring | Weijiang Xiong, Lan Feng, Alexandre Alahi +1 | cs.RO | 2026-08-25 |
| #29 | CARE: Camera-Residual Reserves for First Sightings in Adaptive LiDAR Sensing | Jiachen Gong, Yun Li, Ehsan Javanmardi +2 | cs.CV | 2026-08-25 |
| #30 | GeoWAM: Visual Geometry World Action Models for Autonomous Driving | Yiren Lu, Xin Ye, Jiaming Liu +10 | cs.CV | 2026-08-24 |
| #31 | MomADv2: Reliable Temporal Memory for End-to-End Autonomous Driving | Ziying Song, Shengkai Zhang, Lin Liu +8 | cs.CV | 2026-08-24 |
| #32 | LoViF 2026 The First Challenge on Unified Removal of Raindrops and Reflections: Methods and Results | Zewei He, Xi Tong, Yu Chen +49 | cs.CV | 2026-08-24 |
| #33 | Scaling Curriculum Learning For Autonomous Driving | Cevahir Koprulu, David Paz, Feng Tao +4 | cs.AI | 2026-08-23 |
| #34 | A2DINOv3: Rethinking Multi-Modal Object Detection via Socialized Collaboration | Jiekang Feng, Zhihe Fan, Yunqi Zhu +5 | cs.CV | 2026-08-21 |
| #35 | CoAnchor: Robust Collaborative Perception under Spatio-Temporal Misalignment via Object-Level Anchors | Chi Li, Rui Lin, Aobo Ji +1 | cs.CV | 2026-08-21 |
| #36 | WA-JEPA: Rethinking the Video JEPA Paradigm for World-Action Modeling in Autonomous Driving | Xinlin Wang, Yujiao Xiang, Yuheng Zhou +11 | cs.CV | 2026-08-21 |
| #37 | A Collaborative Multi-Modality Interaction for VLA-based End-to-End Autonomous Driving | Jingtao Sun, Xiaohai He, Yike Zhang +4 | cs.CV | 2026-08-21 |
| #38 | Multi-Modal Traffic Sign Detection with Semantic Attributes for Autonomous Driving | Meda Lazar, Sourab Sridhar, Shashwata Gupta +3 | cs.CV | 2026-08-21 |
| #39 | Multi-Agent Orchestration with the Common-Sense Reasoning Capabilities of LLMs for Autonomous Driving | Mehdi Azarafza, Faezeh Pasandideh, Ali Ehteshami Bejnordi +2 | cs.MA | 2026-08-20 |
| #40 | G-MARK: Grounded Multi-Agent Reasoning for Cooperative Driving via Knowledge Graphs | Bhavya Gupta, Onat Gungor, Tajana Rosing | cs.LG | 2026-08-20 |
| #41 | SCAPE: Scenario-Conditioned Simulation-Augmented Policy Evaluation | Dijie Zhu, Seunghun Oh, Ruopeng Huang +3 | cs.RO | 2026-08-19 |
| #42 | CAViAR: A Causal Video Dataset for Fine-Grained Accident Reasoning in Real-World Scenarios | Sparsh Garg, Yi-Wen Chen, Vijay Kumar B G +1 | cs.CV | 2026-08-19 |
| #43 | SceneGTMM: A Conformal Mapping-based Scene-Aware Transferable GNN-Transformer Dual-Graph Interaction Framework for Map Matching | Yongliang Zhang, Feng Song, Ji Chen +6 | cs.CV | 2026-08-19 |
| #44 | DA-WAM: Decision-Aligned Future Latents for Driving World Models | Ruiguo Zhong, Benshan Ma, Xiaolong Chen +5 | cs.RO | 2026-08-19 |
| #45 | USR-Drive: Unified Driving Scene Representation via Joint Denoising of 3D Gaussians and Boxes | Li-Heng Chen, Haokai Pang, Chengye Su +7 | cs.CV | 2026-08-19 |
| #46 | One-Stage Object Detectors in Autonomous Driving | Jonel Roman, Ryan Sirjue, Peter Nguyen +3 | cs.CV | 2026-08-19 |
| #47 | TestifAI: Tomography-Based Testing for Deep Learning Systems | Arooj Arif, Tobias Hartung, Elena Botoeva +1 | cs.AI | 2026-08-19 |
| #48 | The Impact of CutMix on Reliability and Robustness in Semantic Segmentation | Steven Landgraf, Markus Ulrich | cs.CV | 2026-08-19 |
| #49 | Plug-and-Play Traffic Element Awareness for End-to-End Autonomous Driving | Zongzheng Zhang, Jijun Wang, Saining Zhang +8 | cs.CV | 2026-08-18 |
| #50 | Geo-VLA: Geometry-Aware Vision-Language-Action Planning via Internalization of Map Semantics | Ran Chen, Jiaxing Ren, Zhikun Zhang +3 | cs.RO | 2026-08-18 |