| 1 | Traffic Sign Recognition for Autonomous Driving Using Branched YOLOv2 and Geometric Features | Arefeh Rezaei | cs.CV | 2026-09-18 |
| 2 | An Interpretable Memory Decision Controller for LLM Agents Based on Three-Signal Complementarity: Decoupling Confidence and Consistency | Yiming Zhang, Jinghong Zhang, Haoran Zhao +2 | cs.CL | 2026-09-18 |
| 3 | GALA: Geometry-Aware Latent Action Modeling for Vision-Language-Action Model Pretraining across Embodiments | Yichen Liu, Puzhen Yuan, Xiang Zhu +2 | cs.RO | 2026-09-18 |
| #4 | The Role of Radiometric Features in Cross-Site Leaf-Wood Segmentation of LiDAR Point Clouds | Roman Kaharlytskyi, Derek T. Robinson, Roberto Guglielmi | cs.CV | 2026-09-18 |
| #5 | Geometric Mean Pooling for Equal-Weight Multiplicative Coarse-Graining | Ang-Kun Wu, Fangdi Wen, Jingtao Zhang | cs.LG | 2026-09-18 |
| #6 | Morphology-Aware Ambiguity Learning for Wafer Defect Decision Support | Seungjun Chu, Seokhyun Chung | cs.CV | 2026-09-18 |
| #7 | Federated Deep Clustering Networks for High-Dimensional and Heterogeneous Data | Morris Stallmann, Charalampos S. Kouzinopoulos, Marcin Pietrasik +1 | cs.LG | 2026-09-18 |
| #8 | PointLAM: Local Attentive Mamba for Efficient Point-based 3D Object Detection | Xuanming Shang, Weijia Zhang, Chao Ma | cs.CV | 2026-09-18 |
| #9 | XCalib Depth-Guided Geometric Optimization for Dense Thermal-Visible Video Registration | Aurelien Godet, Gabriel Jobert, Mauro Dalla Mura | cs.CV | 2026-09-18 |
| #10 | SFVO: Decoupled Confidence-Guided Stereo-Flow Visual Odometry with Bidirectional PnP | Kai Zhang, Guoyang Zhao, Jun Ma | cs.CV | 2026-09-18 |
| #11 | ZYT-World: A Real-Time Controllable World Model for Closed-Loop Autonomous-Driving Simulation | Boni Hu, Xiong Wei, Haoming Huang +19 | cs.CV | 2026-09-18 |
| #12 | Beyond Gaussian Worlds: Latent Geometry Matters for JEPAs | Léo Nicollier, Enric Meinhardt-Llopis, Marc Pic +2 | cs.LG | 2026-09-18 |
| #13 | 2D GauSS-MI: Efficient Active Scene Reconstruction with Balanced Visual and Geometric Quality | Yuhan Xie, Jia Pan | cs.CV | 2026-09-18 |
| #14 | Adaptive World Memory 3D Foundation Model for Scalable 3D Mapping, Localization, and Rendering | Tianchen Deng, Guole Shen, Yilin Shen +7 | cs.CV | 2026-09-18 |
| #15 | VoxelTTO: Voxel-Aligned Feed-Forward 3D Gaussian Splatting with Test-Time Optimization | Yibin Zhao, Yihan Pan, Yangwen Li +2 | cs.CV | 2026-09-18 |
| #16 | Driving on Registers, Reasoning on Risk: Risk-Aware Occupancy for Register-Based End-to-End Autonomous Driving | Jiaxing Chen, Hengduo Zou, YuKai Qin +3 | cs.AI | 2026-09-18 |
| #17 | A Scene Language Model for Open-Vocabulary Scene Mapping | Adam Lilja, Fabio Hübel, Siming He +6 | cs.CV | 2026-09-18 |
| #18 | Cube-Splat: High-Fidelity 360° Gaussian Splatting SLAM via Cubemap Factorization and Adjoint-Consistent Optimization | Xiangfei Guo, Hao Shi, Yufan Zhang +4 | cs.CV | 2026-09-18 |
| #19 | VeriFuse: Bounded Vision-Language Arbitration and Reason-Guided Refinement for Cooperative 3D Perception | Hongyi Lin, Yiyao Liu, Qi Kang +4 | cs.CV | 2026-09-18 |
| #20 | Combining Object Detection with Geometry-Aware Clustering to Distinguish Overlapping Plants in UAV Imagery | Ik Jae Lee, Hieu D. Nguyen, Mahbubur Meenar +2 | cs.CV | 2026-09-18 |
| #21 | PlaceReasoner-Beta: Reasoning-Driven Macro Placement and Benchmarking | Qiufeng Li, Chengxuan Wang, Rongqian Chen +6 | cs.AI | 2026-09-18 |
| #22 | Geometry-Aware Diffusion Guidance via Curvature-Adaptive Tubular Correction | Enze Jiang, Jinwei He, Zheng Ma | cs.CV | 2026-09-18 |
| #23 | FOCAL-VLA: Subtask-Guided Geometry Distillation and Implicit World Modeling for Vision-Language-Action Models | Zhiyuan Gao, Di Wen, Yanxiang Zhan +4 | cs.RO | 2026-09-18 |
| #24 | VGGT-CAD: Reconstructing Parametric CAD 3D Model with Geometric Grounding | Chunan Yu, Tianrun Chen, Fu Shen +3 | cs.CV | 2026-09-18 |
| #25 | Robust Structureless Monocular Visual Inertial Initialization Exploiting Line Features and Vanishing Points | Junwan Choi, Woongrae Jo, Dong-Uk Seo +2 | cs.RO | 2026-09-18 |
| #26 | Implicit Rule Induction with Test-Time Task Embeddings in ARC-like Tasks | Adrien Deliège, Claas Beger, Marc Van Droogenbroeck +1 | cs.AI | 2026-09-18 |
| #27 | 4DGS-Fixer: Generative Sparse-View 4D Gaussian Splatting with Iterative Refinement Guided by Video Diffusion Priors | Haitao Huang, Shenghao Zhao, Boyuan Tian +7 | cs.CV | 2026-09-18 |
| #28 | GeoAAC: Geometry-Based Adaptive Action Chunking from Denoising Trajectories in VLA Policies | Xin Chen, Sen Chen, Yujuan Ding +5 | cs.RO | 2026-09-17 |
| #29 | The First-Order Oracle Complexity of Lipschitz Convex Optimization in Nondual Settings | David Martínez-Rubio, Brian Bullins, Cristóbal Guzmán +1 | math.OC | 2026-09-17 |
| #30 | Learning Foresight without Explicit Trajectories for 3D Diffusion Policies | Zhongbo Zhang, Zaibin Zhang, Yifan Wang +3 | cs.RO | 2026-09-17 |
| #31 | DexTouch-WM: Learning Action-Conditioned Tactile World Models from Human Touch for Dexterous Robot Manipulation | Yan Qin, Yue Chen, Wenwei Lin +8 | cs.RO | 2026-09-17 |
| #32 | CoRef-GS: Cooperative Referring Gaussian Splatting for Multi-Agent Scene Understanding | Zhikun Zhou, Kunyu Peng, Runyi Yang +7 | cs.RO | 2026-09-17 |
| #33 | TAP Accuracy Below the Fluctuation Scale and Universal Posterior Geometry in Spherical Linear Models | Jingbo Liu, Zhiyuan Yu | stat.ML | 2026-09-17 |
| #34 | Online Supervised Dimension Reduction with Random Features: Diagnostics and Computational Trade-offs | Zhenlin Yao, Wei Xiong | stat.ML | 2026-09-17 |
| #35 | Xeno-Interpretability: Investigating the Alien Minds of LLMs | F. Pierucci, M. Bracale Syrnikov, M. Prandi +3 | cs.CL | 2026-09-17 |
| #36 | Navi-Agent: Unlocalized Monocular Navigation Agent | Wenyuan Xie, Mengyang Hong, Yongzhong Wang +9 | cs.RO | 2026-09-17 |
| #37 | EliGSiR: Continual RGB-D Mapping with Gaussian Splatting under Bounded Compute | Björn Ellensohn, Elmar Rueckert, Christian Rauch | cs.CV | 2026-09-17 |
| #38 | NeuSOGA3D: A Neuro-Symbolic Framework for Explainable 3D Geometric Reconstruction | Qingde Li, Qingqi Hong, Zihan Li +1 | cs.AI | 2026-09-17 |
| #39 | GRF-Recon: Global Ray-Field Optimization for Long-Sequence Feed-forward Reconstruction | Enpeng Li, Yunzhou Zhang, Zhiyao Zhang +4 | cs.CV | 2026-09-17 |
| #40 | LapaTrack-3D: 6 DoF pre-operative shape tracking for laparoscopic surgery | Jingwei Song, Javid Hussain Jakir, Ray Zhang +4 | cs.RO | 2026-09-17 |
| #41 | DirtyMoCap: Robust Motion Capture from Unconstrained Markers | Long Wang, Shuting Zhao, Shen Yan +5 | cs.CV | 2026-09-17 |
| #42 | CitySTAR: Structured and Topology-Aware Reasoning for Open-Vocabulary Urban 3D Grounding | Shuai Zhang, Hongye Hou, Qinghe Liu +5 | cs.CV | 2026-09-17 |
| #43 | SlugTrails: An Egocentric Benchmark for Floor Plan Localization in Large Buildings | Yunqian Cheng, Roberto Manduchi | cs.CV | 2026-09-17 |
| #44 | PACE: Precise AI Cinematic Expression: A Typed Specification for Script-Grounded Previsualization and Geometric Conformance | Bing Duan, Qiang Guo, Linpu Li +6 | cs.CV | 2026-09-17 |
| #45 | Constraint-Safe Graph-Context Scoring for Stable Point-Feature Labels Under Text-Width and Accessibility-Inspired Profiles | Taimoor Ahmad | cs.AI | 2026-09-17 |
| #46 | SnapPhysics: A Physics-Aware Scene Graph from a Single View for Interactive Mixed Reality Scenes | Suji Kang, Seok-Young Kim, Young Bin Kim +4 | cs.CV | 2026-09-17 |
| #47 | TorchCraft: Unified binder design by inverting an all-atom structure predictor | TorchCraft Team, Yu Liu, Zhouhanyu Shen +8 | cs.AI | 2026-09-17 |
| #48 | OceanMoE: Structured Conditional Sparse Computation for Long-Horizon Multivariate Ocean Forecasting | Yishun Zhu, Jian Wang | cs.LG | 2026-09-17 |
| #49 | GAPrompt++: Multi-Granular Geometry-Aware Point Cloud Prompt for 3D Vision Model | Zixiang Ai, Zhenyu Cui, Yufei Guo +4 | cs.CV | 2026-09-17 |
| #50 | Instance Segmentation and Fine-grained Classification for Urban Buildings with Adaptive Region Dividing and Spatially-Supervised Contrastive Learning | Weiyuan Zhang, Qi Zhang, Hui Huang | cs.CV | 2026-09-17 |