| 1 | TANGO: Humanoid Navigation in Cluttered Environments with a Whole-Body Vision-Language-Action Model | Anqi Li, Yuxin Chen, Zhaobo Li +4 | cs.RO | 2026-09-08 |
| 2 | Point4D: Long-range 4D Motion Reconstruction | Minsik Jeon, Jay Karhade, Deva Ramanan +1 | cs.CV | 2026-09-08 |
| 3 | The Audit Decides the Verdict: Instrument Effects Rival Demographic Bias in LLM Decision Audits | Siddharth Vohra, Manikandan Ravikiran | cs.CL | 2026-09-08 |
| #4 | Flexible Spectral-Normalized Neural Gaussian Process for Dynamic Aperture Prediction | Yousra El-Bachir, Frederik Van der Veken, Davide di Croce +4 | physics.acc-ph | 2026-09-08 |
| #5 | Dynamics of meaning: Towards the Evaluation of Diachronic Semantic Change in Sinhala | Nevidu Jayatilleke, Nisansa de Silva | cs.CL | 2026-09-08 |
| #6 | AURORA: Active Uncertainty-Driven Re-Orientation for In-Hand Reconstruction | Feiyu Zhao, Yuetong Li, Chenxi Xiao | cs.RO | 2026-09-08 |
| #7 | Segment Any Motion with Radar: Robust Multimodal Moving-Object Segmentation and Tracking | Jue Wang, Xuan Wang, Hao Zhou +6 | cs.CV | 2026-09-08 |
| #8 | A Multi-Modal Perception Pipeline for Object Detection and Tracking in Autonomous Racing | Davide Malvezzi, Michele Pestarino, Vittoria Cavicchioli +9 | cs.RO | 2026-09-08 |
| #9 | Supervised Cross-Modal Feature Alignment for Zero-Wearable Freezing of Gait Detection in Parkinsonism | Aryan Singh, Chandan Biswas | cs.CV | 2026-09-08 |
| #10 | Tracking-by-detection in Multi-object Tracking: Survey and Experiments | Yujin Yang, Kyujin Shim, Kangwook Ko +1 | cs.CV | 2026-09-08 |
| #11 | TTGBench: Benchmarking Topological Evolution and Semantic Drift in Text-attributed Temporal Graphs | Longfei Ma, Zemin Liu, Fei Wu | cs.AI | 2026-09-08 |
| #12 | A Black-Box Adversarial Attack on Human Pose Estimation and Keypoint-Based Action Recognition Models | Kacper Mroczek, Michal Kepski | cs.CV | 2026-09-07 |
| #13 | Kalman Delta Networks: Uncertainty-aware Associative Memory | Ngoc Bui, Tinglin Huang, Rex Ying | cs.LG | 2026-09-07 |
| #14 | DroneGround: Open-Vocabulary Drone Payload Characterization Using Synthetic Data and Grounded Vision-Language Models | Ami Pandat, Rajasekhar Punna, Gopika Vinod +1 | cs.CV | 2026-09-07 |
| #15 | LLM Forensics: Where Do Backdoors Hide? Localizing and Controlling Trigger Mechanisms with Sparse Autoencoders | Wissam Antoun, Francis Kulumba, Théo Lasnier +2 | cs.CL | 2026-09-07 |
| #16 | TFTrack: A Template-Free Framework for Efficient 3D Point Cloud Tracking | Zhaofeng Hu, Sifan Zhou, Jiahao Nie +3 | cs.CV | 2026-09-07 |
| #17 | CrowdTraj: A Benchmark for Dense Crowd Trajectory Prediction in Realistic Crowded Environments | Antonius Bima Murti Wijaya, Paul Henderson, Marwa Mahmoud | cs.CV | 2026-09-07 |
| #18 | No-Regret Mixing of LRU and LFU with Optimal Switching Cost | Younes Ben Mazziane, Xinying Zou | cs.LG | 2026-09-07 |
| #19 | Re-engineering SORT-based algorithms for low-cost small object tracking from omnidirectional footage | Xin Shu, Meegan Gower, Yvonne Buckley +1 | cs.CV | 2026-09-07 |
| #20 | Zero-Shot Sim-to-Real Contact-Rich Assembly via Proprioception-Anchored Cross-Modal Pretraining | Yuhan Wang, Yurou Chen, Hongye Jiang +1 | cs.RO | 2026-09-07 |
| #21 | Inferring Urban Mobility Interactions from Aggregated Dynamics | Yi Wang, Jing Li, Jinliang Deng +5 | cs.LG | 2026-09-07 |
| #22 | LightSplat: Real-Time High-Fidelity 3D Gaussian SLAM with Loop Closure | Junze Bao, Ye Gao, Yiming Huang +5 | cs.RO | 2026-09-07 |
| #23 | NutriBench-Kitchen: Benchmarking Embodied AI for Nutrition Management | Yulin Wei, Xiangchen Wang, Jianhui Pan +5 | cs.CV | 2026-09-07 |
| #24 | From LLM-Generated Specifications to Learned Quadruped Locomotion | Merve Atasever, Keyan Azbijari, Cagan Bakirci +5 | cs.RO | 2026-09-07 |
| #25 | Continuous Token-Level Spatio-Temporal Context Modeling for Visual Object Tracking | Ding Xia, Meiqin Liu, Jing Zhou +1 | cs.CV | 2026-09-07 |
| #26 | Aha-Flow Distillation: Flow Markers Matter in LLM Reasoning | Xiaodong Wang, Peixi Peng | cs.CL | 2026-09-07 |
| #27 | Fine-Grained Visual Preprocessing and Dual-Stream Temporal Modeling for Multimodal Sentiment Analysis on Social Media | Su Li, Yigong Zhang, Lei Xiong +1 | cs.CV | 2026-09-07 |
| #28 | Tracking the Moving Frontier: Long-Short Term Advantage Estimator | Xinhao Yao, Lu Yu, Changhao Wang +5 | cs.LG | 2026-09-06 |
| #29 | One Step, One Lead: Mitigating Higher-Order Interference in Multi-Domain Reinforcement Learning via Cross-Step Control | Zihan Lin, Xiaohan Wang, Jie Cao +4 | cs.LG | 2026-09-06 |
| #30 | An Integrated Video-AI Platform for Action-Level Microanastomosis Training and Performance Feedback | Yan Meng, Daniel A. Donoho | cs.CV | 2026-09-06 |
| #31 | CONTINUITY: Security-Context Contracts for Composable LLM Agent Controls | Chris Zheng, Geng Yang | cs.CR | 2026-09-04 |
| #32 | Cross-Domain Tracker Adaptation Without Target-Domain Labels via Vision-Language Agents | Daniel Davila, Ravikumar Balakrishnan, Mike Cochran | cs.CV | 2026-09-04 |
| #33 | Repeated Queries Exhaust an LLM's Brand Recommendations but Not Its Sources | Dmitrij Żatuchin | cs.IR | 2026-09-04 |
| #34 | Qlippy: A Retrieval-Augmented GenAI Assistant for Reproducible Quantum Workflows and Experiment Tracking | Mahee Gamage, Vlad Stirbu | quant-ph | 2026-09-04 |
| #35 | PuTR-CouT: Counting-by-Tracking in Camera-Trap Image Sequences | Fagner Cunha, Juan G. Colonna, Eulanda M. dos Santos | cs.CV | 2026-09-04 |
| #36 | Predicting Spatiotemporal Mobile Sensing-Based PM2.5 Concentrations Using Low-Rank Adapted Spatially Attentive Graph Neural Network | Om Chiddarwar, Priyanka Mandal, Praveen Kumar Chandaliya +1 | cs.AI | 2026-09-04 |
| #37 | Hidden In Plain Gaze: Gaze Representations as Privacy Controls for Utility and Re-identification Risk in XR | Cory Ilo, Brendan-David John, Doug A. Bowman | cs.CV | 2026-09-04 |
| #38 | Temporal Self-Distillation: Learning Visual State Tracking in Videos Without Supervision | Shravan Venkatraman, Wenshuai Zhao, Mohammad Hassan Vali +1 | cs.CV | 2026-09-03 |
| #39 | Last Translation Benchmark | Vilém Zouhar, Niyati Bafna, Mukund Choudhary +241 | cs.CL | 2026-09-03 |
| #40 | IRWOZ 2.0: A Large Language Model-driven Dialogue Dataset for Industrial Robot Conversations | Chen Li, Dimitrios Chrysostomou | cs.AI | 2026-09-03 |
| #41 | ENEAS: Embedding-guided Neural Ensemble for Adaptive Segmentation | Javier del Pino, Salvador Rodríguez, Alejandro Garabito +2 | cs.CV | 2026-09-03 |
| #42 | LeanGRPO: Eliminating Redundant Recomputation in Diffusion RL | Sijie Wang, Zhiqiang Tan, Xinrui Yang +1 | cs.LG | 2026-09-03 |
| #43 | LongCounsel-8: A Benchmark Suite for Longitudinal Depression Tracking from Multi-Session Counseling Dialogues | Jiayi Li, Zhaomin Wu, Bingsheng He | cs.LG | 2026-09-03 |
| #44 | BRIDGE: An Open-Source Humanoid Platform via Morphology-Control Co-Design for Physical AI | Jianren Wang, Letian Qian, Zikai Wang +4 | cs.RO | 2026-09-03 |
| #45 | BMCTrack-d: Pig re-identification and tracking via back marks in challenging camera settings | David Brunner, Maciej Oczak, Marie Bordes +3 | cs.CV | 2026-09-03 |
| #46 | Counting Animals in Camera-Traps Image Sequences without Count Labels: Winning Solution to the iWildCam 2021 Challenge | Fagner Cunha, Juan G. Colonna, Eulanda M. dos Santos | cs.CV | 2026-09-03 |
| #47 | VeriPhy: Agentic Physical Reasoning for World Model Evaluation and Refinement | Wenzhuo Xu, Yuchen Zhu, Chongjian Ge +8 | cs.CV | 2026-09-02 |
| #48 | Distilling deep optical flow stereo methods to retrieve dense three-dimensional wind fields | Thomas J. Vandal, Dong L. Wu, James L. Carr +5 | cs.LG | 2026-09-02 |
| #49 | MuyBridge: Mobile Human Center-of-Mass Estimation from Monocular Video via Sparse Fusion | Aidan Bradshaw, Marco Giordano, David Rode +8 | cs.CV | 2026-09-02 |
| #50 | The Implications of Linguistic Illegibility for LLM Security | James Mickens | cs.LG | 2026-09-02 |