| 1 | Mask Forcing: Improving Autoregressive Video Diffusion Distillation via Dual-Noise Masking Rollout | Zhuoran Zhao, Shengju Qian, Tongtong Liang +7 | cs.CV | 2026-09-08 |
| 2 | Rethinking Learned Occupancy in Autonomous Active Mapping with Observation-Gated Filtering | Jiahui Zhang, Bonian Han, Gongbo Liang +1 | cs.RO | 2026-09-08 |
| 3 | Spheriverse: 3D Scene Understanding from Spherical Observations in the Wild | Fei Teng, Sheng Wu, Mengfei Duan +7 | cs.CV | 2026-09-08 |
| #4 | EgoSIS: From Factorized Visual Ego-Transitions to Motion-Canonical Spatial Evidence for UAV Reasoning | Jingpu Yang, Fengxian Ji, Mingxuan Cui +4 | cs.CV | 2026-09-08 |
| #5 | FRAME: Factored Retrieval via Attribute Readouts for Object-Centric Scene Memory | Woosang Jeon, Sanghyeok Choi, Minwoo Kim +2 | cs.CV | 2026-09-08 |
| #6 | SeGDeP: Semantic- and Geometric-Aware Decoupled Prompts for Reasoning Segmentation | Linnan Zhao, Xu Liu, Lingling Li +3 | cs.CV | 2026-09-08 |
| #7 | Beyond Gait: Person Identification from Millimeter-Wave Point Clouds Across Activities of Daily Living | Xilai Wang, Zixiong Han, Saad Rhanmouni +3 | cs.CV | 2026-09-08 |
| #8 | Adaptive Anisotropic Attention for Axis-Structured Signals | Mahir Jain, Parshva Runwal, Aditya Ray Mishra +3 | cs.LG | 2026-09-08 |
| #9 | Interpretable Hyperspectral Unmixing Framework with Fixed Endmember Prior and Structured Residual Refinement | Ziyi Guan, Jianping Zhang, Qian Liu | cs.CV | 2026-09-08 |
| #10 | CVT-GS: Learning to Simplify 3D Gaussian Splatting with Centroidal Voronoi Tessellation | Bingxian Li, Yilong Li, Jingliang Peng +6 | cs.CV | 2026-09-08 |
| #11 | CoordFormer: Give Me Any Coordinates and I Will Give You Labels | Iacopo Curti, Pierluigi Zama Ramirez, Alioscia Petrelli +1 | cs.CV | 2026-09-08 |
| #12 | TriCCOT: Tri-part Convolutional Conformal Transformer for Onboard Space Object Detection | Adrien Dorise, Marjorie Bellizzi, Julia Cohen +1 | cs.CV | 2026-09-08 |
| #13 | Charts Are Beyond Pixels: Probing for Layer-Wise Chart Understanding and Editing | Xiaochuan Zhong, Yifan Hou, Chenxi Pang +1 | cs.CV | 2026-09-08 |
| #14 | SignRefine: Adapting Foundational Video Models for Sign Language Generation | Anton Pelykh, Edward Fish, Ozge Mercanoglu Sincan +1 | cs.CV | 2026-09-08 |
| #15 | AirAnchor: Bridging Local and Global Spatial Information for Zero-Shot Aerial Vision-and-Language Navigation | Shanwei Fan, Bin Zhang, Zhiwei Xu +4 | cs.CV | 2026-09-08 |
| #16 | From Coordinates to Candidate Regions: Temporal Change Localization via Region Selection in Remote Sensing Multimodal LLMs | Juwan Chung, Sungjune Park, Yeongyun Kim +1 | cs.CV | 2026-09-08 |
| #17 | ReMoMask-2: Latent Retrieval-Augmented Masked Motion Generation | Yiran Wang, Zeyu Zhang, Ling Shao +1 | cs.CV | 2026-09-08 |
| #18 | CoVeR: Coverage-Based Token Pruning for Multi-View 3D Reasoning in VLMs | Nhat-Tan Bui, Varshini Elangovan, Arun Reddy Anugu +7 | cs.CV | 2026-09-08 |
| #19 | RoboCousin: Build Your Own Simulation Playground for Robust Bimanual Robotic Manipulation | Jingxuan Zhu, Jingyi Li, LiangLiang Chen +3 | cs.RO | 2026-09-08 |
| #20 | TRIUNE-Net: Harmonizing Scale, Shape, and Efficiency in Pancreatic Tumor Segmentation | Amir Hossein Saleknia, Alireza Kheyrkhah, Sanaz Karimijafarbigloo +5 | cs.CV | 2026-09-08 |
| #21 | Human-Centric Image Captioning with Subject-Centered Spatial Understanding | Bozhou Li, Jiahang Zhang, Yue Ding +11 | cs.CV | 2026-09-08 |
| #22 | MARS-CLIP: Multi-Resolution and Attention Refined Zero-Shot Image Segmentation | Nagito Saito, Shintaro Ito, Koichi Ito +1 | cs.CV | 2026-09-08 |
| #23 | 3DWay: Generalizing Robot Manipulation via 3D Consistent Waypoints | Ziqin Huang, Yingyue Li, Chenyangguang Zhang +6 | cs.RO | 2026-09-08 |
| #24 | WSPolypNet: Weakly Supervised Polyp Localization in Colonoscopy Videos | Giseong Hwang, Minjae Jo, Yeonghyeon Park +9 | cs.CV | 2026-09-08 |
| #25 | Dual-Layer Semantic-Spatial Belief Mapping for Aerial Object Goal Navigation | Jianqiang Xiao, Xiang Deng, Yuexuan Sun +3 | cs.RO | 2026-09-08 |
| #26 | Hyperspectral Anomaly Detection via Group Sparse Low-Rank Tensor Factorization With Automatic Anomaly Grouping | Quan Yu, Yu-Hong Dai, Xiongjun Zhang | cs.CV | 2026-09-08 |
| #27 | SynthGait-19K: A Physically Grounded Synthetic Video Dataset for Gait Parameter Estimation | Soroush Mehraban, Xin Lei Lin, Vida Adeli +5 | cs.CV | 2026-09-08 |
| #28 | Learning Metamaterial Eigenmodes with Wavelet-Encoded Fourier Neural Operators | Han Zhang, Alexander Ogren, Cynthia Rudin +2 | cs.LG | 2026-09-08 |
| #29 | PocketVE: Stable and Property-Guided Structure-Based Drug Design with Variance-Exploding Diffusion | Peining Zhang, Jinbo Bi | q-bio.BM | 2026-09-08 |
| #30 | Bayesian Matrix-Valued Graphs for Context-Dependent Multivariate Relationships | Papri Dey | stat.ME | 2026-09-07 |
| #31 | A Black-Box Adversarial Attack on Human Pose Estimation and Keypoint-Based Action Recognition Models | Kacper Mroczek, Michal Kepski | cs.CV | 2026-09-07 |
| #32 | Semi-Supervised Learning under Spatially Biased Sampling | Bright Wiredu Nuakoh, Francky Fouedjio, Stephen Bradshaw +4 | cs.LG | 2026-09-07 |
| #33 | Structured Extrema Errors in Classical Surrogates for Viscous Burgers: A Physics-Consistent Interpretation | Youssef Oubari | cs.LG | 2026-09-07 |
| #34 | ReactVAU: A Slow-Fast Decoupled Framework for Streaming Video Anomaly Understanding | Chia-Hui Chen, Shih-Ying Yeh, Fu-En Yang +2 | cs.CV | 2026-09-07 |
| #35 | TDDN: Text-aligned Diffused DINO Network for Puzzle Understanding | Harsha Patnala, Debopriyo Banerjee, Ayush Sunil Munot +1 | cs.CV | 2026-09-07 |
| #36 | JEDI: JEPA-to-Edge Distillation for Efficient Cropland Segmentation from Satellite Imagery | Kishor Kumar Bhaumik, Nicolas Roque dos Santos, Jia Chen +1 | cs.CV | 2026-09-07 |
| #37 | InfluenceField: A Differentiable Field with Interventionally Identifiable Causal Structure for Multimodal World Modeling | Zihao Yang, Zijia Wang, Zhiqiu Huang | cs.LG | 2026-09-07 |
| #38 | Sub-6 GHz Over-the-Air AMC via Curriculum Fine-Tuned CNN-Transformers | Nurettin Safak, Muhammet Sefa Demirel, Alperen Marasli +3 | eess.SP | 2026-09-07 |
| #39 | Spatial Feature-wise Linear Modulation (SpFiLM) for Contrast Agent-Aware Brain Parcellation | Pushpendra Singh, Joshua R. Astley, Roman Rodionov +3 | eess.IV | 2026-09-07 |
| #40 | Privacy Leakage from a Thousand Words: Millipixel Location Recovery from Dot Maps | Yuntao Du, Tanishq Pauskar, Hao Wang +2 | cs.CR | 2026-09-07 |
| #41 | SphereSOD: Geometry-Structure Coupled Learning for 360 Salient Object Detection | Junsong Zhang, Zhijie Shen, Shuai Zheng +4 | cs.CV | 2026-09-07 |
| #42 | Zero-Shot Sim-to-Real Contact-Rich Assembly via Proprioception-Anchored Cross-Modal Pretraining | Yuhan Wang, Yurou Chen, Hongye Jiang +1 | cs.RO | 2026-09-07 |
| #43 | Topologically Consistent Agricultural Parcel Vectorization with Semantic-Guided Diffusion and Topology-Aware Polygonization | Weiqin Jiao, Xiaolong Zuo, Claudio Persello | cs.CV | 2026-09-07 |
| #44 | PICANet: Physics-Informed Cascaded Asymmetric Network for Infrared Small Target Detection | Jingjing Liu, Yinchao Han, Xianchao Xiu +2 | cs.CV | 2026-09-07 |
| #45 | Statistical versus machine learning-based spatial interpolation of post-processed ensemble weather forecasts | Mária Lakatos | cs.LG | 2026-09-07 |
| #46 | CrACK: Adversarial Attacks on Cross-Model Consistency in Collaborative Vision Foundation Models | Feifei Liu, Jintao Cheng, Chi Man Vong +1 | cs.CV | 2026-09-07 |
| #47 | CosmoH2G: A Hand-to-Gripper Transfer Dataset and Baseline Method for Object Manipulation with Complex Spatial Movements | Hongxiang Zhao, Mutian Xu, Zeyu Jin +3 | cs.RO | 2026-09-07 |
| #48 | Unified Vision-Centric Pedestrian Crossing Action Prediction via Adaptive Patch Projection and Proactive Spatial Rectification | Yao Tian, Le Yang, Binglu Wang | cs.CV | 2026-09-07 |
| #49 | Self-Supervised Multi-View 3D Gaze Target Estimation via Probabilistic Ray Marching | Keqi Chen, Vinkle Srivastav, Nicolas Padoy | cs.CV | 2026-09-07 |
| #50 | RelightFormer: Feed-forward Generative Transformer for Multiview Object Relighting | Hejun Wang, Jinxi Li, Junwei Jiang +4 | cs.CV | 2026-09-07 |