| 1 | SolarWM: Open Data and Scalable Training for Long-Horizon Video World Models | Junchao Huang, Guian Fang, Shengju Qian +15 | cs.CV | 2026-09-02 |
| 2 | Discriminative World Models for Web Agents | Kelvin Li, Dhruv Pendharkar, Anish Pahilajani +6 | cs.AI | 2026-09-02 |
| 3 | Graph Machine: Towards Better Pretraining via Edges | Lintai Hou | cs.LG | 2026-09-02 |
| #4 | Towards Trustworthy Autonomous Robots: An Explainable AI-Based Decision Framework | Cagri Temel | cs.RO | 2026-09-02 |
| #5 | RoGe: Novel View Synthesis via End-to-End Implicit Reconstruction and Generation | Xiaolei Lang, Ze Kang, Zehao Huang +1 | cs.CV | 2026-09-02 |
| #6 | Learning Spectral-Like Mesh-Free Discretisations | Lucas Gerken Starepravo, Henry Broadley, Steven Lind +1 | physics.comp-ph | 2026-09-02 |
| #7 | GDB-Reward: From Evaluation Metrics to Training Rewards for Graphic Design | Adrienne Deganutti, Purvanshi Mehta, Simon Hadfield +1 | cs.CV | 2026-09-02 |
| #8 | Bilevel Coordinated Reflection: A Game-Theoretic Approach to Multi-Agent LLM Systems | Yihang Chen, Yuxiang Chen, Yuxuan Huang +3 | cs.AI | 2026-09-02 |
| #9 | Balancing Frequencies and Pixels in Flow Matching | Lucas Degeorge, Paul Couairon, Arijit Ghosh +3 | cs.CV | 2026-09-02 |
| #10 | A Top-Down Framework for Metric-Scale Athlete Localization from Single Broadcast Frames | Thanh-Khoi Nguyen, Hoang-Phuc Nguyen, Linh-Huynh +1 | cs.CV | 2026-09-02 |
| #11 | Eliciting ESG Preferences for Reinforcement Learning-Based Portfolio Optimization | Giovanni Dispoto, Marcello Restelli, Carmine Ventre | q-fin.PM | 2026-09-02 |
| #12 | Query Rewriting for Complex Object Segmentation in 4D Gaussian Representations | Thanh-Khoi Nguyen, Thien-Phuc Tran, Minh-Triet Tran | cs.CV | 2026-09-02 |
| #13 | Differentiable Electricity-Market Clearing for Gradient-Based Planning | Luca Mungo, Maarten P. Scholl, Arnau Quera-Bofarull | cs.LG | 2026-09-02 |
| #14 | Learning to Attract and Repel: Dual Quality Margin Learning for Face Recognition (DQM-Face) | El Ouanas Belabbaci, Bhavesh Wani, Philipp Terhörst | cs.CV | 2026-09-02 |
| #15 | Source Distribution Estimation by Posterior Averaging | Trung-Dung Hoang, Lisa M. Koch | cs.LG | 2026-09-02 |
| #16 | Predictors of Loneliness in Older Adults Using Multimodal Analysis of Speech and Language | Vinmay Khandode, Sai Karthik Kosuri, Neil K. R. Sehgal +6 | cs.CL | 2026-09-02 |
| #17 | MARS: What Retrieval Signals Are Hidden in Multimodal Large Language Models for Text-Video Retrieval? | Uicheol Jung, Juyoung Hong, Geuntaek Lim +1 | cs.CV | 2026-09-02 |
| #18 | Stereo 4D Radar for 3D Object Detection: Integrating Geometric Alignment and Absolute Velocity Estimation | Seung-Hyun Song, Dong-Hee Paek, Woong-Chan Byun +1 | cs.CV | 2026-09-02 |
| #19 | Blending Concepts: Benchmarking Visual Metaphor Generation in Text-to-Image Models | Chuer Chen, Zichen Wang, Yi He +2 | cs.CV | 2026-09-02 |
| #20 | Addressing Trust in AI Systems through Education: A Didactic Perspective | Pierre Haritz, Hendrik Krone, Thomas Liebig | cs.CY | 2026-09-02 |
| #21 | WiFlow: Estimating Optical Flow using WiFi Channel State Information | Thomas Weigel, Simon Kiefhaber, Fabian Portner +2 | cs.CV | 2026-09-02 |
| #22 | Uncertainty-Guided Adverse Weather Restoration via Gated Transformer Network | Zheke Jin, Yuning Cui, Tianle Jin +2 | cs.CV | 2026-09-02 |
| #23 | IFW-BLS: Dual-Robust Broad Learning System with Intuitionistic Fuzzy Wave Loss | Mushir Akhtar, M. Tanveer | cs.LG | 2026-09-02 |
| #24 | Information Density Imbalance in Visual Object Detection | Ziwei Zhao, Yanxi Lu, Yuwei Hu +8 | cs.CV | 2026-09-02 |
| #25 | TempoGround: State-Aware Streaming Visual Grounding with Vision-Language Models | Leqian Ding, Junning Qiu, Manwen Yang +2 | cs.CV | 2026-09-02 |
| #26 | Fair Stable Matching: A Nash Social Welfare Approach | Parth Desai, Rasheed M, Ganesh Ghalme +1 | cs.GT | 2026-09-02 |
| #27 | Towards Zero-Shot Transfer Across Embodiments For Driving VLAs | Caio Azevedo, Stefano Sabatini, Sascha Hornauer +1 | cs.CV | 2026-09-02 |
| #28 | AGI Maze Prediction Datasets: A Compact Benchmark for Learning World Dynamics with Transformers | Alexey Potapov | cs.LG | 2026-09-02 |
| #29 | Poisoning Attacks on the PGM-index | Atsuki Sato, Martin Aumüller, Yusuke Matsui | cs.DB | 2026-09-02 |
| #30 | YesTrack: Referring Multi-Object Tracking via MLLM-based Yes/No Verification | Quansheng Hu, Qin Sun, Qiansen Dai +4 | cs.CV | 2026-09-02 |
| #31 | Domain shift-robust object detection with GenAI image editing | Isabel D. Stein, Thijs A. Eker, Sebastiaan P. Snel +4 | cs.CV | 2026-09-02 |
| #32 | If It Moves, Radar Knows: A Physics-Aware Radar Transformer for Class-Agnostic Moving-Object Detection | Yinghao Sun, Shuguang Li, Jinliang Shao +1 | cs.CV | 2026-09-02 |
| #33 | DiffuSearch: How Hybrid Trajectory Planning Benefits from Aligned Objectives in Diffusion and Action Space | Steffen Hagedorn, Aron Distelzweig, Alexandru P. Condurache | cs.RO | 2026-09-02 |
| #34 | RideSkill: A Hierarchical Algorithm for Generalized Ride Sharing with LLM-Driven Automatic Evolution | Zijian Zhao, Sen Li, Xialiang Tong +1 | cs.MA | 2026-09-02 |
| #35 | Task-Level Natural Language Priors as Learning Signals for Low-Resource LLM Training | Jian Gao, Xiao Zhang, Xun Zhu +2 | cs.AI | 2026-09-02 |
| #36 | Recursive Value Learning for Long-Horizon Offline Goal-Conditioned RL | Hyeonseong Jeon, Youngwoon Lee | cs.LG | 2026-09-02 |
| #37 | InfraPatch: Cross-Task Targeted Grayscale Patch Attacks on Infrared-Adapted Vision-Language Models | Chengyin Hu, Dingyi Lu, Jiaju Han +5 | cs.CV | 2026-09-02 |
| #38 | PEARL: Path-Entity Aligned Relational Learning with Contextual Subgraphs for Inductive Knowledge Graph Completion | Yunchi Yang, Longlong Li, Cunquan Qu | cs.AI | 2026-09-02 |
| #39 | Progressive Pseudo-Label Optimization for Point-Supervised Change Detection | Hailong Ning, Hao Wang, Yimeng Wang +3 | cs.CV | 2026-09-02 |
| #40 | OBJECTION! Lawyer Agents Mitigate Guilty Bias in Legal Judgment Prediction | Jaehoon Jeong, Jay-Yoon Lee | cs.CL | 2026-09-02 |
| #41 | Beyond Modality Harmony: Orthogonal Purification and Topology-Guided MoE for Conflict-Aware Multimodal Recommendation | Jialin Liu, Zhaorui Zhang, Ray C. C. Cheung | cs.IR | 2026-09-02 |
| #42 | Online Non-Monotone DR-Submodular Maximization Matching the Offline $0.401$ Factor | Vaneet Aggarwal, Yiyang Lu | cs.LG | 2026-09-02 |
| #43 | SoK: Where Do Flow Labels Come From? Auditing Label Provenance in Encrypted Traffic Benchmarks | Sizhe Huang, Shujie Yang | cs.NI | 2026-09-02 |
| #44 | Beyond Context Windows: Persistent Discovery Context for Data-Centric Agents | Jalal Mahmud | cs.AI | 2026-09-02 |
| #45 | Scalable Bayesian Optimization of Composite Functions for Image-Based Inverse Problems in Materials Characterization | Dasol Yoon, Poompol Buathong, Chia-Hao Lee +3 | cs.LG | 2026-09-02 |
| #46 | Synergistic Information Disentanglement for Omni-modal Slide Representation Learning in Computational Pathology | Mingxin Liu, Chengfei Cai, Anwen Lu +5 | cs.CV | 2026-09-02 |
| #47 | Semantic Signal-Assisted Inspection and Recovery Allocation in Reverse Logistics | Jiani He, Dingyan Shang, Yihua Xu +4 | cs.AI | 2026-09-02 |
| #48 | A Unified Rate-Distortion Perspective on Vector, Product, and Scalar Quantization | Xianghong Fang, Wenlong Mou, Yuan Yuan +2 | cs.LG | 2026-09-02 |
| #49 | Git4Data: Database-Native Version Control for AI Agents | Hongshen Gou, Zuyu Zhang, Yuze Sun +4 | cs.DB | 2026-09-02 |
| #50 | Modeling What Changes: Sparse, Residual World Models for Object-Centric Manipulation | Param Thakkar, Parsika Paresh Shah, Manisha Sushant Gote | cs.RO | 2026-09-02 |