| 1 | GraphFAS: A Distributed System for Automated Graph Feature Generation and Selection in Industrial Transaction Networks | Yice Luo, Yun Zhu, Xi Chen +7 | cs.LG | 2026-09-08 |
| 2 | Q2D-Web: A Large-Scale Benchmark for Retrieval in Agentic RAG Systems | Maximilian Schall, Sedigheh Eslami, Markus Krimmel +4 | cs.IR | 2026-09-08 |
| 3 | Hyperparameter Scaling Laws Across MoE Sparsity | Changxin Tian, Kunlong Chen, Jia Liu +3 | cs.LG | 2026-09-08 |
| #4 | MoEMB: Scaling Universal Multimodal Embeddings with Efficient Mixture-of-Experts Models | Xuanming Cui, Shlok Kumar Mishra, Wentao Bao +6 | cs.LG | 2026-09-08 |
| #5 | From Where to How: Continuous 4D Interaction Forecasting from Egocentric Video | Qiaohui Chu, Haoyu Zhang, Meng Liu +3 | cs.CV | 2026-09-08 |
| #6 | Flexible Spectral-Normalized Neural Gaussian Process for Dynamic Aperture Prediction | Yousra El-Bachir, Frederik Van der Veken, Davide di Croce +4 | physics.acc-ph | 2026-09-08 |
| #7 | Limitations of Automated Simulatability: LLM Simulators Can Bypass Explanations | Antonin Poché, Fanny Jourdan, Nils Feldhus +6 | cs.CL | 2026-09-08 |
| #8 | SignRefine: Adapting Foundational Video Models for Sign Language Generation | Anton Pelykh, Edward Fish, Ozge Mercanoglu Sincan +1 | cs.CV | 2026-09-08 |
| #9 | Environments as Scaffold: Enriching Feedback to Bootstrap Self-Evolving Agents in Long-Horizon Tasks | Hongbang Yuan, Zhuoran Jin, Yixin Cao | cs.LG | 2026-09-08 |
| #10 | LEBGen: An LLM-Enhanced Bayesian Network Framework for Few-Shot Travel Survey Data Generation | Zijian Shen, Bin Zhou, Jiguang Wang +2 | cs.AI | 2026-09-08 |
| #11 | Agentic ML Exploration (A-MLE) for Ads Ranking | Erwin Gao, Vinodh Kumar Sunkara, Jingyi Guan +36 | cs.AI | 2026-09-08 |
| #12 | 3DWay: Generalizing Robot Manipulation via 3D Consistent Waypoints | Ziqin Huang, Yingyue Li, Chenyangguang Zhang +6 | cs.RO | 2026-09-08 |
| #13 | GPU-Enabled Large-Scale Optimization Using Randomized Linear Algebra | Pratik Rathore, Zachary Frangella, Parth Nobel +2 | cs.LG | 2026-09-08 |
| #14 | DriveMotion: A Large-Scale Multi-Source Benchmark for Driver Motion Sequence Modeling and Forecasting | Yuhang Wang, Chuheng Wei, Jingxin Yang +2 | cs.CV | 2026-09-08 |
| #15 | A Machine Learning Framework for Predicting Restaurant Food Waste to Support Sustainable Food Management | Md Mehedi Hasan Naeem, Md Ashraful Islam, Moumita Barua +2 | cs.LG | 2026-09-08 |
| #16 | Popular Knowledge Propagates More Errors in LLM Knowledge Updating | Yuji Zhang, Weibing Wang, Cheng Qian +5 | cs.CL | 2026-09-08 |
| #17 | SGD in Multiclass Logistic Regression: Sequential Learning and Scaling Laws | Konstantinos Christopher Tsiolis, Denny Wu, Christos Thrampoulidis +1 | stat.ML | 2026-09-07 |
| #18 | SAFIRE: Safety-Critical Benchmark for Fine-grained Fire and Smoke Understanding in Multimodal LLMs | Pengfei Li, Naufal Suryanto, Sicheng Zhang +2 | cs.CV | 2026-09-07 |
| #19 | LLM Agents as Computational Typologists | Changbing Yang, Christopher Hammerly, Freda Shi +1 | cs.CL | 2026-09-07 |
| #20 | Large-Scale User Behavior Analysis in Multimodal AI-Assisted Manual Task Execution | Rafael Ferreira, Diogo Tavares, Diogo Glória-Silva +2 | cs.HC | 2026-09-07 |
| #21 | SeisBench DAS: A machine learning framework for Distributed Acoustic Sensing | Jannes Münchmeyer, Han Xiao, Frederik Tilmann | physics.geo-ph | 2026-09-07 |
| #22 | Qwen-Audio-3.0-ASR Technical Report | Chuanmeng Bian, Daren Chen, Peixin Chen +42 | cs.CL | 2026-09-07 |
| #23 | CosmoH2G: A Hand-to-Gripper Transfer Dataset and Baseline Method for Object Manipulation with Complex Spatial Movements | Hongxiang Zhao, Mutian Xu, Zeyu Jin +3 | cs.RO | 2026-09-07 |
| #24 | Latent-to-Latent Flow for Volumetric Stochastic Segmentation | Omar Todd, Sooha Kim, Raghav Mehta +5 | cs.CV | 2026-09-07 |
| #25 | Revisiting Thinning Methods for Kernel Learning Problems | Blanca Cano-Camarero, Yago R. Aguado-Carrillo-de-Albornoz, Ángela Fernández-Pascual +1 | cs.LG | 2026-09-07 |
| #26 | SPARROW: Scalable Taxonomy Induction via Structure-Preserving Partitioning and Constraint-Guided Merging | Yirui Zhang, Yixuan Tang, Yandong Sun +2 | cs.CL | 2026-09-07 |
| #27 | Query-Aware Token Budgeting for Efficient Late-Interaction Visual Document Retrieval | PS Rishi, Rajeev Ranjan Dwivedi, Vinod K Kurmi | cs.IR | 2026-09-07 |
| #28 | Retrieval-Augmented Multi-Prompt Ensemble for Minor-Grain Breeding Information Extraction | Hang Zhao, Jiahao Wang | cs.CL | 2026-09-07 |
| #29 | Online Draft Co-Training for Speculative Decoding in Large-Scale, Long-Context RL Post-Training | Zili Wang, Zhaopeng Qiu, Yuekai Zhang +2 | cs.LG | 2026-09-07 |
| #30 | Beyond Sparse Rewards: A New Benchmark and Structure-Aware Graph Alignment for Micro-Drama Understanding | Yixin Qin, Shi-Zhe Chen, Zhiqi Yu +4 | cs.AI | 2026-09-07 |
| #31 | Single Image to Textured 3D Object Generation in Frequency Domain: From Theory to Pipeline | Qisen Wang, Yifan Zhao, Jia Li | cs.CV | 2026-09-07 |
| #32 | One for All: Generalist Foundation Model for Cross-Sensor Skeleton Representation Learning | Jeonghyeok Do, Yun Chen, Munchurl Kim | cs.CV | 2026-09-07 |
| #33 | DPSF-Net: A Dual-Prior Spatial-Frequency Network for Real-World Remote Sensing Image Dehazing | Mei Lu, Shangliang Shao, Shanliang Yao | cs.CV | 2026-09-07 |
| #34 | Emo-DVS: A Multimodal Benchmark for Privacy-Aware Emotion Recognition with Event Cameras | Jiaqi Chen, Qinfu Xu, Hao Zhuang +1 | cs.CV | 2026-09-07 |
| #35 | Organization of Valence and Arousal in Vision-Language Representations of Built Environments: Insights from the EMOIS Dataset | Madoka Yonekura, Katsunori Kohda, Nobuhiko Muramoto +1 | cs.CV | 2026-09-06 |
| #36 | Generalist Open-World Temporal Perception | Cristian Sminchisescu | cs.CV | 2026-09-06 |
| #37 | AuthBench: A Large-Scale Multilingual Benchmark for Authorship Representation across Genres and Lengths | MaoXun Huang, Zhenxing Zhang, Claire Cardie | cs.CL | 2026-09-06 |
| #38 | A Novel Semantic Manifold Alignment Attack against Embedding-to-Embedding Obfuscation in Privacy-Preserving LLMs | Sicong Li, Lingfeng Yao, Xingke Yang +7 | cs.CR | 2026-09-06 |
| #39 | Back to the Feature: Zero-Shot 6DoF Pose Estimation via Dense Local Features | Ali Rafiaei, Michael Greenspan | cs.CV | 2026-09-06 |
| #40 | DianShi-RxnDB: A Large-Scale, Fine-Grained Organic Reaction Data Platform Built via a Fully Automated Pipeline for Researchers and AI Agents | Yubin Wang, Xingjian Wei, Jiang Wu +33 | cs.CL | 2026-09-06 |
| #41 | A Grapheme-Aware Indic Tokenizer for Tamil: Large-Scale Training and Intrinsic Evaluation | Hari Krishnan K, Sudarsun Santhiappan | cs.CL | 2026-09-06 |
| #42 | VidaForge: Open Research Infrastructure for Video Pretraining Data Recipes | Yan Ma, Jiadi Su, Zhulin Hu +4 | cs.CV | 2026-09-06 |
| #43 | Thinking with Cameras: Active Visual Reasoning via Dynamic Viewpoint Control for Surveillance Video Understanding | Xiao Zhang, Wang Zeng, Sheng Jin +3 | cs.CV | 2026-09-06 |
| #44 | Large-Scale Pretraining for Improving Deep Learning-Based Geometric Distortion Correction of Diffusion-Weighted Imaging | Saroj Khanal, Yashawant Kumar Yadav, Kritam Bhattarai +12 | cs.CV | 2026-09-06 |
| #45 | Scaling 3D Generative Priors to Large-Scale Scene Meshes from Multi-View Images | SangEun Lee, Wonseok Chae, Hoyoung Yoo +3 | cs.CV | 2026-09-06 |
| #46 | Parameterized and Streaming Algorithms for Euclidean Fair $k$-Center Clustering | Zeyu Lin, Chaoqi Jia, Longkun Guo +1 | cs.LG | 2026-09-06 |
| #47 | RoboSPA: Can VLA Models Go Beyond Simple Scenes and Short-Horizon Tasks? | Zhenxuan Fan, Bo Zhang, Yutong Lin +9 | cs.RO | 2026-09-04 |
| #48 | Scalable Detection of Fossil Palynomorphs in Multifocal Digital Microscopy Images | Abbas Shaikh, Praise Mayor, Patrick Ainlay-Vazquez +7 | cs.CV | 2026-09-04 |
| #49 | Don't Drop Dropout: Optimizing Layer Sparsity for Efficient LLM Training and Inference | Mostafa Elhoushi, Alex Pretko, Nolan Dey +6 | cs.AI | 2026-09-04 |
| #50 | Self-Supervised Lexical Representation Learning for Fast, Large-Scale Phylogenetic Inference | Tim Wientzek | cs.CL | 2026-09-04 |