| 1 | A Common Measure of Communication for Speech Brain-Computer Interfaces | Dulhan Jayalath, Benjamin Ballyk, Oiwi Parker Jones | cs.LG | 2026-09-02 |
| 2 | Discriminative World Models for Web Agents | Kelvin Li, Dhruv Pendharkar, Anish Pahilajani +6 | cs.AI | 2026-09-02 |
| 3 | Graph Machine: Towards Better Pretraining via Edges | Lintai Hou | cs.LG | 2026-09-02 |
| #4 | GRADSOLVE: fast exact gradients for ODE ensembles on GPUs | Alessio Spurio Mancini | cs.MS | 2026-09-02 |
| #5 | Improved Gradient Descent Lower Bounds Beyond Nesterov | Yuhan Ye, Kaizhao Liu | math.OC | 2026-09-02 |
| #6 | The Implications of Linguistic Illegibility for LLM Security | James Mickens | cs.LG | 2026-09-02 |
| #7 | Post-Training Language Models for Gold-Medal Performance in Coding Competitions | Aleksander Ficek, Sean Narenthiran, Mehrzad Samadi +2 | cs.LG | 2026-09-02 |
| #8 | UE5M3 FP4 Block Scaling for Stable Language Model Pretraining | Robert Hu, Carlo Luschi, Paul Balanca | cs.LG | 2026-09-02 |
| #9 | Learning Spectral-Like Mesh-Free Discretisations | Lucas Gerken Starepravo, Henry Broadley, Steven Lind +1 | physics.comp-ph | 2026-09-02 |
| #10 | AI Contextual Measurement for Recovering Individual and Group-Level Effects: Validation Against Survey Measures and an Occupational Application | Wenxin Jiang, Xuyang Wang, Yuxiao Wu | cs.AI | 2026-09-02 |
| #11 | Cliff: Learning Process Rewards from the First Mistake | Peixuan Han, Runhui Wang, Ketan Ramaneti +3 | cs.LG | 2026-09-02 |
| #12 | Dutch Books for Language Models | Isaiah Andrews, Suproteem Sarkar | econ.GN | 2026-09-02 |
| #13 | Full-Model Optimality for Tunable Linear Generative Priors in Compressed Sensing | Zhaoming Li, Paul Hand | stat.ML | 2026-09-02 |
| #14 | CodePoisonRAG: Knowledge Poisoning Attacks on Retrieval-Augmented Code Generation | Varun Gadey, Ziad Marey, Alexandra Dmitrienko | cs.CR | 2026-09-02 |
| #15 | From Reweighting to Rewriting: Unlocking the Intervention Effects of Influential Samples in Training Data Attribution | Yuzhang Luo, Chenpeng Wang, Jianhui Chen +1 | cs.CL | 2026-09-02 |
| #16 | Do Tabular Foundation Models Know Physics? Contamination, Units, and the Deterministic Limit | Wassim Tenachi, Yashar Hezaveh, Laurence Perreault Levasseur +1 | cs.LG | 2026-09-02 |
| #17 | Untangling the Mechanisms of Misleading Context in Medical Question Answering | Robin Linzmayer, Noémie Elhadad | cs.CL | 2026-09-02 |
| #18 | HiPoly: a hierarchical polymer-native AI framework for property prediction and generative design | Ge Sun, Gervasio Zaldivar, Yuan Tian +5 | physics.chem-ph | 2026-09-02 |
| #19 | SPADE: SPaT Attack Detection from the Connected Vehicle's Perspective | James Di Novo, Hany Ragab, Sylvain P. Leblanc | cs.CR | 2026-09-02 |
| #20 | Language Models Can Control Their Own Attention | Namgyu Ho, Huzama Ahmad, Woosung Koh +3 | cs.CL | 2026-09-02 |
| #21 | LoRA-TSD: Tangent-Space Spectral Descent for LoRA via Muon-Style Updates | Dmitrii Andriianov, Andrey Veprikov, Aleksandr Beznosikov | cs.LG | 2026-09-02 |
| #22 | Momentum in large-batch training: Polyak enlarges the critical batch size, Nesterov improves data efficiency | Jia-Nan Wang, Zixun Huang, Kairui Li +1 | stat.ML | 2026-09-02 |
| #23 | Neural operators approximate strongly continuous convex monotone semigroups | Jonas Blessing, Philipp Schmocker, Alessandro Sgarabottolo | math.NA | 2026-09-02 |
| #24 | H3DNAS: Hardware-Aware ONNX-Native 3D Point Cloud Model Compression | Anchit Mulye, Rhythm Baghel, Sujay Kumar Ingle +1 | cs.LG | 2026-09-02 |
| #25 | Eliciting ESG Preferences for Reinforcement Learning-Based Portfolio Optimization | Giovanni Dispoto, Marcello Restelli, Carmine Ventre | q-fin.PM | 2026-09-02 |
| #26 | oHC: Orthogonal Hyper-Connections on SO(4) via Quaternions | Haoqiang Guo, Xuyi Chen, Bo Ke +5 | cs.CL | 2026-09-02 |
| #27 | Dimension Dependent Correlation Gap Bounds under Restricted Independence | Arjun Ramachandra | math.PR | 2026-09-02 |
| #28 | Unfolding the Leech Lattice: Fused Multi-Shell Decoding and VRAM Layouts for 2-Bit LLM Weights | Pier-Jean Malandrino | cs.LG | 2026-09-02 |
| #29 | Loom: Weaving Diagnostic Strands into Free-Text Consensus via Embedding-Space Reweighting | Ron Begleiter, Katya Egert Berg, Gilad Saban +1 | cs.AI | 2026-09-02 |
| #30 | Differentiable Electricity-Market Clearing for Gradient-Based Planning | Luca Mungo, Maarten P. Scholl, Arnau Quera-Bofarull | cs.LG | 2026-09-02 |
| #31 | TaRA: Training-Aware Low-Rank Adaptation Initialization | Taehyeon Kim, Eunhyeok Park | cs.CL | 2026-09-02 |
| #32 | Oracle, will I ever learn? A study of prediction convergence and complementarity across link prediction models | Guillaume Méroué, Fabien Gandon, Pierre Monnin | cs.LG | 2026-09-02 |
| #33 | Scalable Direction-Following TTS via Voice Impression-Guided Pseudo Triplet Construction | Kenichi Fujita, Yusuke Ijima | cs.SD | 2026-09-02 |
| #34 | Source Distribution Estimation by Posterior Averaging | Trung-Dung Hoang, Lisa M. Koch | cs.LG | 2026-09-02 |
| #35 | Learning-Based Reconstruction Attacks on Coordinate-Obfuscated Point Clouds | Mohammad Waquas Usmani, Susmit Shannigrahi, Michael Zink | cs.CR | 2026-09-02 |
| #36 | Online Reinforcement Learning in the Met Office Unified Model through Distributed Model-Agent Coupling | Pritthijit Nath, Sebastian Schemm, Peter Haynes +2 | cs.LG | 2026-09-02 |
| #37 | ProbeMatchDTI: Probe-Driven Multi-Scale Biochemical Pattern Matching for Drug-Target Interaction Prediction | Quan Hao, Mengyue Fan, Zifan Dong +8 | cs.LG | 2026-09-02 |
| #38 | Learn from Whoever Is Right: Answer-Verified Multi-Teacher Distillation for Multi-Domain LLMs | Xixiang He, Xingming Li, Baiqi Wu +4 | cs.LG | 2026-09-02 |
| #39 | TrajMind: Chaining Role-Specialized LoRAs for Fast-and-Slow Collective Trajectory Anomaly Diagnosis | Jiahao Wu, Zhenqun Yang, Chen Jason Zhang +1 | cs.LG | 2026-09-02 |
| #40 | A Comparative Study of Graph Representations for GNN-Based Power Grid Control in L2RPN | Adrian Degenkolb, Qiong Huang, Benjamin Schäfer | cs.LG | 2026-09-02 |
| #41 | Spectral Initialization and Scheduled Graph Smoothness for Uncertain Knowledge Graph Completion | Md Abrar Jahin, Taufikur Rahman Fuad, Jay Pujara +1 | cs.LG | 2026-09-02 |
| #42 | Orthogonal Ensembles and Tested Explanations for Performer-Independent Body-Motion Emotion Recognition | Naoto Nishida, Yoshio Ishiguro | cs.CV | 2026-09-02 |
| #43 | Rethinking the Teacher-Student Framework for Test-Time Adaptation | Damian Sójka, Marc Masana, Bartłomiej Twardowski +1 | cs.LG | 2026-09-02 |
| #44 | Training seeds and model-selection stability in recommender-system evaluation | Juan Manuel Rodriguez, Oleg Lesota, Antonela Tommasel | cs.IR | 2026-09-02 |
| #45 | RINSE: Robust Target-Time Normality Estimation for Zero-Shot Graph Anomaly Detection | Taufikur Rahman Fuad, Md Abrar Jahin, Amir Hussain | cs.LG | 2026-09-02 |
| #46 | DeepAffinity: Long-Term Aspect Preference Prediction in eCommerce using Small Language Models | Yotam Eshel, Guy Hadad, Guy Feigenblat +3 | cs.LG | 2026-09-02 |
| #47 | Scalable Kronecker-Fisher Approximation: Efficient Hessian Analysis for Billion-Parameter Language Models Compression | Viacheslav Yusupov, Daria Cherniuk, Evgeny Frolov | cs.LG | 2026-09-02 |
| #48 | CACTUS: Mask-Guided Semantic Clean-Label Backdoors in Decentralized Federated Learning | Chao Feng, Burkhard Stiller | cs.LG | 2026-09-02 |
| #49 | Towards One-for-All Robustness Across a Continuum of Threat Levels | Zhichao Hou, Xiaorui Liu | cs.LG | 2026-09-02 |
| #50 | When Decodability Is Not Enough: Logical Validity Representations, Behavioral Dissociation, and Causal Tests in Language Models | Smitha Muthya Sudheendra, Jaideep Srivastava | cs.CL | 2026-09-02 |
| #51 | IFW-BLS: Dual-Robust Broad Learning System with Intuitionistic Fuzzy Wave Loss | Mushir Akhtar, M. Tanveer | cs.LG | 2026-09-02 |
| #52 | Coverage, Not Targeting: A Structural Regime in Multi-Turn Agent Credit Assignment | Chenyu Zhou, Qiliang Jiang, Shuning Wu +1 | cs.LG | 2026-09-02 |
| #53 | Evidence for Shared Routing Geometry and Dynamics in Sparse Mixture-of-Experts | Kirill Labzin, Stepan Kulibaba, Artem Dzhalilov +1 | cs.LG | 2026-09-02 |
| #54 | A computational approach to maximum likelihood thresholds for colored Gaussian graphical models | Roser Homs, Olga Kuznetsova, Bernadette J. Stolz | stat.ML | 2026-09-02 |
| #55 | Percolation Dynamics in Optimization : Variance Cascades and Discrete Scale Invariance | Sai Niranjan Ramachandran, Suvrit Sra | cs.LG | 2026-09-02 |
| #56 | Humanoid Safe Stop via Learned Stoppability Value | Junfeng Long, Pieter Abbeel, Koushil Sreenath +3 | cs.RO | 2026-09-02 |
| #57 | AGI Maze Prediction Datasets: A Compact Benchmark for Learning World Dynamics with Transformers | Alexey Potapov | cs.LG | 2026-09-02 |
| #58 | Poisoning Attacks on the PGM-index | Atsuki Sato, Martin Aumüller, Yusuke Matsui | cs.DB | 2026-09-02 |
| #59 | What Is Worth Representing? Representational Empowerment for Continual Model Construction | Fei Dai, Hanqi Zhou, Alison Gopnik +1 | cs.LG | 2026-09-02 |
| #60 | Bayes-Optimal BER and AUC: Estimation and Evaluation of Estimators | Ryota Ushio, Takashi Ishida, Masashi Sugiyama | cs.LG | 2026-09-02 |
| #61 | Improving Evaluation Realism with Inference-Time Compute and Deployment Scaffolds | Axel Ahlqvist, Richard Guan, Juan-Pablo Rivera +6 | cs.AI | 2026-09-02 |
| #62 | SEAL: Reinforcing Global Safety in Mixture-of-Experts through Shared Expert ALignment | Qingyu Meng, Yiwei Zha, Jiahuan Pei +3 | cs.LG | 2026-09-02 |
| #63 | From topology learning to graph generation: A unifying perspective | Xiaowen Dong, Hoi-To Wai, Siheng Chen +2 | stat.ML | 2026-09-02 |
| #64 | Entangled Representations Amplify Collateral Damage in Unlearning | Evžen Wybitul, Tim G. J. Rudner, Christian Schroeder de Witt | cs.LG | 2026-09-02 |
| #65 | Do Large Language Models Capture the Diversity in their Training Data? | Youqi Wu, Farzan Farnia | cs.CL | 2026-09-02 |
| #66 | CAPTURE: Disentangling Preference Drift from Memory Poisoning in Personalized LLM Agents | S M Asif Hossain, Ruksat Khan Shayoni, Md Kishor Morol | cs.LG | 2026-09-02 |
| #67 | Codebook Agent: Amortized Topology Design for LLM Multi-Agent Systems | Jinxi Yu, Yubei Li, Eric Hanchen Jiang +6 | cs.AI | 2026-09-02 |
| #68 | RideSkill: A Hierarchical Algorithm for Generalized Ride Sharing with LLM-Driven Automatic Evolution | Zijian Zhao, Sen Li, Xialiang Tong +1 | cs.MA | 2026-09-02 |
| #69 | LLM-as-a-Judge Is Not an Oracle: Why Self-Improving Agents Need Deterministic Guardrails | Vansh Wahi | cs.AI | 2026-09-02 |
| #70 | Similarity-Aware Personalized Federated Learning in Heterogeneous Environments | Arun Kumar A, Sunil Gupta, Dang Ngyuen +2 | cs.LG | 2026-09-02 |
| #71 | Recursive Value Learning for Long-Horizon Offline Goal-Conditioned RL | Hyeonseong Jeon, Youngwoon Lee | cs.LG | 2026-09-02 |
| #72 | Hardware-Accelerated Instance Segmentation for Resource-Constrained Space Robotics with Criticality Analysis | Siddhant Shete, Hilmi Dogu Kücüker, Udo Frese +1 | cs.RO | 2026-09-02 |
| #73 | Prototype-guided transfer of sparse literature knowledge for electrolyte additive discovery | Weixiang Hong, Hongting Du, Jiayue Tang +4 | physics.chem-ph | 2026-09-02 |
| #74 | SMart: A Multi-source Multi-phase Time Series Representation Transfer Framework | Fang He, Wang-chien Lee | cs.LG | 2026-09-02 |
| #75 | Schrödinger Bridges on Lie Group Manifolds for Probabilistic Intrinsic Generation | Shizhe Zhang, Mingyang Zhao, Lei Ma | stat.ML | 2026-09-02 |
| #76 | Learning the Constitutive Behavior of Materials via Neural Operators and Causal Attention: Case Studies in Plasticity and Damage | Rishabh Arora, Lisa Scheunemann, Tim Brepols +1 | cs.LG | 2026-09-02 |
| #77 | Quantum MeanFlow: single-shot generative sampling on NISQ hardware | Ashish Joshi, Eshaan Mistry, Takahiko Koyama | quant-ph | 2026-09-02 |
| #78 | WeaveMark: Robust and Scalable Multi-bit LLM Watermarking via Coded Payload Spreading | Gang-Hyun Park, Ju-Hyeong Lee, Hee-Youl Kwak +1 | cs.CR | 2026-09-02 |
| #79 | Breadth Beats Depth: Improving GCG-Based Jailbreak Optimization with Breadth-Oriented Suffix Search | Shiliang Xiao, Jingsong Wei, Yuzhi Liang +3 | cs.CL | 2026-09-02 |
| #80 | DMRL: Document-Mediated Reinforcement Learning for Skill Optimization in Advertising Recommendation | Wei Zhang, Hongji Li, Song Sun +4 | cs.LG | 2026-09-02 |
| #81 | GenCAR: Generative Counterfactual Alignment with Risk-Controlled Selection for Out-of-Distribution Recommendation | Qianqian Wang, Yunshan Li, Jiawen Zeng +2 | cs.IR | 2026-09-02 |
| #82 | GeoSPRINT: Geometric Redundancy-Aware Step Pruning for Inference in Diffusion Trajectories | Arpita Joshi | cs.LG | 2026-09-02 |
| #83 | Exact Limits of Random Projections for Preserving Geometry: Distance Recovery, Nearest-Neighbor Rankings, and Covariance Shape in Gaussian Models | Piyush Sao | cs.LG | 2026-09-02 |
| #84 | Online Non-Monotone DR-Submodular Maximization Matching the Offline $0.401$ Factor | Vaneet Aggarwal, Yiyang Lu | cs.LG | 2026-09-02 |
| #85 | A Power Law in Logarithm's Clothing: On the Scalability of Graph-Based Vector Search | Sajad Faghfoor Maghrebi, Navid Eslami, Niv Dayan | cs.DB | 2026-09-02 |
| #86 | SoK: Where Do Flow Labels Come From? Auditing Label Provenance in Encrypted Traffic Benchmarks | Sizhe Huang, Shujie Yang | cs.NI | 2026-09-02 |
| #87 | HyperMC: Multi-Fidelity Hyperparameter Tuning for Stochastic Gradient MCMC | Ming Tan, Xiyun Jiao | stat.ML | 2026-09-02 |
| #88 | Scalable Bayesian Optimization of Composite Functions for Image-Based Inverse Problems in Materials Characterization | Dasol Yoon, Poompol Buathong, Chia-Hao Lee +3 | cs.LG | 2026-09-02 |
| #89 | Disease Burden over Skin Tone: Decomposing the Dermatology-AI Generalization Gap | Nirajan Kunwor, Sanjaya Poudel, Quoc-Huy Trinh +2 | cs.CV | 2026-09-02 |
| #90 | A Computational Comparison of Fourier Spectral Differentiation and Spatial Automatic Differentiation in Periodic Physics-Informed Neural Networks | Xilai Liang, Zhao Zhang | cs.LG | 2026-09-02 |
| #91 | A Unified Rate-Distortion Perspective on Vector, Product, and Scalar Quantization | Xianghong Fang, Wenlong Mou, Yuan Yuan +2 | cs.LG | 2026-09-02 |
| #92 | Federated LoRA Adaptation of BiomedCLIP Across Four International Chest X-Ray Cohorts | Sanjaya Poudel, Nirajan Kunwor, Manish Dhakal +2 | cs.LG | 2026-09-02 |
| #93 | Compositional Spectral Prompts for LLM-based Online Time Series Forecasting | Seungyoon Choi, Hyunchul Kim, Jae-Gil Lee +1 | cs.LG | 2026-09-02 |
| #94 | IDEEA: training-free Input-Dependent stEEring via Activation cluster matching | Zheng Wang, Muchen Li, Renjie Liao +1 | cs.CL | 2026-09-02 |
| #95 | TC-Next: Zero-Shot Multimodal Cyclone Forecasting | Zhe Wang, Sijie Chen, Yiming Luo +2 | cs.LG | 2026-09-02 |
| #96 | XMerge: Cross-Axis Selection and Reconstructive Layer Merging for LLM Depth Compression | Jundong Hu, Shekar Ramachandran | cs.LG | 2026-09-02 |
| #97 | DynG-Diff: A State-Aware Dynamic Guidance Diffusion Framework for Probabilistic Time Series Forecasting | Zhente Zhang, Zhengwei Ni, Wei Fan | cs.LG | 2026-09-02 |
| #98 | DocHop: Benchmarking Out-of-domain Multi-hop Reasoning in Information-Dense Documents | Zhuoran Yu, Le Thien Phuc Nguyen, Jaden Park +5 | cs.AI | 2026-09-02 |
| #99 | The Dynamics of Continuous Mixture Collapse in Language Models | Ali Backour | cs.LG | 2026-09-02 |
| #100 | Act More, Decide Less: Skill-Guided Adaptive Action Chunking for Long-Horizon LLM Agents | Yanting Yang, Can Jin, Jinman Zhao +6 | cs.LG | 2026-09-02 |
| #101 | Source-Free Class Relearning: Diagnosing Forgetting in Class Unlearning | Zahra Dehghani, Pablo Piantanida, Mohammadhadi Shateri | cs.LG | 2026-09-02 |
| #102 | Perceptually Regularized Diffusion Model for Image Super-Resolution | Chuxiangbo Wang, Pavithra Venkatachalapathy, Ying Liang +4 | eess.IV | 2026-09-02 |
| #103 | Train What You Deploy: Closing the MLP Reachability Gap in Low-Rank Clone Distillation | Wenhui Chen, Zhifeng Li, Jie Zhou +5 | cs.LG | 2026-09-02 |
| #104 | Posterior Tempering Explains Variance Inflation in Linear and Generalized Linear Thompson Sampling | Prateek Jaiswal, Debdeep Pati, Anirban Bhattacharya +1 | stat.ML | 2026-09-02 |
| #105 | Linear Fusion MultiDiffusion for Fast Training-Free Spherical Panorama Generation | Akio Hayakawa, Yusuke Mukuta, Tatsuya Harada | cs.CV | 2026-09-02 |
| #106 | CAHR-Net: Condition-Adaptive Hysteresis Reconstruction for Compact and Interpretable Magnetic Core Loss Modeling | Chunye Gong, Cong Yao | cs.LG | 2026-09-02 |
| #107 | Morphology signal in whole slide image foundation models can automatically triage slides | Ayushi Sinha, Shashank Yadav, Benjamin Holmes +9 | cs.CV | 2026-09-02 |
| #108 | A Unified Particle Filter LSTM for Data-Driven Process Simulation | Parvin Malekzadeh, Opher Baron, Dmitry Krass | cs.LG | 2026-09-02 |
| #109 | Post-Training Ternarization of Qwen3-4B Capability, Effective Bit Budget, Storage Compression, and Deployment | Anirudh Malik, M Sparsh Mehra, Poojith Devan | cs.AI | 2026-09-02 |
| #110 | Network-Aware Forecasting on Wireless Access Points | Niloo Bahadori, Swadhin Pradhan, Peiman Amini | cs.NI | 2026-09-02 |
| #111 | FlashKAN: B-Spline KANs via Truncated Power Form | Naveen Mysore | cs.LG | 2026-09-02 |
| #112 | Convergence Theory of Knowledge Distillation in Asynchronous P2P Gossip Learning Network | Lucas Qingyang Fang, Tiyao Liu, Jinhao Jing +4 | cs.LG | 2026-09-01 |
| #113 | On-Policy Distillation Meets Off-Policy GRPO: Training Compact Instruction-Following Rerankers | Vignesh Prabhakar, Jialing Pan, Anil Babu Ankisettipalli | cs.LG | 2026-09-01 |
| #114 | Pushing Forward Multi-Secret-Key Homomorphic Encryption for Private Average Aggregation | Miguel Morona-Mínguez, Fernando Pérez-González, Alberto Pedrouzo-Ulloa | cs.CR | 2026-09-01 |
| #115 | Refining Heuristic-Based Bitcoin Address Clustering with Graph Neural Networks | Hugo Schnoering, Roman Bresson, Michalis Vazirgiannis | cs.LG | 2026-09-01 |
| #116 | Sparse Readout Prism: Explaining Logit-Lens Scores in Features Instead of Tokens | Matteo He, William F. Shen, Xinchi Qiu +1 | cs.CL | 2026-09-01 |
| #117 | OR-Transformer: Scaling Real-Time Decision-Making to 1,000 Items | Shuze Daniel Liu, David Simchi-Levi, Claire Chen +2 | cs.LG | 2026-09-01 |
| #118 | CRISP: Cliff-awaRe Input-adaptive Sparse Prefilling with Structural-Mass-Motivated Routing | Huu Huy Nguyen, Chien Van Nguyen, Franck Dernoncourt +4 | cs.LG | 2026-09-01 |
| #119 | Basin Geometry and Reliable Recall of Dynamical Memories in Reservoir Computing | Ling-Wei Kong, Ying-Cheng Lai | nlin.CD | 2026-09-01 |
| #120 | OutageDiT: A Generative Foundation Model for Power Outage Forecasting and Scenario Simulation | Yunqin Zhu, Feng Qiu, Yao Xie | cs.LG | 2026-09-01 |
| #121 | Latent unified smooth Hamiltonians for excited state chemistry | David Juergens, Martin Stöhr, Andreas E. Hillers-Bendtsen +2 | physics.chem-ph | 2026-09-01 |
| #122 | Import What You Need: Learning When and How to Augment EHR Graphs with External Knowledge | Chen Chen, Mohsen Nayebi Kerdabadi, Dongjie Wang +2 | cs.LG | 2026-09-01 |
| #123 | Interpretable Symptom Vectors for Depression in a Large Language Model | Fangyi Zhu, Ajay Subramanian, Allison Constant +3 | cs.CL | 2026-09-01 |
| #124 | Reinforcement learning to choose optimizers | Martin van der Schelling, Deepesh Toshniwal, Miguel A. Bessa | cs.NE | 2026-09-01 |
| #125 | hLLM: Single Pass Decoding for Generative Reranking | Emil Laftchiev, Prachi Agrawal, Moe Kayali +7 | cs.LG | 2026-09-01 |
| #126 | D-FROST: Decentralized Federated pRompt-tuning via Optimal tranSporT for Non-IID and Imbalanced Data | Quan Minh Nguyen, Hoang M. Ngo, Trong Nghia Hoang +1 | cs.LG | 2026-09-01 |
| #127 | Ten Architectures, One Error: Shared Failure Modes in Hyperspectral Classification under Spatially Disjoint Evaluation | Ehsan Faghih, Fatemeh Ashrafi, Marguerite Moore +1 | cs.CV | 2026-09-01 |
| #128 | Emergence of Fibrations, Compression, and Symmetry Breaking in Artificial Neural Networks | Osvaldo M Velarde, Lucas C Parra, Alireza Hashemi +1 | cs.LG | 2026-09-01 |
| #129 | Toward Explainable and Policy-Aware AI for Carbon Credit Price Prediction: A Research Framework for Emerging Carbon Markets | Summaiya Unnisa Begum, Mohammed Nadeem Ullah, Mohammed Abdul Ghani Khan | cs.LG | 2026-09-01 |
| #130 | Pooling and Drift in Delayed Bandits | Melika Baghi | stat.ML | 2026-09-01 |
| #131 | A Study of Conditional Diffusion Models for Open-Loop Control under Dry Friction and Stiction | Eric Aislan Antonelo | cs.LG | 2026-09-01 |
| #132 | CAT-Flow: Curvature-Adaptive sTeps for Flow Matching | Qinchan Li, Pedro Cisneros-Velarde, Keru Fu +3 | cs.LG | 2026-09-01 |
| #133 | Harness Engineering in LLM Tool Use via Agent-Native Reusable Tool Primitives | Haibo Jin, Suijin Wang, Xucheng Yu +2 | cs.SE | 2026-09-01 |
| #134 | RecKAN: Kolmogorov-Arnold Networks with a Learnable Recursive Polynomial Basis | Amirhosein Azarpour | cs.LG | 2026-09-01 |
| #135 | Hearing the Whispers: Black-Box Membership Inference Attacks on Finetuned TTS Models | Kunlin Cai, Kaiyuan Zhang, Zihang Xiang +4 | cs.CR | 2026-09-01 |
| #136 | Generative Diffusion Surrogates with Analytical Variance Schedule | Patrick Reichherzer, Gianluca Gregori, David N. Hosking +1 | cs.LG | 2026-09-01 |
| #137 | Beyond Scores: Understanding LLM-as-a-Judge Mechanisms in Summarization Evaluation | Himil Vasava, Ming Jiang | cs.CL | 2026-09-01 |
| #138 | Facet-0: A Robotic Foundation Model for Contact-Rich Precise Manipulation | Haoyuan Deng, Haichao Liu, Wenkai Guo +6 | cs.RO | 2026-09-01 |
| #139 | The Structure of Quantization Damage in LLMs: Why the Next Bit Should Be Spent Globally | Jundong Hu, Shekar Ramachandran | cs.LG | 2026-09-01 |
| #140 | Scaling Near-Optimal SFT-RL Annotation Budget Allocation from Small to Large LLMs | Jingtan Wang, Arun Verma, Xiaoqiang Lin +4 | cs.CL | 2026-09-01 |
| #141 | Selective Agent Guidance via Entropy: Learning Autonomous Policies from Imperfect VLM Teachers | Giovanni Bonetta, Matteo Merler, Davide Zago +2 | cs.AI | 2026-09-01 |
| #142 | Gradient-Update Mismatch: Rethinking Conflict-Free Training of Physics-Informed Neural Networks | Jing Xiao, Xinhai Chen, Qinglin Wang +5 | cs.LG | 2026-09-01 |
| #143 | Retrieved but not ranked: surface-form bias in structural retrieval, from mathematics to agent trajectories | Nabira Rashid, Manolis Kellis | cs.LG | 2026-09-01 |
| #144 | Can LLMs Discover Scientific Laws in Real and Parallel Worlds? | Yiming Huang, Ziche Liu, Zhuohang Wu +11 | cs.AI | 2026-09-01 |
| #145 | A Mathematical Theory of Reusable Neural Bases for Network Compression | Binshuai Wang | cs.LG | 2026-09-01 |
| #146 | NashDreamer: Model-Based Reinforcement Learning for Zero-Sum Imperfect-Information Games | Tomáš Holeček, Viliam Lisý | cs.LG | 2026-09-01 |
| #147 | Variable Selection for Feature-Based Newsvendor | Zhaoliang Yuan, Jie Wang | stat.ML | 2026-09-01 |
| #148 | Quantum Sparse Autoencoders for Q-Matrix Estimation in Cognitive Diagnosis | Arif Hassan Zidan, Yi Pan, Bowen Guo +5 | cs.LG | 2026-09-01 |
| #149 | Tri-Band Channel Measurement-Enabled Multi-Layer Digital Twin for Terahertz Wireless Data Centers | Mingjie Zhu, Ziming Yu, Guangjian Wang +1 | cs.LG | 2026-09-01 |
| #150 | Sierpiński--Knopp Wasserstein Distance for Persistence Diagrams and Applications to 2-Wasserstein Approximation | Sebastien Tchitchek, Julien Tierny | cs.CG | 2026-09-01 |
| #151 | LatentPress: Context Compression Beyond Text and Vision | Zhengze Zhou, Hejian Sang | cs.LG | 2026-09-01 |
| #152 | Optimizing Byzantine Node Placement in Decentralized Federated Learning | Edoardo Gabrielli, Gabriele Tolomei | cs.LG | 2026-09-01 |
| #153 | Rethinking Learnability in Offline Data-driven Optimization | Chao Qian, Chen-Guang Wang, Rong-Xi Tan +1 | cs.LG | 2026-09-01 |
| #154 | FairLens: Benchmarking Fairness in Vision-Language Models for High-Stakes Decision-Making | Vahid Reza Khazaie, Ahmed Y. Radwan, Shaina Raza | cs.CV | 2026-09-01 |
| #155 | Does Imitation Learning Preserve Temporal Robustness in Dexterous Manipulation? An Expert-Learner Comparison Across Task Execution Speeds | Clinton Enwerem, John S. Baras, Calin Belta | cs.RO | 2026-09-01 |
| #156 | Diffusion as a Training Curriculum for Timestep-Free Iterative Reasoning | Mariia Drozdova, Aidan Sirbu, Pietro Miotti +4 | cs.LG | 2026-09-01 |
| #157 | Edge-Girth as a Structural Edge Feature for Graph Neural Networks | Lilian Marey, Charlotte Laclau | cs.LG | 2026-09-01 |
| #158 | Efficiently Estimating Optimal Hyperparameter Scaling Laws through Power-Law Entropy Search | Zhiliang Chen, Sebastian Ament, David Eriksson +3 | cs.LG | 2026-09-01 |
| #159 | Learning Sparse Decision Trees via Transformer Variational Auto-Encoders | Giacomo Fidone, Alessio Cascione, Riccardo Guidotti | cs.LG | 2026-09-01 |
| #160 | TRIAGE: Three-level Routing and Intelligent Agent Guidance for Efficient Execution | Ruocan Wei | cs.LG | 2026-09-01 |
| #161 | Semantic-Guided Multimodal Preprocessing for Vision Transformer-Based Clear Cell Renal Cell Carcinoma Grading | Fatemeh Javadian, Zhu Chen, Zahra Aminparast +1 | cs.CV | 2026-09-01 |
| #162 | Median-of-Means as an Extremal Convex Estimator and a Nonconvex Route to the Trimmed Oracle | Angshul Majumdar | cs.LG | 2026-09-01 |
| #163 | CATeye: Coupled Attribute-Topology Invariance Learning for Voucher Abuse Detection | Tian Tian, Shuaicheng Niu, Hao Kuang +3 | cs.LG | 2026-09-01 |
| #164 | Provably Safe Sim-to-Real Transfer | Tingting Ni, Maryam Kamgarpour | cs.LG | 2026-09-01 |
| #165 | Predicting Subsurface Abnormalities Growth using Physics-Informed Neural Networks | Mehrdad Shafiei Dizaji, Hoda Azari | cs.LG | 2026-09-01 |
| #166 | On the Reliability of Generative Augmentation: A Wasserstein-Based Theoretical and Empirical Study | Chathurika S Abeykoon, Mathias Nthiani Muia, Mallory Goldstein | stat.ML | 2026-09-01 |
| #167 | Contribution-Aware Bandwidth Allocation for Multimodal Split Learning | Iason Ofeidis, Leandros Tassiulas | cs.LG | 2026-09-01 |
| #168 | Measuring consistency via ensemble margin and local prediction variability: Auditing decision systems in the presence of predictive multiplicity | Sinjini Banerjee, Tim Marrinan, Anand D. Sarwate | stat.ML | 2026-09-01 |
| #169 | Investigating Linear Probe Robustness to Linguistic Register, Medical Specialty, and Corpus Shifts in Medical QA | Nishant Mishra, Ameen Abu-Hanna, Iacer Calixto | cs.CL | 2026-09-01 |
| #170 | Exact Risk-Complexity Laws for Projective Boundaries in Scenario Optimization and Distribution-Free Certification | Giuseppe C. Calafiore | eess.SY | 2026-09-01 |
| #171 | Where the Verifier Fails: A Category-Level Audit of Reward Signals in RLVR | Esther Xin | cs.CL | 2026-09-01 |
| #172 | Cheap Verifiers, Large Blind Spots: Measuring the Reliability Cost of Cost-Saving Cascades | Dushyant Rajput | cs.AI | 2026-09-01 |
| #173 | SMELT: Scaling Laws for Compute-Matched MoE Looped Transformers | Shaowen Wang, Ge Zhang, Kairong Luo +6 | cs.LG | 2026-09-01 |
| #174 | mzCache: On-Device LLM Memory Management under Multitasking | Hongseung Yu, Minsung Kim, Jongseok Park +1 | cs.OS | 2026-09-01 |
| #175 | Bandits in Prod: Hyperparameter Optimization at Inference Time | Louis Abraham, Tuan-Anh Nguyen, Nicolas Devatine | cs.LG | 2026-09-01 |
| #176 | Exploring Sparse Autoencoders in Text-Based Causal Confounding Adjustment | Mian Zhong, Katherine A. Keith, Anjalie Field | cs.CL | 2026-09-01 |
| #177 | Matched Queries for Curvature and Density at Branching Junctions | Ziqi Zhao, Qingjian Ni | stat.ML | 2026-09-01 |
| #178 | MIDR: Enrichment-Augmented Indexing for Multimodal Document Retrieval | Debanjan Mahata, Atharva Tendle, Daniel Preotiuc-Pietro +2 | cs.IR | 2026-09-01 |
| #179 | One-Layer Transformer Provably Learns Multiclass One-Nearest Neighbor in Context | Skanda Athreya, Yutong Wang | cs.LG | 2026-09-01 |
| #180 | GazeRefine: Expert Gaze as a Test-Time Prompt for Training-Free Medical Image Segmentation | Mohammed Oussama Benyahia, Marouane Tliba, Mohamed Amine Kerkouri +10 | eess.IV | 2026-09-01 |
| #181 | Relational Task Generation Language: A Declarative Specification Framework for Relational Deep Learning | Oleksii Kolesnichenko, Jakub Peleška, Gustav Šír | cs.PL | 2026-09-01 |
| #182 | The Constitutional Coverage Trilemma in AI Governance | Natalija Mitic, Soona Sedahmed A. O., Mamadou Selly Ly +1 | cs.LG | 2026-09-01 |
| #183 | Position: Privacy Is a Claim, Not a Property of Synthetic Data | Jiachen Zhao, Antonia Januszewicz, Taeho Jung | cs.LG | 2026-09-01 |
| #184 | FORGE: Forward-Only Test-Time Adaptation for Integer-Only Vision Models on Microcontrollers | Muhammad Rehan, Haider Ali, Muhammad Ali Munir +1 | cs.CV | 2026-09-01 |
| #185 | Solving In-Table Prediction Problems by Deep Neural Networks with Performance Evaluation Using Synthetic Data | Xiao Zhao, Daniela Oelke | cs.LG | 2026-09-01 |
| #186 | Explore More, Drift Less: Outcome-Only Reinforcement Learning Can Suffice for Long-Horizon Interactive Agents | Liming Pu, Xiaoxia Li, Yifu Liu +2 | cs.LG | 2026-09-01 |
| #187 | Post-Training Science for Supervised Fine-Tuning | Charles O'Neill, Mudith Jayasekara, Harry Partridge | cs.LG | 2026-09-01 |
| #188 | From Language to Behavior: Scaling Sequence Transformers for Industrial Recommendation Ranking with Rec-Native Designs | Jie Chen, Xiangqian Yu, Yanchao Lian +9 | cs.IR | 2026-09-01 |
| #189 | Multi-Head Self Attention is a Parameter Identification Mechanism | W. Ross Morrow | cs.LG | 2026-09-01 |
| #190 | REFACTOR-VLA: Unsupervised Library Learning of Typed Motor Programs | Riyaaz Shaik, Chandru Venkataraman | cs.LG | 2026-09-01 |
| #191 | Recent Developments in Transformer Inference Deployment on FPGA Platforms: A Survey | Arjan Blankestijn, Uraz Odyurt, Amirreza Yousefzadeh | cs.LG | 2026-09-01 |
| #192 | Births are difficult to predict even with rich survey and full-population register data | Elizaveta Sivak, Emily M. Cantrell, Thomas Emery +109 | cs.LG | 2026-09-01 |
| #193 | Pre-carved Niches: The Formation Dynamics of Modular Task Partitions in Early LLM Training | Guangqi Li, Yongxin Li | cs.LG | 2026-09-01 |
| #194 | CopyShield: A Cross-Level Benchmark of Copyright Defenses in LLMs | Maryam Alshehyari, Dushyant Singh Chauhan, Samuele Poppi +3 | cs.LG | 2026-09-01 |
| #195 | Superposed Latent Autoencoder | Quanling Zhao, Jiaying Yang, Tianqi Zhang +4 | cs.LG | 2026-09-01 |
| #196 | Reinforcement Learning and Rule-Based Peer-to-Peer Pricing in Residential PV-BES Communities | Pablo Benalcazar, Maciej Kalka, Wilian Guamán +1 | cs.LG | 2026-09-01 |
| #197 | Scaled Idempotence in Transformer Attention: Paired OV Geometry and Shared-Value Algebras | Jiming Feng, Junliang Li | cs.LG | 2026-09-01 |
| #198 | When Does Online Adaptation Pay on the Edge? A Leakage-Free Evaluation of Warmup, Learning-Rate Selection, and Resource Trade-offs for Time-Series Forecasting | Takumi Fujimoto, Hiroaki Nishi | cs.LG | 2026-09-01 |
| #199 | Replicating TRACE: A Practitioner's Guide to Its Threshold and Particle Budget | Alex Chadyuk, Alicia Zhang, Roy Kucukates | cs.LG | 2026-09-01 |
| #200 | Neural Symbollic Regression Using Deep Learning and Sparse Modelling | Ravi Kumar U, Sumitra S | cs.LG | 2026-09-01 |