| 1 | Zero-Shot Novel Depth Synthesis Using 3D Foundation Models Scene Representations | Denis M. Akola, David F. Fouhey | cs.CV | 2026-09-03 |
| 2 | Translation as a Decision Space: A Multi-Agent Perspective on Low-Resource Dialect Generation | Hasan Alkhder, Mohammad Abboush, Igor Tchappi +2 | cs.CL | 2026-09-03 |
| 3 | LLM4CKD: Large Language Models for Early Stage Chronic Kidney Disease Screening | Muhammad Ashad Kabir, Sirajam Munira | cs.AI | 2026-09-03 |
| #4 | IchthyoNoma: Nomenclature and Context Sensitivity of Zero-Shot Biological Vision--Language Models for Bangladeshi Freshwater Fish Recognition | Nazim-E-Alam, Tarek Rahman, Md Kishor Morol | cs.CV | 2026-09-03 |
| #5 | VI3: Grounding Pretrained 3D Foundation Models with Inertial Cues | Ernesto Lozano, Alberto Jaenal, Javier Civera | cs.CV | 2026-09-03 |
| #6 | Typological Feature Prediction with Large Language Models: An In-Context Learning Approach | Qianwen Wang, York Hay Ng, Aditya Khan +1 | cs.CL | 2026-09-03 |
| #7 | Synthetic Semantic Supervision for Contrastive Code Representation Learning in Small Transformers: An Empirical Study | Kenneth Paulsen, Florian Tambon, Mike Papadakis +1 | cs.AI | 2026-09-03 |
| #8 | SignSeek: Learning Transferable Representations for Sign Dictionary Retrieval | Sobhan Asasi, Ozge Mercanoglu Sincan, Richard Bowden | cs.CV | 2026-09-03 |
| #9 | ARCOS: Zero-shot Boundary Localization for Corneal Layer Segmentation Across Optical Coherence Tomography Devices | Nuno Vivas Brás, Benjamin Memmi, Maëlle Bouhassane +4 | cs.CV | 2026-09-03 |
| #10 | Out-of-Distribution Generalisation with Sequence Models in Offline Multi-Agent Reinforcement Learning | Oussama Hidaoui, Omer Ebead, Ulrich Armel Mbou Sob +14 | cs.LG | 2026-09-03 |
| #11 | SV-WAM: An Efficient Surround-View World-Action Model for End-to-End Autonomous Driving | Jinyang Wang, Shiwei Li, Junjian Wang +12 | cs.CV | 2026-09-03 |
| #12 | KhatianDoc: A Human-Verified Benchmark Diagnosing Multimodal LLM Failure on Bengali Legal Land Records | Tasmiad Hasan, Arafat Zaman Ratul, Sarker Sadman Saalim +3 | cs.CL | 2026-09-03 |
| #13 | An Adversarial Zero-Shot Learning Approach for Anomaly Detection in Multivariate IoT Traffic Data | Mahshid Rezakhani, Tolunay Seyfi, Fatemeh Afghah | cs.LG | 2026-09-03 |
| #14 | To What Extent Do Large Language Models Understand Bangla Idioms? | Mousumi Akter, Md. Faiyaz Abdullah Sayeedi, Nurul Labib Sayeedi +1 | cs.CL | 2026-09-03 |
| #15 | Exploring the Potential of Contrastive Language-Image Pre-training for Multi-Source Remote Sensing Data | Xiangyang Miao, Kelu Yao, Yekai Huang +7 | cs.CV | 2026-09-03 |
| #16 | ALRA: Adaptive Local Relational Alignment for Logit-Based Pre-training Distillation of Autoregressive Language Models | Quang Hoang Trung, Quang Huu Hieu, Nguyen Van Hoang Phuc +1 | stat.ML | 2026-09-03 |
| #17 | P-CORE: Self-Supervised Surface Consistency for Point-Based Neural Editing | Yanshu Zhang, Shichong Peng, Mehran Aghabozorgi +2 | cs.CV | 2026-09-03 |
| #18 | Contextual Tamil Spelling and Grammar Correction Using Progressively Fine-Tuned Sequence-to-Sequence Transformers | Karthikeyan A, Jaya Nirmala S, Sangeetha Sivanesan +4 | cs.CL | 2026-09-03 |
| #19 | Frontier LLMs are effective batch optimizers: Assessing reasoning models in continuous and discrete settings | Frank Hu, Shriram Chennakesavalu, David Graff | cs.LG | 2026-09-02 |
| #20 | Solving the Needle-in-a-Haystack Problem in Mammography Vision-Language Model with Differentiable Subset Sampling | Young Seok Jeon, Beatrice Brown-Mulry, Rohan Satya Isaac +5 | cs.CV | 2026-09-02 |
| #21 | Learnable composition for neural operators | Zituo Chen, Baiming Zhang, Sili Deng | cs.LG | 2026-09-02 |
| #22 | SHELF: A Synthetic Harness for Multi-Task Bibliographic Benchmarking | Michael J. Bommarito | cs.CL | 2026-09-02 |
| #23 | Language Models Can Control Their Own Attention | Namgyu Ho, Huzama Ahmad, Woosung Koh +3 | cs.CL | 2026-09-02 |
| #24 | Choosing a PEFT Variant for Per-Patient Dysarthric ASR: A Single-Speaker Case Study on Two ASR Bases | Bernard Muller, László Tóth, LaVonne Roberts | cs.CL | 2026-09-02 |
| #25 | MV-dVRK: A Multi-Viewpoint Benchmark for Spatial Surgical Perception | Guido Caccianiga, Sergey Prokudin, Yutong Chen +9 | cs.CV | 2026-09-02 |
| #26 | H3DNAS: Hardware-Aware ONNX-Native 3D Point Cloud Model Compression | Anchit Mulye, Rhythm Baghel, Sujay Kumar Ingle +1 | cs.LG | 2026-09-02 |
| #27 | Modern Transformers Are Implicit Hybrids: From Functional Differentiation to Principled Hybrid Architecture Design | Runlin Shi, Bojian Yin, Guoqi Li | cs.LG | 2026-09-02 |
| #28 | Equation Recast for Canonical Operator Learning Across Parametric PDEs | Qiyun Cheng, Valentin Duruisseaux, Cesar F. Clauser +7 | cs.LG | 2026-09-02 |
| #29 | RINSE: Robust Target-Time Normality Estimation for Zero-Shot Graph Anomaly Detection | Taufikur Rahman Fuad, Md Abrar Jahin, Amir Hussain | cs.LG | 2026-09-02 |
| #30 | Debias-SparseGPT: Bias-Aware Pruning for Large Language Models | Irina Proskurina, Guillaume Metzler, Antoine Gourru +1 | cs.CL | 2026-09-02 |
| #31 | Adapting a Foundation Model for Lunar Surface Height Estimation | Patrick Bauer, Marius Schwinning, Melanie Siegel +2 | cs.CV | 2026-09-02 |
| #32 | NE-R1: Enhancing Named Entity Recognition Model via Reinforcement Learning | Meixuan Chen, Hehan Li, Ruizhi Zhao +8 | cs.CL | 2026-09-02 |
| #33 | Structured-Prior-Guided Diffusion Inpainting with Physical Consistency for Traffic Sign Augmentation | Luo Li, Chongchong Huang, Jun Jia +4 | cs.CV | 2026-09-02 |
| #34 | SonicCaps: Large-Scale Diverse and Fine-Grained Captioning for Improved Audio-Retrieval | Zineb Lahrichi, Marc Ferras, Gaël Richard +1 | cs.SD | 2026-09-02 |
| #35 | Towards Zero-Shot Transfer Across Embodiments For Driving VLAs | Caio Azevedo, Stefano Sabatini, Sascha Hornauer +1 | cs.CV | 2026-09-02 |
| #36 | SCX Router: Streaming Zero-Shot Model Selection with a Decoder-KV Classifier and a Real-World Task Ontology | Ihor Stepanov, Aleksandr Smechov, Mykhailo Shtopko +2 | cs.AI | 2026-09-02 |
| #37 | CAPTURE: Disentangling Preference Drift from Memory Poisoning in Personalized LLM Agents | S M Asif Hossain, Ruksat Khan Shayoni, Md Kishor Morol | cs.LG | 2026-09-02 |
| #38 | Signal or Noise? Auditing Rotation-Induced Saliency Drift in Medical and Aerial Imaging | Khawaja Murad ul Hassan, Mehran Ebrahimi | cs.CV | 2026-09-02 |
| #39 | Federated LoRA Adaptation of BiomedCLIP Across Four International Chest X-Ray Cohorts | Sanjaya Poudel, Nirajan Kunwor, Manish Dhakal +2 | cs.LG | 2026-09-02 |
| #40 | TC-Next: Zero-Shot Multimodal Cyclone Forecasting | Zhe Wang, Sijie Chen, Yiming Luo +2 | cs.LG | 2026-09-02 |
| #41 | XMerge: Cross-Axis Selection and Reconstructive Layer Merging for LLM Depth Compression | Jundong Hu, Shekar Ramachandran | cs.LG | 2026-09-02 |
| #42 | DPA: Decoupling Product-Agnostic Anomaly Representations for Zero-shot Anomaly Generation | Hang Yao, Yansheng Fu, Ming Liu +4 | cs.CV | 2026-09-02 |
| #43 | Morphology signal in whole slide image foundation models can automatically triage slides | Ayushi Sinha, Shashank Yadav, Benjamin Holmes +9 | cs.CV | 2026-09-02 |
| #44 | Benchmarking Language Models for Statistical Problem Formulation | Chen Wang, Junzhe Zhao, Xin Cong +2 | cs.AI | 2026-09-02 |
| #45 | Data-Efficient Networks for Multi-Contrast MRI Reconstruction based on a Generalized Content/Style Prior | Chinmay Rao, Efe Ilıcak, Matthias J. P. van Osch +5 | eess.IV | 2026-09-02 |
| #46 | OutageDiT: A Generative Foundation Model for Power Outage Forecasting and Scenario Simulation | Yunqin Zhu, Feng Qiu, Yao Xie | cs.LG | 2026-09-01 |
| #47 | TalkFa: A Unified Benchmark for Farsi Dialogue Generation and Understanding | Neda Jamshidi, Kamyar Zeinalipour, Fahimeh Akbari +3 | cs.CL | 2026-09-01 |
| #48 | DESA-TTA: Dynamic EMA and Source Anchoring for Test-Time Adaptation | Atif Belal, Lilian Hollard, Marco Pedersoli +1 | cs.CV | 2026-09-01 |
| #49 | AlphaRAD: Grounded Zero-Shot Classification in Chest Radiology via $α$-Corrected Binary Cross Entropy and Factorized Latent Supervision | Jianzhong You, Yuan Gao, Chris McIntosh | cs.CV | 2026-09-01 |
| #50 | SpeakPay: Domain-Adaptive LoRA Fine-Tuning of Whisper for Low-Resource Nepali Financial Speech Recognition | Biraj Subedi | cs.CL | 2026-09-01 |