| 1 | Learning to Stop without Learning to Stop: Self-Supervised Confidence Training Improves Reasoning Efficiency | Parsa Hosseini, Akasha Tigalappanavara, Sumit Nawathe +6 | cs.AI | 2026-09-25 |
| 2 | BeatGraph: Self-Supervised Heartbeat Graphs for Infant ECG Representations from the Home Environment | Mohammad Nur Hossain Khan, M. S. Krafczyk, Beverly G. Bolster +3 | cs.LG | 2026-09-25 |
| 3 | Forensic Twins: Self-Supervised Residual Learning for AI-Generated Image Forensics | Javier Muñoz-Haro, Ruben Tolosana, Ruben Vera-Rodriguez +2 | cs.CV | 2026-09-25 |
| #4 | KneePreM: Towards 3D Knee MRI Foundation Models via Large-Scale Unlabeled Pretraining and Label-Efficient Fine-Tuning | Xinxin Wang, Liam Hazan, Jing Li +3 | cs.CV | 2026-09-25 |
| #5 | Towards Whole-Study Screening for Congenital Heart Disease in Fetal Ultrasound Using Multiple Instance Learning | Mohamed Azzam, Ruobing Liu, Esther C. Ugwueke +8 | eess.IV | 2026-09-25 |
| #6 | Progressive Memory Transformer: Memory-Aware Attention for Time-Series | Tord Sture Stangeland, Andreas Köhler, Steffen Mæland +1 | cs.LG | 2026-09-25 |
| #7 | Self-Supervised Representation Learning: From Spectral Foundation Models to Auroral Emission Spectra | Matthieu Le Lain, Gaël Cessateur, Sébastien Lefèvre | cs.LG | 2026-09-25 |
| #8 | BAT-CLIP: Trimodal Alignment of Brain, Audio and Text | Suhyun Kim, Jinmo Han, Danny Dongyeop Han +9 | cs.SD | 2026-09-25 |
| #9 | Self-Supervised Perceptually Interpretable Monocular Depth Estimation | Zain Ul Abidin, George Dimas, Dimitris K. Iakovidis | cs.CV | 2026-09-25 |
| #10 | Reliability-Regulated Trajectory Optimization for Progressive COLMAP-Free 3D Gaussian Splatting | Zijian Wu, Jinliang Wang, Zidian Lin +5 | cs.CV | 2026-09-25 |
| #11 | Combining General and Domain-Specific Pretext Tasks for Brain MR Image Segmentation | Tasneem Nasser, Susanne Schmid, Roberto Souza +1 | cs.CV | 2026-09-25 |
| #12 | Structure-Guided Masked Autoencoders for Ultra-High Resolution Scientific Image Understanding | Enzhi Zhang, Du Wu, Rui Zhong +18 | cs.CV | 2026-09-25 |
| #13 | StarWM: Self-Supervised Trained Attention Routing for Robust World Models | Zeqiang Zhang, Fabian Wurzberger, Maximilian Otte +4 | cs.CV | 2026-09-25 |
| #14 | Can Frozen Hyperspherical Features Guide the Selection of Pseudo Masks? | Xinge Guo, Fengyang Xiao, Dingming Zhang +6 | cs.CV | 2026-09-24 |
| #15 | A Native-Reference Phone-Class Geometry for Second-Language Pronunciation Analysis | Tina Raissi, Nhan Phan, Chenxiao Wang +1 | cs.CL | 2026-09-24 |
| #16 | ConPro: Contrast Projection Pretraining for Label-Efficient Vessel Segmentation in DSA Sequences | Xinge Guo, Yuanhao Wang, Liqi Shu +2 | cs.CV | 2026-09-24 |
| #17 | A New Gap Sequence for Shellsort: RL-Driven Algorithm Discovery Beyond $N^{4/3}$ | Bo Liu | cs.CC | 2026-09-24 |
| #18 | RD-JEPA: Predictive latent pretraining for few-trajectory transfer across reaction--diffusion equations | Chenhao Si, Ming Yan | cs.AI | 2026-09-24 |
| #19 | Learning a Flow to Self-Supervised Representations | Yuling Jiao, Wensen Ma, Houduo Qi +1 | cs.CV | 2026-09-24 |
| #20 | Med-AR: Autoregressive Vision-Language Pretraining for Long-Tailed Chest X-Ray Classification and Uncertainty-Aware Evaluation | Janhavi Prabhu, Sahil, Akshay V +3 | cs.CV | 2026-09-24 |
| #21 | Spooftral: Can Voxtral Audio-Language Model Detect Speech Spoofing? | Avishai Weizman, Yehuda Ben-Shimol, Itshak Lapidot | eess.AS | 2026-09-23 |
| #22 | Physics-Informed Self-Supervised Learning for Joint Wire Calibration and Interaction Position Reconstruction in Multi-Wire Parallel Plate Avalanche Counters | Antoine Lemasson, Maurycy Rejmund | cs.LG | 2026-09-23 |
| #23 | Two Global Crops Suffice: Locating Semantic Emergence in DINO-Style Self-Supervised Learning | Basavaraj Sunagad, Artur Jesslen, Adam Kortylewski | cs.CV | 2026-09-23 |
| #24 | A Native-Reference Coordinate Geometry for L2 Pronunciation Deviation Using Self-Supervised Speech Models | Tina Raissi, Nhan Phan, Mikko Kurimo | cs.CL | 2026-09-23 |
| #25 | Prompt, Probe, Train, or Annotate? Single-camera sports video understanding in amateur settings | Sai Varun Kodathala, Prashanth Pollishetty, Jaylen Cargill | cs.CV | 2026-09-23 |
| #26 | Leakage-Safe Machine Learning for Hydrogen Embrittlement Detection in 316L Stainless Steel: A Region-Held-Out Evaluation of Texture and Deep Features in SEM Micrographs | Muhammad Awais, Muhammad Yaseen, Abdul Shakoor +3 | cs.LG | 2026-09-23 |
| #27 | Vision Foundation Models with Synthetic-Only Training for Monocular Spacecraft Pose Estimation | John Church, Vazghen Nikolian | cs.CV | 2026-09-22 |
| #28 | Latent Commonality Expectation-Maximisation for Box-supervised Tree Crown Instance Segmentation | Thomas Pitts, Kunqi Li, Bin Liang | cs.CV | 2026-09-22 |
| #29 | FLINT: Fast Lightweight Inference for Traversability | William Bonilla, Maxime Boisvert, David-Alexandre Poissant +2 | cs.RO | 2026-09-22 |
| #30 | On the Role of the Projector in Contrastive Self-Supervised Learning: Last-Layer Rank Dynamics Drive Representation Quality | Siladittya Manna, Priyangshu Mandal, Umapada Pal +1 | cs.CV | 2026-09-22 |
| #31 | Less Is More in the Long Tail: Stage-Adaptive Sample Selection for Annotation-Efficient Dense Prediction | Xiaofei Du, Lei Zhang, Shuyu Yan +2 | cs.CV | 2026-09-22 |
| #32 | Self-Supervised Combinatorial Optimization with Constraints via Frank-Wolfe | Akbar Rafiey, Yifei Xu, Nikolaos Karalias | cs.LG | 2026-09-22 |
| #33 | Decoupling Disease, Covariates, and Individual Variability: A Unified Disentanglement Framework for Medical Image Classification | Shengjie Zhang, Jinglin Zhang, Zhuangzhuang Jiang +10 | cs.CV | 2026-09-22 |
| #34 | EMGBlend: Heterogeneity-Aware Self-Supervised Pretraining for Gesture and Force Decoding | Yuwei Jia, Cheng Zhong, Jinyang Yu +1 | cs.LG | 2026-09-22 |
| #35 | RootQuantV2: Adapting a Vision Foundation Model for Root-Trait Regression from Minirhizotron Imagery | Kinjalk Parth, Sebastian Varela, Andrew D. B. Leakey | cs.CV | 2026-09-22 |
| #36 | Learning Defensive Policies against Diverse Inference Attacks for Smart Meter Privacy | Ruichang Zhang, Mustafa A. Mustafa | cs.LG | 2026-09-21 |
| #37 | MT-ProtBERT: Multi-task Learning ProtBERT for Intrinsically Disordered Proteins Classification with Scarce Data | Jian Sun, Kingshuk Ghosh, Lilianna Houston +1 | cs.LG | 2026-09-21 |
| #38 | Toward a foundation model for forest point clouds | Yuanwen Yue, Stefano Puliti, Damien Robert +7 | cs.CV | 2026-09-21 |
| #39 | GraphSVR: q-Space--Aware Graph-Based Slice-to-Volume Registration for Diffusion MRI | Noga Kertes, Daphna Link Sourani, Alex M. Bronstein +1 | cs.CV | 2026-09-21 |
| #40 | CMAMBADEPTH: Self-supervised Monocular Depth Estimation with Channel Mamba and Hybrid Attention | Xuezhi Xiang, Jiayao Liu, Heqi Xiang +3 | cs.CV | 2026-09-21 |
| #41 | End-to-end Jordanian dialect speech-to-text self-supervised learning framework | Ali A. Safieh, Ibrahim Abu Alhaol, Rawan Ghnemat | cs.CL | 2026-09-21 |
| #42 | Tactile-JEPA: Topology-Aware Self-Supervised Representation Learning for Distributed Tactile Sensors | Elizaveta Kovtun, Matvey Konovalov, Andrey Sakhovskiy +1 | cs.RO | 2026-09-21 |
| #43 | TReViS: Temporal Repetition Structure Aware Video Synthesis for Self-supervised Repetitive Action Counting | Fanqi Yu, Shengming Ma, Stefano Fiorini +4 | cs.CV | 2026-09-21 |
| #44 | SAFe: Segment-guided Aggregation of Feature Densities for Anomaly-aware Segmentation | Anja Delić, Jurica Runtas, Marin Oršić +2 | cs.CV | 2026-09-21 |
| #45 | Positive Pair Geometry Matters: Optimal Transport for Contrastive Learning of Visual Representations | Akshit Nanda, Shahzad Ahmad, Ram Prasad Padhy | cs.CV | 2026-09-21 |
| #46 | Vision Transformers versus convolutional neural networks for fine-grained orchid genus identification in a species-rich, data-poor flora: a controlled benchmark on the Orchidaceae of New Guinea | Reza Saputra, Diah Harnoni Apriyanti, André Schuiteman +4 | cs.CV | 2026-09-21 |
| #47 | Enhancing Shrimp Disease Detection via Deep Learning and Data Refinement for Resilient Aquaculture | Vinh Canh-Thanh Truong, Hai-Binh Pham, Ngoc Hong Tran | cs.CV | 2026-09-20 |
| #48 | BiView-Touch: Learning Bimanual Tactile Representations by Cross-Hand Completion | Chenxin Liang, Youchen Lai, Chuqiao Lyu +3 | cs.CV | 2026-09-20 |
| #49 | Latent Telepathy: Multi-Robot Communication with Self-Supervised Perceptual Latents | Howard Wang, Han Zheng, Cathy Wu | cs.RO | 2026-09-20 |
| #50 | Enhancing speech representation learning with cross-modal knowledge transfer with HGNN under low resource settings: the case study of Yemba | Yannick Yomie Nzeuhang, Paulin Melatagia Yonta, Marie Tahon | cs.CL | 2026-09-19 |