| 1 | Graph Machine: Towards Better Pretraining via Edges | Lintai Hou | cs.LG | 2026-09-02 |
| 2 | Towards Trustworthy Autonomous Robots: An Explainable AI-Based Decision Framework | Cagri Temel | cs.RO | 2026-09-02 |
| 3 | UE5M3 FP4 Block Scaling for Stable Language Model Pretraining | Robert Hu, Carlo Luschi, Paul Balanca | cs.LG | 2026-09-02 |
| #4 | Efficient All-in-One Weather Restoration using Spectral Harmonization | Paula Garrido-Mellado, Daniel Feijoo, Yuning Cui +2 | cs.CV | 2026-09-02 |
| #5 | Trace as State: Reasoning Traces as Conditional States for Long-Context Transformers | Xu Zou, Jie Tang | cs.CL | 2026-09-02 |
| #6 | GaLe: memory-efficient Global Approximate and Local Exact features | Alberto Ancilotto, Elisabetta Farella | cs.CV | 2026-09-02 |
| #7 | oHC: Orthogonal Hyper-Connections on SO(4) via Quaternions | Haoqiang Guo, Xuyi Chen, Bo Ke +5 | cs.CL | 2026-09-02 |
| #8 | UnCapsTSR: An Unsupervised Transformer-based Image Super-Resolution Approach for Capsule Endoscopy Images | Anjali Sarvaiya, Shubh Kawa, Lalit Agrawal +3 | cs.CV | 2026-09-02 |
| #9 | When Decodability Is Not Enough: Logical Validity Representations, Behavioral Dissociation, and Causal Tests in Language Models | Smitha Muthya Sudheendra, Jaideep Srivastava | cs.CL | 2026-09-02 |
| #10 | Uncertainty-Guided Adverse Weather Restoration via Gated Transformer Network | Zheke Jin, Yuning Cui, Tianle Jin +2 | cs.CV | 2026-09-02 |
| #11 | MultiGhostBench: A Multilingual Benchmark for Long-Form LLM-Generated Text Attribution under Distribution Shifts | Matteo Greco, Anudeex Shetty, Andrea Tagarelli +1 | cs.CL | 2026-09-02 |
| #12 | GlyphAnchor: Enhancing Visual Text Rendering via Position-Anchored Glyph Priors | Qiang Xiang, Shuang Sun, Binglei Li +4 | cs.CV | 2026-09-02 |
| #13 | AGI Maze Prediction Datasets: A Compact Benchmark for Learning World Dynamics with Transformers | Alexey Potapov | cs.LG | 2026-09-02 |
| #14 | If It Moves, Radar Knows: A Physics-Aware Radar Transformer for Class-Agnostic Moving-Object Detection | Yinghao Sun, Shuguang Li, Jinliang Shao +1 | cs.CV | 2026-09-02 |
| #15 | SMart: A Multi-source Multi-phase Time Series Representation Transfer Framework | Fang He, Wang-chien Lee | cs.LG | 2026-09-02 |
| #16 | Lightweight Adaptation of General-Purpose VLMs for Multispectral and SAR Image Understanding | Shanji Liu, Kelu Yao, Junxiao Xue +5 | cs.CV | 2026-09-02 |
| #17 | C$^{3}$T: Counterfactual Causal Reasoning for Sentiment Shifts in Social-Media Conversation Trees | S M Rafiuddin, Atriya Sen | cs.CL | 2026-09-02 |
| #18 | XMerge: Cross-Axis Selection and Reconstructive Layer Merging for LLM Depth Compression | Jundong Hu, Shekar Ramachandran | cs.LG | 2026-09-02 |
| #19 | The Dynamics of Continuous Mixture Collapse in Language Models | Ali Backour | cs.LG | 2026-09-02 |
| #20 | Aggregating Neighbor Embedding Projection and Rank-Based Manifold Learning for Image Retrieval | Vinicius Atsushi Sato Kawai, Gustavo Rosseto Leticio, Lucas Pascotti Valem +1 | cs.CV | 2026-09-02 |
| #21 | OR-Transformer: Scaling Real-Time Decision-Making to 1,000 Items | Shuze Daniel Liu, David Simchi-Levi, Claire Chen +2 | cs.LG | 2026-09-01 |
| #22 | Looped Transformers under the Jacobian Lens: Does the Global Workspace Survive Recurrence? | Wenlong Wang, Fergal Reid | cs.AI | 2026-09-01 |
| #23 | Latent unified smooth Hamiltonians for excited state chemistry | David Juergens, Martin Stöhr, Andreas E. Hillers-Bendtsen +2 | physics.chem-ph | 2026-09-01 |
| #24 | Improved Automatic Target Recognition in Synthetic Aperture Sonar Imagery Using Large Deep Neural Networks | C. J. Moore, Alex Hurt, Jordan Malof | cs.CV | 2026-09-01 |
| #25 | Designing Versatile Samples for Learned Trajectory Scoring | Yaguang Li, Jiaru Zhang, Chuheng Wei +2 | cs.RO | 2026-09-01 |
| #26 | CoViT: Instance-Correspondence Contrastive Learning for Vision Transformer | Yisen Wang, Zhirong Wu, Limin Wang | cs.CV | 2026-09-01 |
| #27 | Ten Architectures, One Error: Shared Failure Modes in Hyperspectral Classification under Spatially Disjoint Evaluation | Ehsan Faghih, Fatemeh Ashrafi, Marguerite Moore +1 | cs.CV | 2026-09-01 |
| #28 | Emergence of Fibrations, Compression, and Symmetry Breaking in Artificial Neural Networks | Osvaldo M Velarde, Lucas C Parra, Alireza Hashemi +1 | cs.LG | 2026-09-01 |
| #29 | Swin Meets EfficientNet: Lightweight Architectures for GAN-Based Face Forensics | Sejuti Basu, Ashima Sood, Vijay Kumar +1 | cs.CV | 2026-09-01 |
| #30 | ZipTok3D: High-Fidelity 3D Tokenization with Compact Token Prefixes | Mingda Lin, Weijie Wang, Zeyu Zhang +7 | cs.CV | 2026-09-01 |
| #31 | What, Where, and How: Probing Spatiotemporal Representations in Video Foundation Models | Sharon S. Musa, Fereshteh Forghani, Harrish Thasarathan +3 | cs.CV | 2026-09-01 |
| #32 | Benchmarking Spatial, Spectral, and Self-Supervised Cues for Face Forgery Detection under Realistic Degradation | Lucas Cunha, Lucas Sotomaior, Lucas Gasperin +3 | cs.CV | 2026-09-01 |
| #33 | Does Imitation Learning Preserve Temporal Robustness in Dexterous Manipulation? An Expert-Learner Comparison Across Task Execution Speeds | Clinton Enwerem, John S. Baras, Calin Belta | cs.RO | 2026-09-01 |
| #34 | Learning Sparse Decision Trees via Transformer Variational Auto-Encoders | Giacomo Fidone, Alessio Cascione, Riccardo Guidotti | cs.LG | 2026-09-01 |
| #35 | Semantic-Guided Multimodal Preprocessing for Vision Transformer-Based Clear Cell Renal Cell Carcinoma Grading | Fatemeh Javadian, Zhu Chen, Zahra Aminparast +1 | cs.CV | 2026-09-01 |
| #36 | Measuring consistency via ensemble margin and local prediction variability: Auditing decision systems in the presence of predictive multiplicity | Sinjini Banerjee, Tim Marrinan, Anand D. Sarwate | stat.ML | 2026-09-01 |
| #37 | Multimodal RGB-Infrared Combination for UAV-Based Wildfire Segmentation: A Comparative Study on FLAME3 | Matheus F. Kovaleski, Luís Garrote, Cristiano Premebida +2 | cs.CV | 2026-09-01 |
| #38 | Polish ModernBERT: The Long and Short of Polish Language Understanding | Michał Perełkiewicz, Sławomir Dadas, Rafał Poświata +1 | cs.CL | 2026-09-01 |
| #39 | SMELT: Scaling Laws for Compute-Matched MoE Looped Transformers | Shaowen Wang, Ge Zhang, Kairong Luo +6 | cs.LG | 2026-09-01 |
| #40 | One-Layer Transformer Provably Learns Multiclass One-Nearest Neighbor in Context | Skanda Athreya, Yutong Wang | cs.LG | 2026-09-01 |
| #41 | HiLRP: Toward One Trustworthy Explanation for Vision Transformer: Conservation-Valid Attribution via Attention Primitives | Sathiyamohan Nishankar, Pubudu Sanjeewani, Asanka Perera +1 | cs.CV | 2026-09-01 |
| #42 | Solving In-Table Prediction Problems by Deep Neural Networks with Performance Evaluation Using Synthetic Data | Xiao Zhao, Daniela Oelke | cs.LG | 2026-09-01 |
| #43 | From Language to Behavior: Scaling Sequence Transformers for Industrial Recommendation Ranking with Rec-Native Designs | Jie Chen, Xiangqian Yu, Yanchao Lian +9 | cs.IR | 2026-09-01 |
| #44 | Position Matters: Feature Inversion Attacks in ViT Split Inference with Token Reduction and Shuffling | Stefano Leggio, Giulio Rossolini, Alessandro Biondi | cs.CR | 2026-09-01 |
| #45 | Multi-Head Self Attention is a Parameter Identification Mechanism | W. Ross Morrow | cs.LG | 2026-09-01 |
| #46 | Recent Developments in Transformer Inference Deployment on FPGA Platforms: A Survey | Arjan Blankestijn, Uraz Odyurt, Amirreza Yousefzadeh | cs.LG | 2026-09-01 |
| #47 | Births are difficult to predict even with rich survey and full-population register data | Elizaveta Sivak, Emily M. Cantrell, Thomas Emery +109 | cs.LG | 2026-09-01 |
| #48 | Scaled Idempotence in Transformer Attention: Paired OV Geometry and Shared-Value Algebras | Jiming Feng, Junliang Li | cs.LG | 2026-09-01 |
| #49 | ViTAMINS: An Empirical Study of Training Self-Supervised Vision Transformers with Synthetic Hard Negatives | Nikos Giakoumoglou, Andreas Floros, Kleanthis-Marios Papadopoulos +1 | cs.CV | 2026-09-01 |
| #50 | Low-Quality Face Recognition using Center Aligned Representations and Local Margin Constraints | Vedat Can Dilaver, Benjamin S. Riggan | cs.CV | 2026-09-01 |