| 1 | Measurement-Driven Sub-Network Selection for On-Premise Retrieval-Augmented Factory Agents | Vasileios Rizeakos, Georgios Paisios, Alexandros Machairas +2 | cs.AI | 2026-09-02 |
| 2 | H3DNAS: Hardware-Aware ONNX-Native 3D Point Cloud Model Compression | Anchit Mulye, Rhythm Baghel, Sujay Kumar Ingle +1 | cs.LG | 2026-09-02 |
| 3 | Debias-SparseGPT: Bias-Aware Pruning for Large Language Models | Irina Proskurina, Guillaume Metzler, Antoine Gourru +1 | cs.CL | 2026-09-02 |
| #4 | Scalable Kronecker-Fisher Approximation: Efficient Hessian Analysis for Billion-Parameter Language Models Compression | Viacheslav Yusupov, Daria Cherniuk, Evgeny Frolov | cs.LG | 2026-09-02 |
| #5 | VoRTeC: Taming Foundation Flow for One-step Real time Video Compression | Yichong Xia, Qinhong Wu, Qinhong Wu +3 | cs.CV | 2026-09-02 |
| #6 | CC-4DGS: Computational Deformation and Point-Cloud Compression for Storage-Efficient Dynamic Gaussian Splatting | Kyungdae Park, Chae Eun Rhee | cs.CV | 2026-09-02 |
| #7 | A Unified Rate-Distortion Perspective on Vector, Product, and Scalar Quantization | Xianghong Fang, Wenlong Mou, Yuan Yuan +2 | cs.LG | 2026-09-02 |
| #8 | XMerge: Cross-Axis Selection and Reconstructive Layer Merging for LLM Depth Compression | Jundong Hu, Shekar Ramachandran | cs.LG | 2026-09-02 |
| #9 | Train What You Deploy: Closing the MLP Reachability Gap in Low-Rank Clone Distillation | Wenhui Chen, Zhifeng Li, Jie Zhou +5 | cs.LG | 2026-09-02 |
| #10 | Post-Training Ternarization of Qwen3-4B Capability, Effective Bit Budget, Storage Compression, and Deployment | Anirudh Malik, M Sparsh Mehra, Poojith Devan | cs.AI | 2026-09-02 |
| #11 | When Does Information Sharing Improve Decentralized Discovery? Aggregation, Independent Rescue, and Equilibrium Selection | Yohei Nakajima | cs.AI | 2026-09-01 |
| #12 | Emergence of Fibrations, Compression, and Symmetry Breaking in Artificial Neural Networks | Osvaldo M Velarde, Lucas C Parra, Alireza Hashemi +1 | cs.LG | 2026-09-01 |
| #13 | A Mathematical Theory of Reusable Neural Bases for Network Compression | Binshuai Wang | cs.LG | 2026-09-01 |
| #14 | Can LLMs Design Video Coding Tools? A Case Study on Planar Mode | Yingwen Zhang, Meng Wang, Liqiang He +1 | cs.MM | 2026-09-01 |
| #15 | Benchmarking Spatial, Spectral, and Self-Supervised Cues for Face Forgery Detection under Realistic Degradation | Lucas Cunha, Lucas Sotomaior, Lucas Gasperin +3 | cs.CV | 2026-09-01 |
| #16 | LatentPress: Context Compression Beyond Text and Vision | Zhengze Zhou, Hejian Sang | cs.LG | 2026-09-01 |
| #17 | Contribution-Aware Bandwidth Allocation for Multimodal Split Learning | Iason Ofeidis, Leandros Tassiulas | cs.LG | 2026-09-01 |
| #18 | Exact Risk-Complexity Laws for Projective Boundaries in Scenario Optimization and Distribution-Free Certification | Giuseppe C. Calafiore | eess.SY | 2026-09-01 |
| #19 | One Prompt Is Enough: Watermark Laundering Through Foundation Image Models | Jidong Yang, Qi Li, Wei Zong +5 | cs.CV | 2026-09-01 |
| #20 | Superposed Latent Autoencoder | Quanling Zhao, Jiaying Yang, Tianqi Zhang +4 | cs.LG | 2026-09-01 |
| #21 | On the Design Fundamentals of Pixel Text Representation Learning | Chaohao Yuan, Ruifeng Yuan, Zhuoxu Huang +4 | cs.CV | 2026-09-01 |
| #22 | ClinTraceBench: Source-Verifiable Longitudinal Clinical Reasoning over EHR-Derived Dialogues | Huimin Wang, Zhengyi Zhao, Yutian Zhao | cs.CL | 2026-09-01 |
| #23 | Stochastic Optimization of Tree Tensor Networks | Marius Willner, Maximilian Scharf, André Uschmajew +2 | math.OC | 2026-09-01 |
| #24 | MemoryWalker: Stop Training Agents on Contexts They Never Saw | Zinco J, Xunjie Zhu, Shen Huang +3 | cs.LG | 2026-09-01 |
| #25 | Can Large Language Models Forecast What Researchers Study Next? | Fenghai Li, Zihan Tang, Haofei Yu +2 | cs.CL | 2026-09-01 |
| #26 | A Closed-Loop Evaluation of Capability Loss and Recovery in Compressed Driving Policies | Ahmad Alfan Alfian Irfan, Nur Ahmad Khatim, Mansur Arief | cs.AI | 2026-09-01 |
| #27 | Residual Sparsification via Output Importance for Compressing Mixture-of-Experts LLMs | Seungwoo Jung, Dohyeok Kwon, Seungmin Cha +4 | cs.AI | 2026-09-01 |
| #28 | Manifold-Aware General Coded Computing for Straggler-Resilient Distributed Computing | Parsa Moradi, Mohammad Ali Maddah-Ali | cs.LG | 2026-09-01 |
| #29 | A hybrid quantum-classical neural network for learning to route | Marcus Rolf Peter Ritt, Alexsandro Santos da Rosa Júnior, Marcos Vinicius Reballo +2 | cs.LG | 2026-08-31 |
| #30 | Sharp Approximation Rates for Neural Networks with Affine Latent Parameterizations | Shijun Zhang | cs.LG | 2026-08-31 |
| #31 | Good Memory Has ECC: Evaluating the Memory of Vision-Language Models Beyond Accuracy | Shmuel Berman, Jia Deng | cs.LG | 2026-08-31 |
| #32 | Every Token Leaves a Ripple in the Stream of Thought: Eliciting Model-Internal Token Saliency for Chain-of-Thought Compression | Tianyi Zhao, Yinhan He, Wendy Zheng +1 | cs.CL | 2026-08-31 |
| #33 | Measure Before You Manage: Evaluating Agent Working Memory in Coding Agents | Le Chen, Zishen Wan, Baixi Sun +6 | cs.AI | 2026-08-31 |
| #34 | Faithfulness Is Not Free: Auditing Offline KV-Cache Quantization in Retrieval-Augmented Generation | Atta Ul Asad, Ahsan Bilal, Muhammad Ali +2 | cs.CL | 2026-08-31 |
| #35 | LightNav-0: Eliciting VLM Spatial Intelligence for Generalist Embodied Navigation | Shaoan Wang, Aocheng Luo, Fei Huang +17 | cs.RO | 2026-08-31 |
| #36 | VCAR: Training-Free 3DGS Segmentation via View Completeness and Axis-Aware Boundary Refinement | Kun Cao, Di Wang, Haibin Zhu +4 | cs.CV | 2026-08-31 |
| #37 | TopoCompress: Long Context Compression via Graph-Wired Semantic Trajectories | Daniel Agyei Asante, Yang Li | cs.CL | 2026-08-31 |
| #38 | SkillZip Pro: Execution-Aware Dynamic Compression of Progressively Loaded Skills for Self-Evolving Agents | Xiaofan Bai, Chao Liu, Hongqiang Lin +5 | cs.AI | 2026-08-31 |
| #39 | Which Rules Matter Now? Policy-Centroid Routing Before an Intelligent System Acts | Thomson D. Nguy | cs.AI | 2026-08-31 |
| #40 | Functional Degeneracy in Neural Networks: Measurement and Pruning | Maria Matveev, Pascal Esser, Ayush Bharadwaj +2 | cs.LG | 2026-08-31 |
| #41 | What It Costs to Compose, Rebuild, and Correct Precomputed Memory | Asa Shepard | cs.CL | 2026-08-31 |
| #42 | Tensor Methods for Language Models: From Token Representation to Training, Adaptation, Inference, Compression, and Interpretability | Matvei Tarasov, Salman Ahmadi-Asl, Andre L. F. de Almeida +1 | cs.LG | 2026-08-31 |
| #43 | DASC: Decay-Aware State Compression for Hybrid Linear-Attention Serving | Yanqi Yu, Pingwei Sun, Jianchao Tan +4 | cs.LG | 2026-08-31 |
| #44 | Tail-Replay: Escaping the Curse of Linear Attention in Prefix Caching for Hybrid LLMs | Yirui Liu, Ruoling Qi, Xuaner Wu +2 | cs.LG | 2026-08-31 |
| #45 | Multivariate Scientific Data Compression with Learned Cross-Variable Latent Decorrelation and Autoregressive Entropy Modeling | Liangji Zhu, Anand Rangarajan, Sanjay Ranka | cs.LG | 2026-08-31 |
| #46 | Exact Recovery Thresholds for Weighted Data Selection in Vector-Valued Linear Regression | Guangjian Zhang | cs.LG | 2026-08-31 |
| #47 | LaMoC: Loss-Aware Modular Compression for LLMs | Mohanad Odema, Jacob Song | cs.AI | 2026-08-31 |
| #48 | Budget-Aware Compression Pipeline for Single-GPU LLM Inference: Methods, Trade-offs, and Coupling Effects | Hongyu Yu, Yifei Shen | cs.CL | 2026-08-30 |
| #49 | Spatial Matryoshka Training for Multi-Granularity Visual Document Retrieval | Trishan Singha Roy, Arkadeep Acharya, Vishwajeet Kumar +2 | cs.AI | 2026-08-30 |
| #50 | Compression-Aware Abstention: Teaching LLMs to Refuse When KV-Compression Masks Remove Answer Evidence | Mohammadali Khodabandehlou, Bhaskar Krishnamachari | cs.CL | 2026-08-30 |