| 1 | Cross-Model Agreement as a Deployment-Time Reliability Signal for Automatic Polyp Segmentation | Siddharth Gupta, Jitin Singla | cs.CV | 2026-09-09 |
| 2 | Why Is Video Still So Expensive? A Survey of Inference-Efficiency Mechanisms in Video and Audiovisual LLMs | Killian Steunou, Yannis Tevissen, Mounîm A. El Yacoubi | cs.CV | 2026-09-09 |
| 3 | LinearMask-GS: Stable-Mask Importance Pruning for Compact 3D Gaussian Splatting | Donghun Ryu, Minhyeok Lee | cs.CV | 2026-09-09 |
| #4 | Elastoformer: Enabling Dynamic Adaptivity via Elastic Model Transformation | Sudaksh Kalra, Dolly Sapra | cs.CV | 2026-09-09 |
| #5 | CLFTv2: Efficient Camera-LiDAR Fusion for Semantic Segmentation via Hierarchical Feature Pyramids | Toomas Tahves, Mauro Bellone, Raivo Sell | cs.CV | 2026-09-09 |
| #6 | RealSimLoop: Online Real-to-Sim Adaptation via Differentiable Reduced-Order Simulation with Vision Feedback | Zhihao Cen, Chuhua Xian, Hailin Sun +6 | cs.GR | 2026-09-09 |
| #7 | Pairit: A Platform for Live Experiments on Human-AI Collaboration | Harang Ju, Sinan Aral | cs.HC | 2026-09-09 |
| #8 | Arti-JEPA: Adapting Video World Model to Real-Time MRI of the Vocal Tract for Speech-Production Analysis | Hong Nguyen, Sean Foley, Christina Hagedorn +4 | cs.SD | 2026-09-09 |
| #9 | HiRAD: A Flexible Large-Scale AGV Routing System | Yunjie Huang, Ruizhong Wu, Mengxuan Zhang +3 | cs.RO | 2026-09-09 |
| #10 | StreamAlign: Streaming Text-Aligned Speech Tokenization | Kang-wook Kim, Jinyoung Park, Jinsoo Kim +3 | cs.CL | 2026-09-09 |
| #11 | ALIGN-HOLD: Experience Alignment for Real-Time Hold Control in Large-Scale Ride-Hailing Matching at DiDi | Zuhao Zhang, Xu Liu, Kai Wan +3 | cs.LG | 2026-09-09 |
| #12 | X2-NativeCursor: Native-Token Text Progress Tracking for Incremental-Text Streaming Codec TTS | Zehan Liu, Carl Chen, Rime Wen +7 | cs.CL | 2026-09-09 |
| #13 | Recovering Biomechanical Signals from Missing Keypoints Using Temporal Interpolation in Monocular Gait Analysis | Shubham Jariwala | cs.CV | 2026-09-09 |
| #14 | SCCM : Stream Cruise Control Method for Automated Drift Detection and Adaptation | Mohammad Abu-Shaira, Weishi Shi | cs.LG | 2026-09-08 |
| #15 | Real-time and adaptive anomaly detection algorithm for cyclostationary models | Justyna Witulska, Tomasz Barszcz, Ireneusz Jabłoński +1 | stat.ME | 2026-09-08 |
| #16 | Improving 5G AI-RAN MCS Selection by Predicting Retransmissions | Tamerlan Aghayev, Maxime Elkael, Michele Polese +5 | cs.NI | 2026-09-08 |
| #17 | Mask Forcing: Improving Autoregressive Video Diffusion Distillation via Dual-Noise Masking Rollout | Zhuoran Zhao, Shengju Qian, Tongtong Liang +7 | cs.CV | 2026-09-08 |
| #18 | Physics-Informed Deep Learning for False Ventricular Tachycardia Alarm Reduction in the ICU | Athanasios Papastathopoulos-Katsaros, Alexandra Stavrianidi, Zhandong Liu | cs.LG | 2026-09-08 |
| #19 | ArmPoser: Real-Time, Calibration-Free Arm Pose Estimation from Smartwatch IMU | Bishnu Dev, Vasco Xu, Xi-Aan Loh +3 | cs.CV | 2026-09-08 |
| #20 | CVT-GS: Learning to Simplify 3D Gaussian Splatting with Centroidal Voronoi Tessellation | Bingxian Li, Yilong Li, Jingliang Peng +6 | cs.CV | 2026-09-08 |
| #21 | BAFF: Bid-Aware Filter Family for Mitigating Training Data Interference in RTB A/B Tests | Jeonglyul Oh, Ikkyu Choi, Inseop Youn +1 | cs.LG | 2026-09-08 |
| #22 | TontaubeV1: Streaming Text-to-Speech with Hierarchical Codec Modeling and Bounded Context | Fritz Cremer, Jonathan Cremer | cs.SD | 2026-09-08 |
| #23 | X2Streaming-ASR: wait when uncertain, emit when ready for streaming ASR | Zhiwei Lin, Kaiqi Fu, Rime Wen +5 | cs.SD | 2026-09-08 |
| #24 | MFVINS: Multiple Fisheye Camera-Based Visual Inertial System | Eunseong Jang, YuJin Chung, Sang Jun Lee +2 | cs.RO | 2026-09-08 |
| #25 | ACEA: An Adversarial Co-Evolution Arena for Head-to-Head Red-Team and Blue-Team LLM Testing | Yi Ting Shen, Kentaroh Toyoda, Alex Leung | cs.CR | 2026-09-08 |
| #26 | ReactVAU: A Slow-Fast Decoupled Framework for Streaming Video Anomaly Understanding | Chia-Hui Chen, Shih-Ying Yeh, Fu-En Yang +2 | cs.CV | 2026-09-07 |
| #27 | Explainable Temporal Attention-based Defect Detection For Fillet Joints in Real-Time Gas Metal Arc Welding Based on Multi-modal Data | Mobina Mobaraki, Mahyar Asadi, Klaske Van Heusden +1 | cs.AI | 2026-09-07 |
| #28 | Scene Graph-Driven Haptic Feedback for Safety Enhancement in Robotic Ophthalmic Surgery via Physically Simulated iOCT | Danial Arbabi, Korab Hoxha, Angelo Henriques +2 | cs.RO | 2026-09-07 |
| #29 | The OCUDU dApp Platform: An Open Runtime and E3 Interface for Real-Time AI-RAN | Timothy O'Shea, Matthew Pennybacker, Andriy Kharchenko | cs.NI | 2026-09-07 |
| #30 | TFTrack: A Template-Free Framework for Efficient 3D Point Cloud Tracking | Zhaofeng Hu, Sifan Zhou, Jiahao Nie +3 | cs.CV | 2026-09-07 |
| #31 | A Tool-Augmented, GPT-4 Chatbot for Real-Time Repository Data Analysis | Muhammad Jawad Chowdhury, Md. Sakib Khan | cs.AI | 2026-09-07 |
| #32 | RAFM-SER++: A Lightweight Multimodal Emotion Recognition Framework for Real-Time Behavioral Monitoring in Surveillance Systems | Ngo Truong Dinh, Tung-Lam Bui, Chi-Trung Duong +3 | cs.AI | 2026-09-07 |
| #33 | Inferring Urban Mobility Interactions from Aggregated Dynamics | Yi Wang, Jing Li, Jinliang Deng +5 | cs.LG | 2026-09-07 |
| #34 | CRISP: Corneal Confocal Microscopy Real-Time Image Stitching Pipeline | Qincheng Qiao, Puli Zhang, Jian Zhou +1 | cs.CV | 2026-09-07 |
| #35 | PLATOS: A Power and Latency-Aware Task-Oriented Scheduling Strategy for Healthcare IoT in Fog Computing | Mohammed Alaa Ala'anzy, Zulfiqar Ahmad, Zhanar Mukash | cs.DC | 2026-09-07 |
| #36 | LightSplat: Real-Time High-Fidelity 3D Gaussian SLAM with Loop Closure | Junze Bao, Ye Gao, Yiming Huang +5 | cs.RO | 2026-09-07 |
| #37 | EmoMed: An Emotionally-Aware Agent for Multimodal Medical Support with Real-Time Information Retrieval | Ivan Nasonov, Nikita Glazkov, Ivan Makovetskiy +4 | cs.AI | 2026-09-07 |
| #38 | AF-Mamba: Efficient Long-Term Signal Modeling for Early Prediction of Atrial Fibrillation Onset | Yongbin Lee, Ki H. Chon | cs.LG | 2026-09-07 |
| #39 | Novel Methods for Catheter and Guidewire Segmentation in X-ray Fluoroscopy under a Federated Learning Setting | Chayun Kongtongvattana | cs.CV | 2026-09-06 |
| #40 | FSAN: Flow State Attention Network for Aerodynamic Prediction | Wenxuan Jin, Jianguo Yao, Haibing Guan +1 | cs.CV | 2026-09-06 |
| #41 | Selective Knowledge Control for Continual GUI Agent Learning over Application Streams | Zirui Shang, Xin Shu, Yang Liu +3 | cs.CV | 2026-09-06 |
| #42 | A Cloud-Based Hybrid Model for Real-Time Detection of BRTA-Approved Licence Plates Using YOLO Tiny and Haar Cascade | Debashis Kar Suvra, Tahsina Farah Sanam | cs.CV | 2026-09-06 |
| #43 | One MLLM, One Call: Efficient Zero-Shot Vision-and-Language Navigation via Spatial-Aware Waypoints | Shiqi Pan, Qi Zheng, Hanqin Sun +3 | cs.CV | 2026-09-06 |
| #44 | A Hybrid Predictive Ensemble of Machine Learning and Deep Neural Networks for Early Cardiovascular Disease Risk Assessment | Balaji Venkateswaran | cs.AI | 2026-09-04 |
| #45 | ARIA - An Agentic Framework for Autonomous Testing of Infotainment Systems | António Azevedo, Bruno Lima, João Pascoal Faria | cs.SE | 2026-09-04 |
| #46 | CHAMP: Cross-domain Hybrid Architecture for Matchmaking and Prediction in Online Multi-Player Games | Kai Wang, Ge Fan, Chaoyun Zhang +2 | cs.AI | 2026-09-04 |
| #47 | LUMIN: Lightweight Universal Manufacturing Inspection Network for Anomaly Detection | Pengfei Yang | cs.CV | 2026-09-04 |
| #48 | Resilience Beyond Stationary Client Unavailability: Unlocking Efficient and Unbiased Federated Learning | Ming Xiang, Stratis Ioannidis, Edmund Yeh +2 | cs.LG | 2026-09-04 |
| #49 | Predicting Spatiotemporal Mobile Sensing-Based PM2.5 Concentrations Using Low-Rank Adapted Spatially Attentive Graph Neural Network | Om Chiddarwar, Priyanka Mandal, Praveen Kumar Chandaliya +1 | cs.AI | 2026-09-04 |
| #50 | Occlusion-Robust Multimodal Emotion Recognition in VR via Fusion of Facial Images and EMG | Birgit Nierula, Karam Tomotaki-Dawoud, Mert Akguel +5 | cs.CV | 2026-09-03 |