| 1 | CLoSeR: Closing the Loop for Long-Context Streaming Reconstruction | Moyang Li, Zihan Zhu, Wei Zhang +2 | cs.CV | 2026-10-01 |
| 2 | OneStreamer: Unifying Perception, Memory, and Proactive Response in Streaming Video Interaction | Xiangyu Zeng, Yuandong Yang, Zhiqiu Zhang +21 | cs.CV | 2026-10-01 |
| 3 | Learning PDE Dynamics between Submanifolds Using Green's Observation Operators | Jan Tauberschmidt, Jephte Abijuru, Samuel Okon +5 | cs.LG | 2026-10-01 |
| #4 | Which LLM to pick? Online Active Model Selection for Large Language Models | Alessandro Turrin, Patrik Okanovic, Torsten Hoefler +1 | cs.CL | 2026-10-01 |
| #5 | Streaming algorithms for robust max-min diversification | Andrea Pietracaprina, Geppino Pucci, Stefano Zanon | cs.LG | 2026-10-01 |
| #6 | FlashBack: Knowing When to Remember in Streaming Vision-Language Models | Yi Chen, MingMing Yu, Rui-Qi Wang +5 | cs.CV | 2026-10-01 |
| #7 | Variational Streaming Flow: Probabilistic Forecasting in Physical Time | Hans Hao-Hsun Hsu, Minseon Gwak, Soon Hoe Lim +2 | cs.LG | 2026-10-01 |
| #8 | Memorizon: Training World Models Beyond Their Context Window | Tingting Liao, Xuezhi Liang, Hao Li +1 | cs.CV | 2026-09-30 |
| #9 | I Have a Stream: Making Self-Supervised Learning Work on Continuous Video | Ivan Martinović, Lukas Knobel, Yuki M. Asano | cs.CV | 2026-09-30 |
| #10 | StreamRig: Exploiting Intra-Rig Geometry for Streaming Multi-Camera Odometry | Yufei Wei, Shuhao Ye, Qi Wang +4 | cs.CV | 2026-09-30 |
| #11 | LOCI: Spatial Linear Memory for Streaming World Models | Ji Xia, Tingting Liao, Xuezhi Liang +2 | cs.CV | 2026-09-30 |
| #12 | Enhancing Autoregressive Video Generation via Representation Adversarial Distillation | Fangyu Lin, Xingtong Ge, Lunjie Zhu +8 | cs.CV | 2026-09-30 |
| #13 | LEAP: Learned Block-wise Evidence Retrieval for Long Audio-Video Perception | Juyi Lin, Zhiqiang Lao, Jiali Cui +9 | cs.CL | 2026-09-30 |
| #14 | CommunityKV: Efficient Long-Context Decoding via Graph Partitioning | Joe McKenna, Anastasios Alexandridis, Nathan Susanj +1 | cs.LG | 2026-09-30 |
| #15 | DuplexAct-Bench: Broadening Full-Duplex Speech Evaluation toward Proactive Interaction across Diverse Behavioral Requirements | Keyue Xing, Wentao Ding, Mengmeng Wang +3 | cs.CL | 2026-09-30 |
| #16 | What Streaming Anomaly Detection Finds (and Misses) in Industrial Time Series | Magali Parrino, Antoine Ajenjo, Emmanuel Remy +2 | cs.LG | 2026-09-30 |
| #17 | In a Streaming World, Should You Stand Still? A Comprehensive Benchmark of Anomaly Detection in Streams | Magali Parrino, Antoine Ajenjo, Emmanuel Remy +3 | cs.LG | 2026-09-30 |
| #18 | What Should an Agent Remember? Disentangling Retention from Retrieval in Bounded-Memory Evaluation | Juli Huang | cs.AI | 2026-09-30 |
| #19 | DeCoPrune: Efficient KV-Cache Pruning for Autoregressive Video Diffusion via Denoising Consistency | Zeqi Xiao, Qingle Liu, Kaiwen Zhang +3 | cs.CV | 2026-09-30 |
| #20 | SCIC: Scope- and Codebook-Aware Instruction Conditioning for Speaker-Adapted Expressive TTS | Longyu Lu, Zongwei Du, Mengtao Xing +4 | cs.SD | 2026-09-30 |
| #21 | Association profile conditioning in a set-temporal transformer for cross-session intracortical motor decoding | Xinyuan Zhang, Handong Mo, Pengfei Wen +5 | q-bio.NC | 2026-09-30 |
| #22 | MEMO: Multi-Level Entity-Aware Memory for Streaming Video Understanding | Yinying Li, Yuqian Fu, Yulin Dai +3 | cs.CV | 2026-09-30 |
| #23 | VOSSA: Voiceprint Optimization for Streaming Speech Architectures | Mu-Ruei Tseng, Waris Quamer, Ghady Nasrallah +1 | eess.AS | 2026-09-30 |
| #24 | No Corners Cut: State-Grounded Transitions for Mid-Stream Prompt Switches in Video Generation | Zejing Rao, Ketong Ren, Xiaoqiang Liu +3 | cs.CV | 2026-09-30 |
| #25 | NeurDuo-EEG: A Long-Sequence EEG Foundation Model with Persistent State and Explicit Memory | Yifan Wang, Haiping Liu, Yang Cui +13 | cs.LG | 2026-09-29 |
| #26 | MOBA-VL: Event-Localized Multi-Turn Reinforcement Learning for Real-Time MOBA Commentary | Shengyun Zhong, Xinkang Zhao, Ziyuan Chu +1 | cs.CV | 2026-09-29 |
| #27 | HelixWorld: A Real-time Interactive Audio-Visual World Model | Lei Ke, Jiahao Pan, Zeyue Tian +13 | cs.CV | 2026-09-29 |
| #28 | Self-Aligned Forcing: Streaming Video Diffusion with Differentiable Noisy History | Weiqiang Wang, Zhuokun Chen, Yusheng Dai +5 | cs.CV | 2026-09-29 |
| #29 | RelayVSR: Large-Small Model Collaboration for Efficient Real-World Video Super-Resolution | Xijun Wang, Xin Li, Zirui Lang +3 | cs.CV | 2026-09-29 |
| #30 | ReCaVSR: One-Step Streaming Diffusion Video Super-Resolution with Recycled Latents and Learned Cache Routing | Xijun Wang, Xin Li, Suhang Yao +3 | cs.CV | 2026-09-29 |
| #31 | APM-Bench: Benchmarking Cross-session Persistent Memory for Egocentric Streaming Video Assistants | Jianguo Huang, Jinming Liu, Qiyao Wang +9 | cs.CV | 2026-09-29 |
| #32 | RawVLA: Embodied Neural Image Signal Processor For Robotic Manipulation | Shuhong Liu, Heng Zhou, Lingfeng Qian +7 | cs.RO | 2026-09-29 |
| #33 | When to Retrieve, When to Stay: Uncertainty-Aware Temporal Evidence Allocation for Streaming Video-LLMs | Xiang Hu, Jiazuo Yu, Lu Zhang +2 | cs.CV | 2026-09-29 |
| #34 | Watch-Think-Interact: Bootstrapping Long-Horizon Multi-Turn Streaming Video Reasoning with Reinforcement Learning | Ziheng Huang, Yicheng Bao, Xueheng Li +8 | cs.AI | 2026-09-29 |
| #35 | Salt++: Context-Aligned Post-Training for Few-Step Streaming Multimodal Generation | Xingtong Ge, Yutong Wang, Lunjie Zhu +7 | cs.CV | 2026-09-29 |
| #36 | EGSD: Event-Grounded Self-Distillation for Streaming Video Understanding | Yuwei Miao, Xuesheng Zhang, Wenhao Zou +5 | cs.CV | 2026-09-29 |
| #37 | FastVR: Efficient Streaming Video Restoration with One-Step Diffusion | Xiaoxu Chen, Qin Yang, Haoran Bai +2 | cs.CV | 2026-09-29 |
| #38 | Video2Skill: From Streaming Experience to Reusable Embodied Skills | Jianshu Zhang, Ce Zhang, Xiyuan Yang +6 | cs.CL | 2026-09-29 |
| #39 | MyoCodec: A Streaming Neural Codec for Electromyography | Jihwan Lee, Kleanthis Avramidis, Junhyeok Lee +3 | eess.SP | 2026-09-29 |
| #40 | You Only Reprogram Once: Rethinking Prolonged Training for Visual Reprogramming | Zizhao Li, Mohammed Yaqoob Ansari, Xinyu Su +3 | cs.CV | 2026-09-29 |
| #41 | MemEvo: Automatic Discovery of Streaming Video Memory Mechanisms | Guohong Liu, Jialei Ye, Shanhui Zhao +2 | cs.AI | 2026-09-29 |
| #42 | HiTS-CL: A Continual Learning Framework for Long-Horizon Temporal Knowledge Graph Extrapolation | Yansong Liu, Rui Liu, Yuan Zuo +6 | cs.LG | 2026-09-29 |
| #43 | Quantum Computing for Network Security Classification: Near-Term Classification and Long-Term Memory Efficiency | Yuqing Li, Poonam Bala Nehru, Yunpeng Zhang +4 | quant-ph | 2026-09-29 |
| #44 | Staircase Policy: Streaming Inference for World-Action Models with Large Action Chunks | Guoheng Sun, Chen Chen, Jin Wang +2 | cs.RO | 2026-09-29 |
| #45 | SMat-Attention: Structured Long-Context Sequence Modeling | Emile Anand, Abdullah Ateyeh, Archer Wang +1 | cs.AI | 2026-09-28 |
| #46 | KV-streams for Efficient Compaction in Agentic Reinforcement Learning | Emiliano Penaloza, Dane Malenfant, Dheeraj Vattikonda +15 | cs.LG | 2026-09-28 |
| #47 | InfiniHand: Streaming World-Space Hand Motion Estimation from Egocentric Video | Kerui Ren, Kaiwen Song, Weiguang Zhao +8 | cs.CV | 2026-09-28 |
| #48 | FlowAct-R2: Beyond Talking Avatar via Streaming Multimodal References and Proactive Agent Planning | Ziyao Huang, Zhengkun Rong, Shiyang Qin +5 | cs.CV | 2026-09-28 |
| #49 | When Does a Spoken Agent Have Enough Evidence to Act? The PACT-SLM Contract Test | Mengzhe Geng | cs.SD | 2026-09-28 |
| #50 | Simultaneous Translation between Sign Languages | Zetian Wu, Bowen Xie, Stefan Lee +1 | cs.CV | 2026-09-28 |