| 1 | Thinking in Pictures: A Systematic Benchmark for Reasoning-driven Image Generation | Yutong Liu, Nan Huang, Xu Cao +1 | cs.CV | 2026-09-02 |
| 2 | Eliciting ESG Preferences for Reinforcement Learning-Based Portfolio Optimization | Giovanni Dispoto, Marcello Restelli, Carmine Ventre | q-fin.PM | 2026-09-02 |
| 3 | Spatially Aware World Action Model via Geometric Latent Diffusion | Javier Alejandro Lopetegui Gonzalez, Paul Pacaud, Cordelia Schmid | cs.CV | 2026-09-02 |
| #4 | Training seeds and model-selection stability in recommender-system evaluation | Juan Manuel Rodriguez, Oleg Lesota, Antonela Tommasel | cs.IR | 2026-09-02 |
| #5 | Learning to Track from Privileged Target Appearances | Xin Chen, Jiao Xu, Dong Wang +2 | cs.CV | 2026-09-02 |
| #6 | CivBench: A Long-Horizon Benchmark for Tool-Mediated Agents in Civilization VI | Austin Tudor David Andrews, Liam Wilkinson, Jamie Heagerty +3 | cs.AI | 2026-09-02 |
| #7 | The Diagnosis a Reporter Leaves Unspoken: Surfacing Frozen Tumor Features for Brain-Tumor MRI Reporting | Khawaja Murad ul Hassan, Ruqiyya Adil, Adil Qayyum +4 | cs.CV | 2026-09-02 |
| #8 | ProSR: Semantic-Prototype-Guided Discrete Modeling for Physically Consistent SAR Super-Resolution | Byoungwoo Kim, Munchurl Kim | cs.CV | 2026-09-02 |
| #9 | AGI Maze Prediction Datasets: A Compact Benchmark for Learning World Dynamics with Transformers | Alexey Potapov | cs.LG | 2026-09-02 |
| #10 | VoRTeC: Taming Foundation Flow for One-step Real time Video Compression | Yichong Xia, Qinhong Wu, Qinhong Wu +3 | cs.CV | 2026-09-02 |
| #11 | CAPTURE: Disentangling Preference Drift from Memory Poisoning in Personalized LLM Agents | S M Asif Hossain, Ruksat Khan Shayoni, Md Kishor Morol | cs.LG | 2026-09-02 |
| #12 | Propose to Learn, Learn to Propose: Evaluability-Aware Assistance under Bounded Rationality | Yifan Zhu, Sammie Katt, Samuel Kaski | cs.AI | 2026-09-02 |
| #13 | GeoSPRINT: Geometric Redundancy-Aware Step Pruning for Inference in Diffusion Trajectories | Arpita Joshi | cs.LG | 2026-09-02 |
| #14 | EmoStance: Response-Side Affective-Orientation Control for Empathetic Response Generation via Emoji Weak Supervision | Ziyuan Jin, Yuxuan Ge, Zheng Tian | cs.AI | 2026-09-02 |
| #15 | Synergistic Information Disentanglement for Omni-modal Slide Representation Learning in Computational Pathology | Mingxin Liu, Chengfei Cai, Anwen Lu +5 | cs.CV | 2026-09-02 |
| #16 | A Unified Rate-Distortion Perspective on Vector, Product, and Scalar Quantization | Xianghong Fang, Wenlong Mou, Yuan Yuan +2 | cs.LG | 2026-09-02 |
| #17 | The Dynamics of Continuous Mixture Collapse in Language Models | Ali Backour | cs.LG | 2026-09-02 |
| #18 | SelfLift: Accelerating Few-Step Diffusion via Self-Recovering Resolution Transition | Tingyan Wen, Chenqian Yan, Xurui Peng +4 | cs.CV | 2026-09-02 |
| #19 | InstEditSeg: Instruction-Driven Image Editing for Polyp and Skin Lesion Segmentation | Ziquan Liu, Zhewei Zhu, Xuyang Shi | cs.CV | 2026-09-02 |
| #20 | Linear Fusion MultiDiffusion for Fast Training-Free Spherical Panorama Generation | Akio Hayakawa, Yusuke Mukuta, Tatsuya Harada | cs.CV | 2026-09-02 |
| #21 | A Unified Particle Filter LSTM for Data-Driven Process Simulation | Parvin Malekzadeh, Opher Baron, Dmitry Krass | cs.LG | 2026-09-02 |
| #22 | Looped Transformers under the Jacobian Lens: Does the Global Workspace Survive Recurrence? | Wenlong Wang, Fergal Reid | cs.AI | 2026-09-01 |
| #23 | Latent unified smooth Hamiltonians for excited state chemistry | David Juergens, Martin Stöhr, Andreas E. Hillers-Bendtsen +2 | physics.chem-ph | 2026-09-01 |
| #24 | Belief-Calibrated Optimization: An Explicit World Model for Agentic Optimization | Yuhan Chen, Zhihua Tian, Mahavir Dabas +7 | cs.AI | 2026-09-01 |
| #25 | AlphaRAD: Grounded Zero-Shot Classification in Chest Radiology via $α$-Corrected Binary Cross Entropy and Factorized Latent Supervision | Jianzhong You, Yuan Gao, Chris McIntosh | cs.CV | 2026-09-01 |
| #26 | ZipTok3D: High-Fidelity 3D Tokenization with Compact Token Prefixes | Mingda Lin, Weijie Wang, Zeyu Zhang +7 | cs.CV | 2026-09-01 |
| #27 | H3-World: Turning Language Understanding into World Control | Danze Chen, Zeqing Wang, Ziyue Lin +2 | cs.CV | 2026-09-01 |
| #28 | What, Where, and How: Probing Spatiotemporal Representations in Video Foundation Models | Sharon S. Musa, Fereshteh Forghani, Harrish Thasarathan +3 | cs.CV | 2026-09-01 |
| #29 | Quantum Sparse Autoencoders for Q-Matrix Estimation in Cognitive Diagnosis | Arif Hassan Zidan, Yi Pan, Bowen Guo +5 | cs.LG | 2026-09-01 |
| #30 | EvoSCM: Scientific Belief Revision Through Causal Model Evolution and Experimentation | Qing Zhao, Haowei Li, Weijian Deng +2 | cs.AI | 2026-09-01 |
| #31 | LatentPress: Context Compression Beyond Text and Vision | Zhengze Zhou, Hejian Sang | cs.LG | 2026-09-01 |
| #32 | Gaussian Core LoRA: Distribution-Aware Dynamic Adaptation for Broad Concept Erasure | Qinghui Gong, Xunlei Chen, Yu-Xuan Zhang +2 | cs.CV | 2026-09-01 |
| #33 | Learning Sparse Decision Trees via Transformer Variational Auto-Encoders | Giacomo Fidone, Alessio Cascione, Riccardo Guidotti | cs.LG | 2026-09-01 |
| #34 | Neuro-Symbolic Geometric Abstraction (NeuSOGA): From Observations to Symbolic Mathematical Representations | Qingde Li, Qingqi Hong, Zihan Li +1 | cs.AI | 2026-09-01 |
| #35 | ExBind: A Controlled Diagnostic Benchmark for Visual-to-Executable Correspondence | Ziqian Wang, Yuxiao Cheng, Tingxiong Xiao +1 | cs.CV | 2026-09-01 |
| #36 | TimeSteer: Inference-Time Speech Scheduling in Joint Audio-Visual Diffusion Models | Chao Zhou, Yiling Chen, Qi Chu +3 | cs.CV | 2026-09-01 |
| #37 | Seeing the World and the Self from Egocentric Video | Kai Guan, Minchao Jiang, Ruichen WangLi +2 | cs.CV | 2026-09-01 |
| #38 | REFACTOR-VLA: Unsupervised Library Learning of Typed Motor Programs | Riyaaz Shaik, Chandru Venkataraman | cs.LG | 2026-09-01 |
| #39 | Superposed Latent Autoencoder | Quanling Zhao, Jiaying Yang, Tianqi Zhang +4 | cs.LG | 2026-09-01 |
| #40 | Latent Recurrent Thoughts: Recurrent Refinement of Proposed Latents for Reasoning with Frozen LLMs | Zhaoliang Chen, Jie Fu | cs.AI | 2026-09-01 |
| #41 | OUTLETS: Output-Length Prediction from Speculative Decoding Backbones | Weihuang Wen, Yingying Liu, Yichuan Liu +5 | cs.CL | 2026-09-01 |
| #42 | QILP-0: Constructing Observational Declarative Twins of Quantum Circuits | Marina de la Cruz Echeandía, César Luis Alonso, Tony Ribeiro +1 | cs.AI | 2026-09-01 |
| #43 | From Truncation to Commitment: Persistent Context in Uniform Discrete Diffusion | Satoshi Hayakawa | cs.LG | 2026-09-01 |
| #44 | ReFlowSET: Representation-Aligned Latent Flow Matching for SAR-to-EO Image Translation | Jeonghyeok Do, Seungchul Lee, Munchurl Kim | cs.CV | 2026-09-01 |
| #45 | PredErase: Training-Free Object-and-Effect Removal with Predictive Latent Guidance | Waikit Xiu, Qiang Lu, Junbiao Chen +1 | cs.CV | 2026-09-01 |
| #46 | A Dataset for Modeling Iterative Problem-Solving | Fagun Patel, Sang T. Truong, Duc Q. Nguyen +4 | cs.CL | 2026-09-01 |
| #47 | Poisson-Gamma Dynamical Systems with Time-varying Transition Dynamics | Jiahao Wang, Yijun Wang, Nan Fang +1 | cs.LG | 2026-09-01 |
| #48 | Beyond the Clock: Measuring the Value of Adaptive Revision | Ayushi Chadha | cs.AI | 2026-09-01 |
| #49 | Training-Free Inpainting Across Domains with a Frozen Text-to-Image Diffusion Model | Zhenhuan Wang, Fengyi Yuan | cs.CV | 2026-09-01 |
| #50 | FLaG: Frequency-Domain Latent-attention Gated Pooling for Token Aggregation | Kewei Li, Rongying Zhang, Xueli Wang +6 | cs.AI | 2026-09-01 |