| 1 | Zero-Shot Novel Depth Synthesis Using 3D Foundation Models Scene Representations | Denis M. Akola, David F. Fouhey | cs.CV | 2026-09-03 |
| 2 | Adaptive Vision-Language Grasping via Composable Foundation Priors and Generalizable Grasp Synthesis | Sixu Yan, Shikang Wang, Binhua Huang +11 | cs.RO | 2026-09-03 |
| 3 | Alignment-Free Text-Audiobox for Voice Dubbing and Full-Duplex Dialogue Synthesis | Sanyuan Chen, Min-Jae Hwang, Sho Inoue +12 | cs.CL | 2026-09-03 |
| #4 | Sharpening the Ensemble: An SSIM-Aligned Residual Refiner for Brain-MRI Inpainting Post-Processing | Kubilay Kağan Kömürcü, İlkay Öksüz | cs.CV | 2026-09-03 |
| #5 | Proactive Service Agents: A Unified Decision Framework, Methods, and Evaluation | Yan Tang, Tingyu Cao, Yuanbo Tang +2 | cs.AI | 2026-09-03 |
| #6 | Stabilizing Camera-Controlled Novel View Synthesis at Inference Time | Prajwal Singh, Arjun Badola, Seema Kumari +2 | cs.CV | 2026-09-03 |
| #7 | Auditing Patient Privacy in Medical Generative Models: Scalable Memorization Detection with DeepSSIM++ | Antonio Scardace, Francesco Guarnera, Sebastiano Battiato +1 | cs.CV | 2026-09-03 |
| #8 | LevelSyn: Physical-Aware Logic Synthesis via Level-Asynchronous Graph Neural Networks | Jingyi Zhou, Zhengyuan Shi, Ziyang Zheng +1 | cs.AR | 2026-09-03 |
| #9 | Text2Thermal: Physics-Aware Thermal Image Synthesis from Textual Priors | Tayeba Qazi, Brejesh Lall, Prerana Mukherjee | cs.CV | 2026-09-03 |
| #10 | TruncGradGS: Improved 3D Gaussian Splatting via Truncated Gradient Updates | Theo Morales, Nhat-Quynh Le-Pham, Robin Atkins +1 | cs.CV | 2026-09-03 |
| #11 | Mudragen: Geometrically Supervised Generation of Interacting Two-Hand Mudras for Preserving Indian Classical Dance Heritage | Jagadish Kashinath Kamble, Jayanta Mukhopadhyay, Debaditya Roy +1 | cs.CV | 2026-09-03 |
| #12 | SurgeGen: A Hybrid Generative Diffusion Framework for Storm Surge Scenario Synthesis | Shunan Zheng, John J. Hasenbein | math.DS | 2026-09-03 |
| #13 | PointGT: Simultaneous Geometry and Texture Editing for Point-Based Representations | Yanshu Zhang, George Shramko, Pratul P. Srinivasan +1 | cs.CV | 2026-09-03 |
| #14 | VoxReason: Listener-Free Evaluation of Source-Grounded Speech Planning Before Synthesis | Mengzhe Geng | cs.SD | 2026-09-02 |
| #15 | CRAW: Codec Robust Audio Watermarking | David Chernin, Ethan Fetaya | cs.SD | 2026-09-02 |
| #16 | Advances in Machine Learning for Directed Evolution: A Five-Year Retrospective | Bruce J. Wittmann | q-bio.BM | 2026-09-02 |
| #17 | Population-Calibrated Graph Screening at 835-Million-Address Scale, with Label-Free Transfer to New Chains | Yury Korolev | cs.CR | 2026-09-02 |
| #18 | Thinking in Pictures: A Systematic Benchmark for Reasoning-driven Image Generation | Yutong Liu, Nan Huang, Xu Cao +1 | cs.CV | 2026-09-02 |
| #19 | Towards Trustworthy Autonomous Robots: An Explainable AI-Based Decision Framework | Cagri Temel | cs.RO | 2026-09-02 |
| #20 | RoGe: Novel View Synthesis via End-to-End Implicit Reconstruction and Generation | Xiaolei Lang, Ze Kang, Zehao Huang +1 | cs.CV | 2026-09-02 |
| #21 | GDB-Reward: From Evaluation Metrics to Training Rewards for Graphic Design | Adrienne Deganutti, Purvanshi Mehta, Simon Hadfield +1 | cs.CV | 2026-09-02 |
| #22 | Genesis: A Generative Engine for Hierarchical Satellite Image Synthesis | Subash Khanal, Yangzhi Cui, Daniel Cher +4 | cs.CV | 2026-09-02 |
| #23 | Loom: Weaving Diagnostic Strands into Free-Text Consensus via Embedding-Space Reweighting | Ron Begleiter, Katya Egert Berg, Gilad Saban +1 | cs.AI | 2026-09-02 |
| #24 | CrashDiffuser: VLM-Guided Collision Intent Reasoning for Fine-Grained Safety-Critical Traffic Scenario Generation | Shucheng Zhang, Yuang Zhang, Bingzhang Wang +3 | cs.RO | 2026-09-02 |
| #25 | Retrosynthesis of Synthetic Media for Explainable AI Provenance Forensics | Yijie Lin, Ching-Chun Chang, Isao Echizen +2 | cs.CR | 2026-09-02 |
| #26 | CC-4DGS: Computational Deformation and Point-Cloud Compression for Storage-Efficient Dynamic Gaussian Splatting | Kyungdae Park, Chae Eun Rhee | cs.CV | 2026-09-02 |
| #27 | Rendering-in-the-Loop: An Execution-Driven Agent for Interactive Web Development | Yilong Guo, Hanqi Chen, Zixiao Ye +3 | cs.CV | 2026-09-02 |
| #28 | NS-Copilot: An LLM-Driven Agent System for Autonomous Neuroscience Analysis | Wuche Liu, Yiran Qiao, Linlin Hou +4 | cs.CL | 2026-09-02 |
| #29 | The Ceiling Is in the Channel: Auditing Learner Gaps and Measurement Frontiers in Clinical Prediction | Sayeed Shafayet Chowdhury, Nusrat Jahan, Snehasis Mukhopadhyay +2 | cs.AI | 2026-09-01 |
| #30 | Dictionary-Guided Mutation Operators for Automated HDL Repair | Maisha Mastora, Dean Sullivan | cs.ET | 2026-09-01 |
| #31 | Hearing the Whispers: Black-Box Membership Inference Attacks on Finetuned TTS Models | Kunlin Cai, Kaiyuan Zhang, Zihang Xiang +4 | cs.CR | 2026-09-01 |
| #32 | SpatialGuard: Harness-Guided Verifiable Spatial Reasoning for Text-to-Image Generation | Ziyun Qian, Zizhi Chen, Yizhou Liu +3 | cs.CV | 2026-09-01 |
| #33 | DualDiff3D: Dual Structure-Appearance Diffusion Priors for Reliability-Enhanced 3D Gaussian Splatting | Qian Wang, Yu Wang, Weiqi Li +4 | cs.CV | 2026-09-01 |
| #34 | Neuro-Symbolic Geometric Abstraction (NeuSOGA): From Observations to Symbolic Mathematical Representations | Qingde Li, Qingqi Hong, Zihan Li +1 | cs.AI | 2026-09-01 |
| #35 | Autonomous discovery of new structure-plausibility laws for explainable and rapid crystal diagnosis and screening | Zhilong Song, Lixue Cheng | cond-mat.mtrl-sci | 2026-09-01 |
| #36 | On Synthesis of Metric Interval Temporal Logics | Hsi-Ming Ho, Shankaranarayanan Krishna, Khushraj Madnani | cs.LO | 2026-09-01 |
| #37 | Phrase-Localized Language-Contrastive Guidance: Training-Free Localized Accent Control for Code-Switching Text-to-Speech | Che Hyun Lee, Sangkwon Park, Donghun Kang +4 | cs.CL | 2026-09-01 |
| #38 | EvoGS: Modeling Deformation Evolution for Dynamic Gaussian Splatting | Wei Dong, Shahram Shirani, Jun Chen +1 | cs.CV | 2026-09-01 |
| #39 | Verifiable Disaster Storylines and Causal Knowledge Graphs: A Citation-Grounded Pipeline from Heterogeneous Humanitarian Sources | Ivan Decostanzi, Michele Ronco, Sergio Consoli +8 | cs.AI | 2026-09-01 |
| #40 | Reflect-SQL: A Self-Reflection Based Framework for Text-to-SQL | Anupreksha Jain, Manish Shrivastava | cs.IR | 2026-09-01 |
| #41 | Differentially Private Paired Table-Image Multimodal Synthesis | Kai Chen, Josephine Lamp, Somesh Jha +1 | cs.CR | 2026-09-01 |
| #42 | Physically Plausible Video Generation via Visual-Semantic Chain-of-Events Conditioning | Zixuan Wang, Yixin Hu, Wen Li +4 | cs.CV | 2026-09-01 |
| #43 | It Takes Two to Match: Co-Evolving Generative Retriever with Reinforcement Learning | Runpeng Dai, Kaili Huang, Changsung Kang +1 | cs.IR | 2026-09-01 |
| #44 | Streaming4D: Accelerate 4D World Models via Block-wise Video Generation and Incremental Reconstruction | Xiaoyan Liu, Jiaxin Liu, Kangrui Li +1 | cs.CV | 2026-09-01 |
| #45 | When the Algorithm Becomes the Brand Crisis: A Sociotechnical Theory of Distributed Responsibility and Accountable Transparency | Mohammad Saleh Torkestani, Taha Mansouri | cs.AI | 2026-09-01 |
| #46 | Conversation Coach: A Voice-enabled AI System that Helps Practice Difficult Workplace Conversations | Fanyou Wu, Suraj Maharjan, Ainur Yessenalina +3 | cs.AI | 2026-08-31 |
| #47 | Puppeteer: Object-Grounded Posture-Aware Co-Speech Gesture Generation | Vida Adeli, Soroush Mehraban, Jacob Rommann +3 | cs.CV | 2026-08-31 |
| #48 | Distributed Implicit Harm: A Compositional Safety Blind Spot in MLLM-Based Video Moderation | Ruotong Wang, Zihao Zhu, Siwei Lyu +2 | cs.CV | 2026-08-31 |
| #49 | Beyond Textual Chain-of-Thought: A Survey on Action-Grounded Reasoning in Autonomous Driving | Zhengxu Tang, Xiaozhou Zhang, Guofeng Cui +10 | cs.CV | 2026-08-31 |
| #50 | Flawed in Nature, Perfect through Evolution | J. M. Diederik Kruijssen | cs.LG | 2026-08-31 |