| 1 | One Editor, Many Edits: A Unified Training-Free Framework for Diverse Video Editing | Adheesh Sunil Juvekar, Onkar Kishor Susladkar, Kiet A. Nguyen +6 | cs.CV | 2026-09-03 |
| 2 | Persistent Identity Preservation in Generative Image Models: A Benchmark and Evaluation System | Mengwei Ren, Xuaner Zhang, Zhihao Xia | cs.CV | 2026-09-03 |
| 3 | When Models Edit Too Much: On the Fidelity of Minimal Code Edits | Tongyao Zhu, Wei Hern Lim, Min-Yen Kan | cs.SE | 2026-09-03 |
| #4 | Editable Visual Design | Junyan Ye, Wei Liu, Dongzhi Jiang +9 | cs.CV | 2026-09-03 |
| #5 | Instruction Duplication as an Inference-Time Control Primitive | Victor Lavrenko | cs.AI | 2026-09-03 |
| #6 | CROCODIL: Cross-Model Code Editing with LLMs | Linghan Zhong, Aditya Thimmaiah, Jayanth Srinivasa +2 | cs.CL | 2026-09-03 |
| #7 | LLaDA-Image: Building Strong Image Generators with Fully Open Training Recipes | Chuyan Chen, Haoxing Chen, Kun Chen +27 | cs.CV | 2026-09-03 |
| #8 | P-CORE: Self-Supervised Surface Consistency for Point-Based Neural Editing | Yanshu Zhang, Shichong Peng, Mehran Aghabozorgi +2 | cs.CV | 2026-09-03 |
| #9 | PointGT: Simultaneous Geometry and Texture Editing for Point-Based Representations | Yanshu Zhang, George Shramko, Pratul P. Srinivasan +1 | cs.CV | 2026-09-03 |
| #10 | SLIDEFORGE: An LLM Agent for Controllable Editing of Slides as Structured Artifacts | Haozhen Zheng, Fulin Wang, Tianhu Xiong +6 | cs.CV | 2026-09-02 |
| #11 | Multi-Tool Image Editing Attribution in Facial Forgery | Sheng Liu, Qiang Sheng, Danding Wang +3 | cs.CV | 2026-09-02 |
| #12 | AffectDelta: Beyond Emotion Labels for Image Editing | Xingzu Zhan, Lin Gu, Ruogu Fang | cs.CV | 2026-09-02 |
| #13 | RGB-to-IR image translation for infrared vehicle detection in unseen UAV domains | Thijs A. Eker, Ella P. Fokkinga, Jan Erik van Woerden +4 | cs.CV | 2026-09-02 |
| #14 | SR-Edit: Region-Aware Image Editing via Self-Refinement | Andong Wang, Zehua Chen, Yuxuan Jiang +1 | cs.CV | 2026-09-02 |
| #15 | GlyphAnchor: Enhancing Visual Text Rendering via Position-Anchored Glyph Priors | Qiang Xiang, Shuang Sun, Binglei Li +4 | cs.CV | 2026-09-02 |
| #16 | Domain shift-robust object detection with GenAI image editing | Isabel D. Stein, Thijs A. Eker, Sebastiaan P. Snel +4 | cs.CV | 2026-09-02 |
| #17 | DMRL: Document-Mediated Reinforcement Learning for Skill Optimization in Advertising Recommendation | Wei Zhang, Hongji Li, Song Sun +4 | cs.LG | 2026-09-02 |
| #18 | Selective Knowledge Edit Reversal via Gated Singular Vector Shrinkage | Weifeng Jiang, Ruirui Chen, Qianren Mao +3 | cs.CL | 2026-09-02 |
| #19 | Rendering-in-the-Loop: An Execution-Driven Agent for Interactive Web Development | Yilong Guo, Hanqi Chen, Zixiao Ye +3 | cs.CV | 2026-09-02 |
| #20 | InstEditSeg: Instruction-Driven Image Editing for Polyp and Skin Lesion Segmentation | Ziquan Liu, Zhewei Zhu, Xuyang Shi | cs.CV | 2026-09-02 |
| #21 | AVERT: Audio-Verified Adjudication for Spoken Dialogue State Tracking | Chunggi Lee, Hanspeter Pfister | cs.CL | 2026-09-01 |
| #22 | CameraEditor: Camera-Controlled Image Editing via Video-Prior Sequential Modeling | Xin Shen, Chengyou Jia, Keshuo Xing +6 | cs.CV | 2026-09-01 |
| #23 | Gaussian Core LoRA: Distribution-Aware Dynamic Adaptation for Broad Concept Erasure | Qinghui Gong, Xunlei Chen, Yu-Xuan Zhang +2 | cs.CV | 2026-09-01 |
| #24 | EdiTikZ: Scientific Figure Editing from Revision Trajectories | Christian Greisinger, Zhixue Zhao, Steffen Eger | cs.AI | 2026-09-01 |
| #25 | ExBind: A Controlled Diagnostic Benchmark for Visual-to-Executable Correspondence | Ziqian Wang, Yuxiao Cheng, Tingxiong Xiao +1 | cs.CV | 2026-09-01 |
| #26 | What Does an Agentic Software Engineering Benchmark Measure? Profiling Task Demands and Agent Behaviour Beyond What Category Labels Reveal | Radin Shayanfar, Keheliya Gallaba, Ahmed E. Hassan | cs.SE | 2026-09-01 |
| #27 | One Prompt Is Enough: Watermark Laundering Through Foundation Image Models | Jidong Yang, Qi Li, Wei Zong +5 | cs.CV | 2026-09-01 |
| #28 | Dotting the Eye: An Intent-Driven Image Retouching Agent for Visual Focus Enhancement | Chujie Qin, Zilong Zhang, Zewei Chang +5 | cs.CV | 2026-09-01 |
| #29 | Figures as Programs: Recursive Generation of Editable Scientific Figures | Yepeng Liu, Dasen Dai, Chengzhi Liu +7 | cs.AI | 2026-09-01 |
| #30 | PredErase: Training-Free Object-and-Effect Removal with Predictive Latent Guidance | Waikit Xiu, Qiang Lu, Junbiao Chen +1 | cs.CV | 2026-09-01 |
| #31 | SCoNE: Selective Context-aware Neuron Editing for Robust Retrieval-Augmented Generation | Chaewon Kim, Seo Yeon Park | cs.CL | 2026-09-01 |
| #32 | GenScale: A Benchmark for Relative Object Scale in Image Generation and Editing | Lingxiao Li, Max Whitton, Ledell Wu +1 | cs.CV | 2026-09-01 |
| #33 | Don't Let the Model Write the YAML: Deterministic, Minimal-Diff GitOps Remediation from LLM-Proposed Field Changes | Pruthvi Davineni | cs.SE | 2026-08-31 |
| #34 | Aspire: Can Models Self-Evolve from Vague Goals? | Yuhao Wu, Jingyuan Zhang, Jiajun Shi +18 | cs.CL | 2026-08-31 |
| #35 | CogEvol: Towards Efficient and Reliable Learning Environment Generation | Shangqing Tu, Daniel Zhang-Li, Yucheng Wang +20 | cs.CL | 2026-08-31 |
| #36 | ObjectSplat: Improving Mesh Fidelity and Interactivity for 3D Scenes via Object-Level Mesh Splatting | Minhas Kamal, Hiranya Garbha Kumar, Mahedi Kamal +1 | cs.CV | 2026-08-31 |
| #37 | Beneath the Diff: Diagnosing and Mitigating Algorithmic Mode Collapse in Code-Level Autonomous Research Loops | Bowei He, Weixu Zhang, Yili Jin +1 | cs.CL | 2026-08-31 |
| #38 | Co-Annotator: Expert-Distilled ViT and VLM for Visual and Documentation Guidance in Age-Related Macular Degeneration | Ziheng "Leo" Li, Benjamin Freeman, Akshay Raman +5 | cs.AI | 2026-08-31 |
| #39 | TAKE 85: Testing Audiovisual filmmaKer's intEnt across 85 Hours of Film | Kaishuu Shinozaki-Conefrey, Olivier Pascaud, Robin Courant +3 | cs.CV | 2026-08-30 |
| #40 | Pak3H: Evaluating the Cost of Cultural Mismatch in LLM Alignment with a Human-Contextualized Urdu Benchmark | Abdullah Hashmat, Usman Naseem, Agha Ali Raza | cs.CL | 2026-08-30 |
| #41 | Dior: Drawing the Light of Image via Material-Decoupled Illumination Representation | Xuanpu Zhang, Xuesong Niu, Haoxiang Cao +3 | cs.CV | 2026-08-30 |
| #42 | OrnaStyler: Ornament-Aware Latent Editing for Content-Preserving 3D Stylization | Tomohiro Aizawa, Shigeru Kuriyama, Chunzhi Gu | cs.CV | 2026-08-30 |
| #43 | En-ViMedNER: An English-Vietnamese Parallel Biomedical Corpus with UMLS Semantic Type Annotations | Nhu Vo, Phuong Nguyen, Nu Uyen Phuong Le +4 | cs.CL | 2026-08-30 |
| #44 | FRAMEWORKERS: A Dynamic Multi-Agent Framework for AI-Generated Video Production | Zhendong Li, Lei Sun, Letian Shi +8 | cs.AI | 2026-08-30 |
| #45 | RegionCache: Semantic-Aware Region Reuse for Efficient Multi-Turn Image Generation | Peizheng Li, Xin Ai, Hanyuan Liu +2 | cs.CV | 2026-08-30 |
| #46 | RePair: Turning Retrieval Failures into Counterfactual Hard Pairs | Siyi Liu, Xiaorong Zhu, Enjun Du +6 | cs.IR | 2026-08-30 |
| #47 | Can escalation channels redirect reward hacking toward defect disclosure? | Francesca Gomez | cs.AI | 2026-08-29 |
| #48 | Test-Time Scaling for Video Diffusion Models via Diagnosis-Guided Candidate Recycling | Hangzhou He, Lunhao Duan, Shanshan Zhao +4 | cs.CV | 2026-08-29 |
| #49 | LightFuse: Relightable Interactive Gaussian Scene Reconstruction via Multi-Scan Fusion and 2D Gaussian Ray Tracing | Haonan Zhou, Gaoxiang Linghu, Youlin Jia +5 | cs.CV | 2026-08-29 |
| #50 | Chat-Edit-3D++: Interactive 3D and 4D Scene Editing via Large Language Models | Shuangkang Fang, Yufeng Wang, Yi-Hsuan Tsai +4 | cs.CV | 2026-08-29 |