| 1 | GDB-Reward: From Evaluation Metrics to Training Rewards for Graphic Design | Adrienne Deganutti, Purvanshi Mehta, Simon Hadfield +1 | cs.CV | 2026-09-02 |
| 2 | SafeEvolve: Harness-Policy Co-Evolution from Agent Experience for Safety Alignment | Qinghua Mao, Wanying Qu, Dadi Guo +8 | cs.AI | 2026-09-02 |
| 3 | CodePoisonRAG: Knowledge Poisoning Attacks on Retrieval-Augmented Code Generation | Varun Gadey, Ziad Marey, Alexandra Dmitrienko | cs.CR | 2026-09-02 |
| #4 | Multi-Tool Image Editing Attribution in Facial Forgery | Sheng Liu, Qiang Sheng, Danding Wang +3 | cs.CV | 2026-09-02 |
| #5 | DKL: Decoupled Knowledge Learning for Instruction-Tuned Language Models | Kushagra Bhushan, Meghanadh Pulivarthi, Sai Krishna Reddy Sathi +7 | cs.CL | 2026-09-02 |
| #6 | RGB-to-IR image translation for infrared vehicle detection in unseen UAV domains | Thijs A. Eker, Ella P. Fokkinga, Jan Erik van Woerden +4 | cs.CV | 2026-09-02 |
| #7 | When Persona Attributes Improve Population Alignment in Large Language Models | Leon Fröhling, Jens Rupprecht, Markus Strohmaier +1 | cs.CL | 2026-09-02 |
| #8 | Blending Concepts: Benchmarking Visual Metaphor Generation in Text-to-Image Models | Chuer Chen, Zichen Wang, Yi He +2 | cs.CV | 2026-09-02 |
| #9 | Debias-SparseGPT: Bias-Aware Pruning for Large Language Models | Irina Proskurina, Guillaume Metzler, Antoine Gourru +1 | cs.CL | 2026-09-02 |
| #10 | DeepAffinity: Long-Term Aspect Preference Prediction in eCommerce using Small Language Models | Yotam Eshel, Guy Hadad, Guy Feigenblat +3 | cs.LG | 2026-09-02 |
| #11 | PolERo: Studying Political Evasion in Romanian | Gabriel Stefan, Sergiu Nisioi | cs.CL | 2026-09-02 |
| #12 | The Missing Temporal Link: Temporal Context Routing for Script-Driven Audio-Video Generation | Yichen Liu, Quanwei Zhang, Haozhe Wang +7 | cs.MM | 2026-09-02 |
| #13 | Structured-Prior-Guided Diffusion Inpainting with Physical Consistency for Traffic Sign Augmentation | Luo Li, Chongchong Huang, Jun Jia +4 | cs.CV | 2026-09-02 |
| #14 | SonicCaps: Large-Scale Diverse and Fine-Grained Captioning for Improved Audio-Retrieval | Zineb Lahrichi, Marc Ferras, Gaël Richard +1 | cs.SD | 2026-09-02 |
| #15 | SEAL: Reinforcing Global Safety in Mixture-of-Experts through Shared Expert ALignment | Qingyu Meng, Yiwei Zha, Jiahuan Pei +3 | cs.LG | 2026-09-02 |
| #16 | Do Large Language Models Capture the Diversity in their Training Data? | Youqi Wu, Farzan Farnia | cs.CL | 2026-09-02 |
| #17 | T2LSC-Bench: Benchmarking Localized Semantic Control in Text-to-Image Generation | Yan Wang, Xinyi Hou, Weiguo Lin +2 | cs.CV | 2026-09-02 |
| #18 | APEx: Distillation of Agent Procedural Experience for Adaptive Deep Research Question Answering | Jie Ding, Rui Sun, Xinyuan Zhang +2 | cs.AI | 2026-09-02 |
| #19 | LLM-as-a-Judge Is Not an Oracle: Why Self-Improving Agents Need Deterministic Guardrails | Vansh Wahi | cs.AI | 2026-09-02 |
| #20 | Task-Level Natural Language Priors as Learning Signals for Low-Resource LLM Training | Jian Gao, Xiao Zhang, Xun Zhu +2 | cs.AI | 2026-09-02 |
| #21 | PhoenixNest-Video: Evidence-Grounded Multimodal Agent Framework for Automated Video Interview Assessment | Fan Yuxuan, Huang Miaojun, Zhang Haimei +2 | cs.AI | 2026-09-02 |
| #22 | ASCII Attack: Recontextualising Harmful Requests as Artistic Critique in Large Language Models | Da Cheng Gu, Yifei Dong, Xinghao Yang +2 | cs.AI | 2026-09-02 |
| #23 | Lightweight Adaptation of General-Purpose VLMs for Multispectral and SAR Image Understanding | Shanji Liu, Kelu Yao, Junxiao Xue +5 | cs.CV | 2026-09-02 |
| #24 | DMRL: Document-Mediated Reinforcement Learning for Skill Optimization in Advertising Recommendation | Wei Zhang, Hongji Li, Song Sun +4 | cs.LG | 2026-09-02 |
| #25 | OmegaUse-SOP: SOP Engineering for Professional Computer Use from Human Demonstrations | Yixiong Xiao, Lang An, Hucheng Yang +9 | cs.HC | 2026-09-02 |
| #26 | C$^{3}$T: Counterfactual Causal Reasoning for Sentiment Shifts in Social-Media Conversation Trees | S M Rafiuddin, Atriya Sen | cs.CL | 2026-09-02 |
| #27 | text2ql: Multi-Target Natural Language Querying via a Language-Agnostic Intermediate Representation | Ritesh Kumar | cs.CL | 2026-09-02 |
| #28 | Evidence-Guided Detection, Localization and Explanation for Text-Centric Image Forensics | Peifeng Liu, Bin Li, Qingsong Zhang +3 | cs.CV | 2026-09-02 |
| #29 | Compositional Spectral Prompts for LLM-based Online Time Series Forecasting | Seungyoon Choi, Hyunchul Kim, Jae-Gil Lee +1 | cs.LG | 2026-09-02 |
| #30 | ToolGate: An Executable Acceptance Pipeline for Tool-Dependent Scientific Benchmark Construction | Ke Zhang, Yankang Liu, Roya Zandi +1 | cs.AI | 2026-09-02 |
| #31 | Test-Time Logit Prompting for Source-Free Missing Modality Adaptation | Taixi Chen, Nancy Guo | cs.CV | 2026-09-02 |
| #32 | Benchmarking Language Models for Statistical Problem Formulation | Chen Wang, Junzhe Zhao, Xin Cong +2 | cs.AI | 2026-09-02 |
| #33 | GAPS: Dimension-Level Gates for Conditional Activation Steering | Moghis Fereidouni, Muhammad Umair Haider, Hassan Sajjad +1 | cs.CL | 2026-09-01 |
| #34 | D-FROST: Decentralized Federated pRompt-tuning via Optimal tranSporT for Non-IID and Imbalanced Data | Quan Minh Nguyen, Hoang M. Ngo, Trong Nghia Hoang +1 | cs.LG | 2026-09-01 |
| #35 | How Do Prompt Variations Affect Energy Consumption in On-Device LLMs? | Wei Hu, Xiaolong Tu, Dawei Chen +3 | cs.CL | 2026-09-01 |
| #36 | HEAT: Faster Fully Homomorphic Inference via Approximations-Weights Co-Adaptation | Alessandro Zirilli, Davide Marincione, Evgenios M. Kornaropoulos +2 | cs.CR | 2026-09-01 |
| #37 | SpatialGuard: Harness-Guided Verifiable Spatial Reasoning for Text-to-Image Generation | Ziyun Qian, Zizhi Chen, Yizhou Liu +3 | cs.CV | 2026-09-01 |
| #38 | Closing Cost-Quality Gap in Document VLMs: Difficulty-Aware Data Curation and Quality-Adjusted Deployment Economics | Maksim Evdokimov, Matvey Ivanov, Dmitrii Tsiupin +3 | cs.CL | 2026-09-01 |
| #39 | Selective Agent Guidance via Entropy: Learning Autonomous Policies from Imperfect VLM Teachers | Giovanni Bonetta, Matteo Merler, Davide Zago +2 | cs.AI | 2026-09-01 |
| #40 | From Confusion to Clarity: Confusion-Aware Retrieval and Knowledge Injection for Text Classification | Manish Gupta, Chaitanya Giri, Jayasimha Talur | cs.CL | 2026-09-01 |
| #41 | SDARE-Bench: Evaluating Large Language Models on Conversational Stigma Detection and Response in Dyadic and Group Dialogue | Stephanie Fong, Yiwen Jiang, Zimu Wang +12 | cs.CL | 2026-09-01 |
| #42 | When Guardrails Look Effective: Construct Validity Failures in LLM Agent Commerce Evaluation | Peiying Zhu, Sidi Chang | cs.AI | 2026-09-01 |
| #43 | RadMatch: Auditable Radiology Report Evaluation via Finding-Level Matching | Charles Corbière, Léo Machado, Aubin Charley +3 | cs.CV | 2026-09-01 |
| #44 | Parsing the Stream: A Live Trace Model for Long-Horizon Agents and Their Observers | Egor Pakhomov, Erik Nijkamp | cs.AI | 2026-09-01 |
| #45 | Gaussian Core LoRA: Distribution-Aware Dynamic Adaptation for Broad Concept Erasure | Qinghui Gong, Xunlei Chen, Yu-Xuan Zhang +2 | cs.CV | 2026-09-01 |
| #46 | MegaStyle++: Scaling Image Style Space through Hierarchical Style Definition | Junyao Gao, Sibo Liu, Jiaxing Li +4 | cs.CV | 2026-09-01 |
| #47 | Evaluating Multimodal LLMs as Generalist Vision-Language-Action Agents for Drone Control: Commanding, Approaching, Tracking and Searching | Jaewoo Park, Minyoung Lee, Sukmin Seo +11 | cs.RO | 2026-09-01 |
| #48 | EDGE: Error Dependency Graph-Guided Multi-Error Attribution in Multi-Agent LLM Systems | Jun Hou, Priya Pitre, Yi Fang +1 | cs.AI | 2026-09-01 |
| #49 | Bandits in Prod: Hyperparameter Optimization at Inference Time | Louis Abraham, Tuan-Anh Nguyen, Nicolas Devatine | cs.LG | 2026-09-01 |
| #50 | Automated Event Log Generation from Unstructured Text Using Finetuned LLMs | Maximilian Seeth, Gabriel Marques Tavares, Daniel Schuster | cs.AI | 2026-09-01 |