| 1 | Towards Trustworthy Autonomous Robots: An Explainable AI-Based Decision Framework | Cagri Temel | cs.RO | 2026-09-02 |
| 2 | PlantC2USeg: Cross-Scale Consistent Pre-Training for Few-Shot Unified Plant Point Cloud Segmentation | Yu Tian, Xintong Jiang, Jan Franklin Adamowski +2 | cs.CV | 2026-09-02 |
| 3 | CodePoisonRAG: Knowledge Poisoning Attacks on Retrieval-Augmented Code Generation | Varun Gadey, Ziad Marey, Alexandra Dmitrienko | cs.CR | 2026-09-02 |
| #4 | HyperStyler: Low-resource Authorship Style Transfer via Context-aware Style Navigation and Hypernetworks | Jongkyung Shin, Minguk Jeon, Chanwoo Park +1 | cs.CL | 2026-09-02 |
| #5 | Bilevel Coordinated Reflection: A Game-Theoretic Approach to Multi-Agent LLM Systems | Yihang Chen, Yuxiang Chen, Yuxuan Huang +3 | cs.AI | 2026-09-02 |
| #6 | RVSD: Retrieval Vision Sparse Decoding for Mitigating Visual Hallucinations in Large Vision-Language Models | Canjie Liu, Jiawen Kang, Jinbo Wen +1 | cs.CV | 2026-09-02 |
| #7 | A Top-Down Framework for Metric-Scale Athlete Localization from Single Broadcast Frames | Thanh-Khoi Nguyen, Hoang-Phuc Nguyen, Linh-Huynh +1 | cs.CV | 2026-09-02 |
| #8 | From Tokens to Semantics: Leveraging Complementary Signals for Hallucination Detection in Black-Box LLMs | Urja Pawar, Rajitha Ramanayake, Owen O'Neill +4 | cs.CL | 2026-09-02 |
| #9 | Query Rewriting for Complex Object Segmentation in 4D Gaussian Representations | Thanh-Khoi Nguyen, Thien-Phuc Tran, Minh-Triet Tran | cs.CV | 2026-09-02 |
| #10 | Characterizing Text Branch Sensitivity in Medical Vision-Language Segmentation via Evidence Decoupling | Ziquan Liu, Zhewei Zhu, Xuyang Shi | cs.CV | 2026-09-02 |
| #11 | Learning to Attract and Repel: Dual Quality Margin Learning for Face Recognition (DQM-Face) | El Ouanas Belabbaci, Bhavesh Wani, Philipp Terhörst | cs.CV | 2026-09-02 |
| #12 | Automated Vulnerability Injection in Smart Contracts Using Large Language Models | Luca Migliaccio, Roberto Natella, Naghmeh Ivaki +2 | cs.SE | 2026-09-02 |
| #13 | AffectDelta: Beyond Emotion Labels for Image Editing | Xingzu Zhan, Lin Gu, Ruogu Fang | cs.CV | 2026-09-02 |
| #14 | Deeply Interleaved Text-Image Contexts for Multimodal LLMs Assessment | Zihao Wang, Xi Xiang, Yuwen Sun +5 | cs.CV | 2026-09-02 |
| #15 | TrajMind: Chaining Role-Specialized LoRAs for Fast-and-Slow Collective Trajectory Anomaly Diagnosis | Jiahao Wu, Zhenqun Yang, Chen Jason Zhang +1 | cs.LG | 2026-09-02 |
| #16 | Rethinking the Teacher-Student Framework for Test-Time Adaptation | Damian Sójka, Marc Masana, Bartłomiej Twardowski +1 | cs.LG | 2026-09-02 |
| #17 | ViSAR: Training-Free Adaptive-$k$ Retrieval for Visual Document Question Answering | Adrien Mialland, Marc Plantevit, Julien Gallois +1 | cs.IR | 2026-09-02 |
| #18 | CACTUS: Mask-Guided Semantic Clean-Label Backdoors in Decentralized Federated Learning | Chao Feng, Burkhard Stiller | cs.LG | 2026-09-02 |
| #19 | When Decodability Is Not Enough: Logical Validity Representations, Behavioral Dissociation, and Causal Tests in Language Models | Smitha Muthya Sudheendra, Jaideep Srivastava | cs.CL | 2026-09-02 |
| #20 | ProSR: Semantic-Prototype-Guided Discrete Modeling for Physically Consistent SAR Super-Resolution | Byoungwoo Kim, Munchurl Kim | cs.CV | 2026-09-02 |
| #21 | LookStep: Efficient Vision-Language Navigation with Linguistic Foresight and Event Driven Memory | Kun-Yang Yu, Yingzhe Li, Hongyu Xu +8 | cs.CV | 2026-09-02 |
| #22 | Structured-Prior-Guided Diffusion Inpainting with Physical Consistency for Traffic Sign Augmentation | Luo Li, Chongchong Huang, Jun Jia +4 | cs.CV | 2026-09-02 |
| #23 | SonicCaps: Large-Scale Diverse and Fine-Grained Captioning for Improved Audio-Retrieval | Zineb Lahrichi, Marc Ferras, Gaël Richard +1 | cs.SD | 2026-09-02 |
| #24 | SALA: Semantic-Aware Logical Alignment for Complex Reasoning in In-Context Learning | Zhao Ji, Wenqing Chen, Zhixuan Chu +4 | cs.AI | 2026-09-02 |
| #25 | CrashDiffuser: VLM-Guided Collision Intent Reasoning for Fine-Grained Safety-Critical Traffic Scenario Generation | Shucheng Zhang, Yuang Zhang, Bingzhang Wang +3 | cs.RO | 2026-09-02 |
| #26 | T2LSC-Bench: Benchmarking Localized Semantic Control in Text-to-Image Generation | Yan Wang, Xinyi Hou, Weiguo Lin +2 | cs.CV | 2026-09-02 |
| #27 | InfraPatch: Cross-Task Targeted Grayscale Patch Attacks on Infrared-Adapted Vision-Language Models | Chengyin Hu, Dingyi Lu, Jiaju Han +5 | cs.CV | 2026-09-02 |
| #28 | PhoenixNest-Video: Evidence-Grounded Multimodal Agent Framework for Automated Video Interview Assessment | Fan Yuxuan, Huang Miaojun, Zhang Haimei +2 | cs.AI | 2026-09-02 |
| #29 | PEARL: Path-Entity Aligned Relational Learning with Contextual Subgraphs for Inductive Knowledge Graph Completion | Yunchi Yang, Longlong Li, Cunquan Qu | cs.AI | 2026-09-02 |
| #30 | TAME: Temporal-Aware Mixture-of-Experts for Text-Video Retrieval | Uicheol Jung, Juyoung Hong, Hojung Kwon +1 | cs.CV | 2026-09-02 |
| #31 | Beyond Modality Harmony: Orthogonal Purification and Topology-Guided MoE for Conflict-Aware Multimodal Recommendation | Jialin Liu, Zhaorui Zhang, Ray C. C. Cheung | cs.IR | 2026-09-02 |
| #32 | OmegaUse-SOP: SOP Engineering for Professional Computer Use from Human Demonstrations | Yixiong Xiao, Lang An, Hucheng Yang +9 | cs.HC | 2026-09-02 |
| #33 | SoK: Where Do Flow Labels Come From? Auditing Label Provenance in Encrypted Traffic Benchmarks | Sizhe Huang, Shujie Yang | cs.NI | 2026-09-02 |
| #34 | AI agents reshape consensus formation in human groups | Lin Chen, Ziyi Liu, Xia Hu +1 | cs.CL | 2026-09-02 |
| #35 | Semantic Signal-Assisted Inspection and Recovery Allocation in Reverse Logistics | Jiani He, Dingyan Shang, Yihua Xu +4 | cs.AI | 2026-09-02 |
| #36 | text2ql: Multi-Target Natural Language Querying via a Language-Agnostic Intermediate Representation | Ritesh Kumar | cs.CL | 2026-09-02 |
| #37 | DocHop: Benchmarking Out-of-domain Multi-hop Reasoning in Information-Dense Documents | Zhuoran Yu, Le Thien Phuc Nguyen, Jaden Park +5 | cs.AI | 2026-09-02 |
| #38 | Test-Time Logit Prompting for Source-Free Missing Modality Adaptation | Taixi Chen, Nancy Guo | cs.CV | 2026-09-02 |
| #39 | How Output Format Confounds Data Quality and Capability in Instruction Tuning | Chengguang Gan, Hanjun Wei, Yunhao Liang +3 | cs.CL | 2026-09-02 |
| #40 | InsightSeg: Reusing Correction Insights for Guideline-Consistent Segmentation | Vanshika Vats, Ashwani Rathee, James Davis | cs.CV | 2026-09-02 |
| #41 | ClaimReceipt: Verifying Evidence Sufficiency and Coverage in Agent Evaluations | Peiying Zhu, Sidi Chang | cs.AI | 2026-09-02 |
| #42 | Aggregating Neighbor Embedding Projection and Rank-Based Manifold Learning for Image Retrieval | Vinicius Atsushi Sato Kawai, Gustavo Rosseto Leticio, Lucas Pascotti Valem +1 | cs.CV | 2026-09-02 |
| #43 | Import What You Need: Learning When and How to Augment EHR Graphs with External Knowledge | Chen Chen, Mohsen Nayebi Kerdabadi, Dongjie Wang +2 | cs.LG | 2026-09-01 |
| #44 | Architecting Conversational Data Systems for Stateless LLM APIs: The Hydration Proxy Pattern | Joseph Axisa | cs.AI | 2026-09-01 |
| #45 | Interpretable Symptom Vectors for Depression in a Large Language Model | Fangyi Zhu, Ajay Subramanian, Allison Constant +3 | cs.CL | 2026-09-01 |
| #46 | CoViT: Instance-Correspondence Contrastive Learning for Vision Transformer | Yisen Wang, Zhirong Wu, Limin Wang | cs.CV | 2026-09-01 |
| #47 | Dictionary-Guided Mutation Operators for Automated HDL Repair | Maisha Mastora, Dean Sullivan | cs.ET | 2026-09-01 |
| #48 | AlphaRAD: Grounded Zero-Shot Classification in Chest Radiology via $α$-Corrected Binary Cross Entropy and Factorized Latent Supervision | Jianzhong You, Yuan Gao, Chris McIntosh | cs.CV | 2026-09-01 |
| #49 | Uncovering Understanding-Generation Synergy in Native Unified Multimodal Models: From Representation, Task to System | Penghao Wu, Haiwen Diao, Weichen Fan +3 | cs.CV | 2026-09-01 |
| #50 | Efficient SWE Agent Benchmarking via Trajectory-Aware Evaluation | Kefeng Duan, Dewu Zheng, Yanlin Wang +7 | cs.SE | 2026-09-01 |