| 1 | Knowledge Acquisition During Pre-training? Large Language Models Learn Better With Auxiliary Views | Joseph Lee, Yidi Huang, Dokyoon Kim +2 | cs.CL | 2026-09-03 |
| 2 | Efficient Test-Time Adaptation through Human-AI Interaction | Zora Zhiruo Wang, Apurva Gandhi, Rulin Shao +22 | cs.AI | 2026-09-03 |
| 3 | IRWOZ 2.0: A Large Language Model-driven Dialogue Dataset for Industrial Robot Conversations | Chen Li, Dimitrios Chrysostomou | cs.AI | 2026-09-03 |
| #4 | FiMI Banking: A Sovereign Model for Indian Retail Banking | NPCI AI Research Team, Aman Kumar, Asit Desai +15 | cs.AI | 2026-09-03 |
| #5 | STAIR (STructure Aware Information Retriever): A novel dataset and LLM based retriever for document structure augmentation | Vineet Kumar, Meghanadh Pulivarthi, vishwajeet kumar +3 | cs.AI | 2026-09-03 |
| #6 | Semantic Bayesian World Models | Tommaso Soru | cs.AI | 2026-09-03 |
| #7 | Urban Boundaries, Social Barriers: A Benchmark and Vision-Centric Framework for Mapping Gated Communities and Equity Implications | Minwei Zhao, Weiming Zhang, Jiawang Du +4 | cs.CV | 2026-09-03 |
| #8 | OBER+: Continuity-Aware Reporting and Traceable Continuous Improvement in Outcome-Based Education | Elakkiya Rajasekar | cs.LG | 2026-09-03 |
| #9 | Rent-a-RAG: Embedding-Space Watermarks for Auditing Third-Party RAG | Alexandr Goultiaev Tolstokorov, Kyriakos Mouratidis, Javad Dogani +1 | cs.CR | 2026-09-03 |
| #10 | Can LLMs Extract Architectural Design Decisions from Source Code Commits? - A Preliminary Exploratory Study | Amey Karan, Rudra Dhar, Mohamed Soliman +1 | cs.SE | 2026-09-03 |
| #11 | Artificial Intelligence for Energy Optimization in Data Centers | Mohammed Basharath Ullah, Summaiya Unnisa Begum, Mohammed Nadeem Ullah | cs.AI | 2026-09-03 |
| #12 | Enhancing Financial Question Answering: A Novel Benchmark Dataset of Banks' financial statements | Arianna Miola, Bruno Spaccavento, Lorenzo Silotto +2 | cs.CL | 2026-09-03 |
| #13 | KhatianDoc: A Human-Verified Benchmark Diagnosing Multimodal LLM Failure on Bengali Legal Land Records | Tasmiad Hasan, Arafat Zaman Ratul, Sarker Sadman Saalim +3 | cs.CL | 2026-09-03 |
| #14 | How Far Can Synthetic Data Take Thai OCR? | Kunat Pipatanakul | cs.CL | 2026-09-03 |
| #15 | GPS-Bench: A Governance Policy Benchmark for Automating Policy Analysis | Linh Le, Melanie Bui, My Chiffon Nguyen +2 | cs.AI | 2026-09-03 |
| #16 | OCR-EDR: Rendering-Aware Diagnosis and Repair for Closed-Loop OCR Improvement | Linnan Zhao, Kang Liu, Hao Yu +3 | cs.CV | 2026-09-03 |
| #17 | Spruce: Scalable Private Outsourced Retrieval Using Compact Embeddings | Peichun Hua, Yunming Xiao | cs.CR | 2026-09-03 |
| #18 | From Zero to Hero: An Open LLM Ecosystem for Armenian | Erik Arakelyan, Khatun Avetisyan, Meri Davtyan +5 | cs.LG | 2026-09-03 |
| #19 | B2B Customer Conversion Prediction: A Document Representation, Graph Theory, and CatBoost Driven Methodology | Tianqi Wang, Sheikh Shams Azam, Wan Eih Huang +3 | cs.LG | 2026-09-03 |
| #20 | The Analyst in the Prompt: Role, Retrieval, and Memory Biases in LLM Financial Analysis | Ahmed Asaad, Amr Mohamed, Yang Zhang +1 | cs.CL | 2026-09-02 |
| #21 | Jina-OCR-v1: Efficient Document Parsing with Speculative Decoding and Dense Verifiable Rewards | Alejandro Barón García, Feng Wang, Emilia Garcia Casademont +1 | cs.CL | 2026-09-02 |
| #22 | SLIDEFORGE: An LLM Agent for Controllable Editing of Slides as Structured Artifacts | Haozhen Zheng, Fulin Wang, Tianhu Xiong +6 | cs.CV | 2026-09-02 |
| #23 | IDSPACE: A Novel Document Generator for Reliable Evaluation of Digital Identity Verification Systems [Extended Technical Report] | Lulu Xie, Yancheng Wang, Kanchan Chowdhury +3 | cs.CV | 2026-09-02 |
| #24 | SHELF: A Synthetic Harness for Multi-Task Bibliographic Benchmarking | Michael J. Bommarito | cs.CL | 2026-09-02 |
| #25 | Unifying Conformal Language Tasks with In-Context Ensembles | Xiao Shi Huang, Chen-Yuan Lin, Bruce Kuwahara +2 | cs.CL | 2026-09-02 |
| #26 | Towards Trustworthy Autonomous Robots: An Explainable AI-Based Decision Framework | Cagri Temel | cs.RO | 2026-09-02 |
| #27 | CodePoisonRAG: Knowledge Poisoning Attacks on Retrieval-Augmented Code Generation | Varun Gadey, Ziad Marey, Alexandra Dmitrienko | cs.CR | 2026-09-02 |
| #28 | Measurement-Driven Sub-Network Selection for On-Premise Retrieval-Augmented Factory Agents | Vasileios Rizeakos, Georgios Paisios, Alexandros Machairas +2 | cs.AI | 2026-09-02 |
| #29 | Incremental Pooled LLM Evaluation for Cost-Effective Retrieval Model Selection | Max Nelson, Hanoz Bhathena, Aviral Joshi +1 | cs.IR | 2026-09-02 |
| #30 | From Tokens to Semantics: Leveraging Complementary Signals for Hallucination Detection in Black-Box LLMs | Urja Pawar, Rajitha Ramanayake, Owen O'Neill +4 | cs.CL | 2026-09-02 |
| #31 | ViSAR: Training-Free Adaptive-$k$ Retrieval for Visual Document Question Answering | Adrien Mialland, Marc Plantevit, Julien Gallois +1 | cs.IR | 2026-09-02 |
| #32 | Learning to Fuse LLMs with Ontology Rankers for Rare-Disease Diagnosis | Zhaoyang Jiang, Zhizhong Fu, Yunsoo Kim +5 | cs.CL | 2026-09-02 |
| #33 | Improving Health Literacy through Lay Summarization of Radiological Reports: An Evaluation of BioNER and Retrieval-Augmented Generation | Egecan Çelik Evgin, İlknur Karadeniz, Olcay Taner Yıldız | cs.CL | 2026-09-02 |
| #34 | Counter-GEO-Bench: Evaluating Defenses Against Information-Distorting Generative Engine Optimization | Bing Zheng, Zongyao Zhao, Wenming Yang | cs.IR | 2026-09-02 |
| #35 | SkillGLoW: Procedural-Family Skill Consolidation for Self-Improving Agents on Long-Horizon Task Streams | Ao Yan, Xin Zhang, Jiawei Du +1 | cs.AI | 2026-09-02 |
| #36 | LeakageBench: Document-Level Leakage Risk for Redacting Personally Identifiable Information in Document Images | Vishnu Prasad Vijaya Kumar, Santhosh Venkatesh, Ivan P. Yamshchikov | cs.CV | 2026-09-02 |
| #37 | When Optimization Becomes Manipulation: Defending Generative Search against Malicious Generative Engine Optimization | Haozhang Li, Yangguang Shao, Xinjie Lin +3 | cs.CR | 2026-09-02 |
| #38 | DMRL: Document-Mediated Reinforcement Learning for Skill Optimization in Advertising Recommendation | Wei Zhang, Hongji Li, Song Sun +4 | cs.LG | 2026-09-02 |
| #39 | OBJECTION! Lawyer Agents Mitigate Guilty Bias in Legal Judgment Prediction | Jaehoon Jeong, Jay-Yoon Lee | cs.CL | 2026-09-02 |
| #40 | DocHop: Benchmarking Out-of-domain Multi-hop Reasoning in Information-Dense Documents | Zhuoran Yu, Le Thien Phuc Nguyen, Jaden Park +5 | cs.AI | 2026-09-02 |
| #41 | Privacy Washing: Detecting Internal Contradictions in Privacy Policies | Thomas Brackin | cs.CY | 2026-09-02 |
| #42 | Epistemic Sybil Resistance: Multiplying AI Agents Without Multiplying Evidence | Marc Bara | cs.AI | 2026-09-01 |
| #43 | Belief-Calibrated Optimization: An Explicit World Model for Agentic Optimization | Yuhan Chen, Zhihua Tian, Mahavir Dabas +7 | cs.AI | 2026-09-01 |
| #44 | The Memory Trust Gap: Capability-Dependent Failures in Persistent-Memory Agents | Jundong Hu, Shekar Ramachandran | cs.AI | 2026-09-01 |
| #45 | SSAKG 2.0: An Open-Source Package for Structural Associative Sequence Memory and Context-Based Retrieval | Przemysław Stokłosa, Janusz A. Starzyk, Paweł Raif | cs.AI | 2026-09-01 |
| #46 | Toward Explainable and Policy-Aware AI for Carbon Credit Price Prediction: A Research Framework for Emerging Carbon Markets | Summaiya Unnisa Begum, Mohammed Nadeem Ullah, Mohammed Abdul Ghani Khan | cs.LG | 2026-09-01 |
| #47 | SpeakPay: Domain-Adaptive LoRA Fine-Tuning of Whisper for Low-Resource Nepali Financial Speech Recognition | Biraj Subedi | cs.CL | 2026-09-01 |
| #48 | Closing Cost-Quality Gap in Document VLMs: Difficulty-Aware Data Curation and Quality-Adjusted Deployment Economics | Maksim Evdokimov, Matvey Ivanov, Dmitrii Tsiupin +3 | cs.CL | 2026-09-01 |
| #49 | A systematic Approach to constructing a Chance-and-Risk Matrix for Semiconductor Supply Chains | Ema Salkić, Alexander Fichtl, Philipp Ulrich +3 | cs.CL | 2026-09-01 |
| #50 | LatentPress: Context Compression Beyond Text and Vision | Zhengze Zhou, Hejian Sang | cs.LG | 2026-09-01 |