PaperScope
LIVE · 2026-09-04 05:40 UTC

document 11 papers this week · -25% WoW

papers mentioning "document" in title/abstract · 30d window

Latestcs.CLcs.LGcs.AIcs.CV

Mentions per Day (30d)

Latest Papers

#TitleAuthorsCatDate
1Knowledge Acquisition During Pre-training? Large Language Models Learn Better With Auxiliary ViewsJoseph Lee, Yidi Huang, Dokyoon Kim +2cs.CL2026-09-03
2Efficient Test-Time Adaptation through Human-AI InteractionZora Zhiruo Wang, Apurva Gandhi, Rulin Shao +22cs.AI2026-09-03
3IRWOZ 2.0: A Large Language Model-driven Dialogue Dataset for Industrial Robot ConversationsChen Li, Dimitrios Chrysostomoucs.AI2026-09-03
#4FiMI Banking: A Sovereign Model for Indian Retail BankingNPCI AI Research Team, Aman Kumar, Asit Desai +15cs.AI2026-09-03
#5STAIR (STructure Aware Information Retriever): A novel dataset and LLM based retriever for document structure augmentationVineet Kumar, Meghanadh Pulivarthi, vishwajeet kumar +3cs.AI2026-09-03
#6Semantic Bayesian World ModelsTommaso Sorucs.AI2026-09-03
#7Urban Boundaries, Social Barriers: A Benchmark and Vision-Centric Framework for Mapping Gated Communities and Equity ImplicationsMinwei Zhao, Weiming Zhang, Jiawang Du +4cs.CV2026-09-03
#8OBER+: Continuity-Aware Reporting and Traceable Continuous Improvement in Outcome-Based EducationElakkiya Rajasekarcs.LG2026-09-03
#9Rent-a-RAG: Embedding-Space Watermarks for Auditing Third-Party RAGAlexandr Goultiaev Tolstokorov, Kyriakos Mouratidis, Javad Dogani +1cs.CR2026-09-03
#10Can LLMs Extract Architectural Design Decisions from Source Code Commits? - A Preliminary Exploratory StudyAmey Karan, Rudra Dhar, Mohamed Soliman +1cs.SE2026-09-03
#11Artificial Intelligence for Energy Optimization in Data CentersMohammed Basharath Ullah, Summaiya Unnisa Begum, Mohammed Nadeem Ullahcs.AI2026-09-03
#12Enhancing Financial Question Answering: A Novel Benchmark Dataset of Banks' financial statementsArianna Miola, Bruno Spaccavento, Lorenzo Silotto +2cs.CL2026-09-03
#13KhatianDoc: A Human-Verified Benchmark Diagnosing Multimodal LLM Failure on Bengali Legal Land RecordsTasmiad Hasan, Arafat Zaman Ratul, Sarker Sadman Saalim +3cs.CL2026-09-03
#14How Far Can Synthetic Data Take Thai OCR?Kunat Pipatanakulcs.CL2026-09-03
#15GPS-Bench: A Governance Policy Benchmark for Automating Policy AnalysisLinh Le, Melanie Bui, My Chiffon Nguyen +2cs.AI2026-09-03
#16OCR-EDR: Rendering-Aware Diagnosis and Repair for Closed-Loop OCR ImprovementLinnan Zhao, Kang Liu, Hao Yu +3cs.CV2026-09-03
#17Spruce: Scalable Private Outsourced Retrieval Using Compact EmbeddingsPeichun Hua, Yunming Xiaocs.CR2026-09-03
#18From Zero to Hero: An Open LLM Ecosystem for ArmenianErik Arakelyan, Khatun Avetisyan, Meri Davtyan +5cs.LG2026-09-03
#19B2B Customer Conversion Prediction: A Document Representation, Graph Theory, and CatBoost Driven MethodologyTianqi Wang, Sheikh Shams Azam, Wan Eih Huang +3cs.LG2026-09-03
#20The Analyst in the Prompt: Role, Retrieval, and Memory Biases in LLM Financial AnalysisAhmed Asaad, Amr Mohamed, Yang Zhang +1cs.CL2026-09-02
#21Jina-OCR-v1: Efficient Document Parsing with Speculative Decoding and Dense Verifiable RewardsAlejandro Barón García, Feng Wang, Emilia Garcia Casademont +1cs.CL2026-09-02
#22SLIDEFORGE: An LLM Agent for Controllable Editing of Slides as Structured ArtifactsHaozhen Zheng, Fulin Wang, Tianhu Xiong +6cs.CV2026-09-02
#23IDSPACE: A Novel Document Generator for Reliable Evaluation of Digital Identity Verification Systems [Extended Technical Report]Lulu Xie, Yancheng Wang, Kanchan Chowdhury +3cs.CV2026-09-02
#24SHELF: A Synthetic Harness for Multi-Task Bibliographic BenchmarkingMichael J. Bommaritocs.CL2026-09-02
#25Unifying Conformal Language Tasks with In-Context EnsemblesXiao Shi Huang, Chen-Yuan Lin, Bruce Kuwahara +2cs.CL2026-09-02
#26Towards Trustworthy Autonomous Robots: An Explainable AI-Based Decision FrameworkCagri Temelcs.RO2026-09-02
#27CodePoisonRAG: Knowledge Poisoning Attacks on Retrieval-Augmented Code GenerationVarun Gadey, Ziad Marey, Alexandra Dmitrienkocs.CR2026-09-02
#28Measurement-Driven Sub-Network Selection for On-Premise Retrieval-Augmented Factory AgentsVasileios Rizeakos, Georgios Paisios, Alexandros Machairas +2cs.AI2026-09-02
#29Incremental Pooled LLM Evaluation for Cost-Effective Retrieval Model SelectionMax Nelson, Hanoz Bhathena, Aviral Joshi +1cs.IR2026-09-02
#30From Tokens to Semantics: Leveraging Complementary Signals for Hallucination Detection in Black-Box LLMsUrja Pawar, Rajitha Ramanayake, Owen O'Neill +4cs.CL2026-09-02
#31ViSAR: Training-Free Adaptive-$k$ Retrieval for Visual Document Question AnsweringAdrien Mialland, Marc Plantevit, Julien Gallois +1cs.IR2026-09-02
#32Learning to Fuse LLMs with Ontology Rankers for Rare-Disease DiagnosisZhaoyang Jiang, Zhizhong Fu, Yunsoo Kim +5cs.CL2026-09-02
#33Improving Health Literacy through Lay Summarization of Radiological Reports: An Evaluation of BioNER and Retrieval-Augmented GenerationEgecan Çelik Evgin, İlknur Karadeniz, Olcay Taner Yıldızcs.CL2026-09-02
#34Counter-GEO-Bench: Evaluating Defenses Against Information-Distorting Generative Engine OptimizationBing Zheng, Zongyao Zhao, Wenming Yangcs.IR2026-09-02
#35SkillGLoW: Procedural-Family Skill Consolidation for Self-Improving Agents on Long-Horizon Task StreamsAo Yan, Xin Zhang, Jiawei Du +1cs.AI2026-09-02
#36LeakageBench: Document-Level Leakage Risk for Redacting Personally Identifiable Information in Document ImagesVishnu Prasad Vijaya Kumar, Santhosh Venkatesh, Ivan P. Yamshchikovcs.CV2026-09-02
#37When Optimization Becomes Manipulation: Defending Generative Search against Malicious Generative Engine OptimizationHaozhang Li, Yangguang Shao, Xinjie Lin +3cs.CR2026-09-02
#38DMRL: Document-Mediated Reinforcement Learning for Skill Optimization in Advertising RecommendationWei Zhang, Hongji Li, Song Sun +4cs.LG2026-09-02
#39OBJECTION! Lawyer Agents Mitigate Guilty Bias in Legal Judgment PredictionJaehoon Jeong, Jay-Yoon Leecs.CL2026-09-02
#40DocHop: Benchmarking Out-of-domain Multi-hop Reasoning in Information-Dense DocumentsZhuoran Yu, Le Thien Phuc Nguyen, Jaden Park +5cs.AI2026-09-02
#41Privacy Washing: Detecting Internal Contradictions in Privacy PoliciesThomas Brackincs.CY2026-09-02
#42Epistemic Sybil Resistance: Multiplying AI Agents Without Multiplying EvidenceMarc Baracs.AI2026-09-01
#43Belief-Calibrated Optimization: An Explicit World Model for Agentic OptimizationYuhan Chen, Zhihua Tian, Mahavir Dabas +7cs.AI2026-09-01
#44The Memory Trust Gap: Capability-Dependent Failures in Persistent-Memory AgentsJundong Hu, Shekar Ramachandrancs.AI2026-09-01
#45SSAKG 2.0: An Open-Source Package for Structural Associative Sequence Memory and Context-Based RetrievalPrzemysław Stokłosa, Janusz A. Starzyk, Paweł Raifcs.AI2026-09-01
#46Toward Explainable and Policy-Aware AI for Carbon Credit Price Prediction: A Research Framework for Emerging Carbon MarketsSummaiya Unnisa Begum, Mohammed Nadeem Ullah, Mohammed Abdul Ghani Khancs.LG2026-09-01
#47SpeakPay: Domain-Adaptive LoRA Fine-Tuning of Whisper for Low-Resource Nepali Financial Speech RecognitionBiraj Subedics.CL2026-09-01
#48Closing Cost-Quality Gap in Document VLMs: Difficulty-Aware Data Curation and Quality-Adjusted Deployment EconomicsMaksim Evdokimov, Matvey Ivanov, Dmitrii Tsiupin +3cs.CL2026-09-01
#49A systematic Approach to constructing a Chance-and-Risk Matrix for Semiconductor Supply ChainsEma Salkić, Alexander Fichtl, Philipp Ulrich +3cs.CL2026-09-01
#50LatentPress: Context Compression Beyond Text and VisionZhengze Zhou, Hejian Sangcs.LG2026-09-01

all trends · matching is case-insensitive substring after tokenization