| 1 | EarlyEval: Cheaper Agent Evaluation via Early Outcome Prediction | Yuling Shi, Zhensu Sun, Junsen Dong +3 | cs.CL | 2026-09-02 |
| 2 | MV-dVRK: A Multi-Viewpoint Benchmark for Spatial Surgical Perception | Guido Caccianiga, Sergey Prokudin, Yutong Chen +9 | cs.CV | 2026-09-02 |
| 3 | Trace as State: Reasoning Traces as Conditional States for Long-Context Transformers | Xu Zou, Jie Tang | cs.CL | 2026-09-02 |
| #4 | ProbeMatchDTI: Probe-Driven Multi-Scale Biochemical Pattern Matching for Drug-Target Interaction Prediction | Quan Hao, Mengyue Fan, Zifan Dong +8 | cs.LG | 2026-09-02 |
| #5 | Spectral Initialization and Scheduled Graph Smoothness for Uncertain Knowledge Graph Completion | Md Abrar Jahin, Taufikur Rahman Fuad, Jay Pujara +1 | cs.LG | 2026-09-02 |
| #6 | ViSAR: Training-Free Adaptive-$k$ Retrieval for Visual Document Question Answering | Adrien Mialland, Marc Plantevit, Julien Gallois +1 | cs.IR | 2026-09-02 |
| #7 | Learning to Fuse LLMs with Ontology Rankers for Rare-Disease Diagnosis | Zhaoyang Jiang, Zhizhong Fu, Yunsoo Kim +5 | cs.CL | 2026-09-02 |
| #8 | CivBench: A Long-Horizon Benchmark for Tool-Mediated Agents in Civilization VI | Austin Tudor David Andrews, Liam Wilkinson, Jamie Heagerty +3 | cs.AI | 2026-09-02 |
| #9 | CACTUS: Mask-Guided Semantic Clean-Label Backdoors in Decentralized Federated Learning | Chao Feng, Burkhard Stiller | cs.LG | 2026-09-02 |
| #10 | IFW-BLS: Dual-Robust Broad Learning System with Intuitionistic Fuzzy Wave Loss | Mushir Akhtar, M. Tanveer | cs.LG | 2026-09-02 |
| #11 | Before the Script, Set the Stage: How Worldview Simulation Amplifies Psychologically Grounded Persuasion in Multi-Turn Jailbreaking | Siyu Chen, Haoran Wang, Xiaojian Li +3 | cs.CL | 2026-09-02 |
| #12 | Propose to Learn, Learn to Propose: Evaluability-Aware Assistance under Bounded Rationality | Yifan Zhu, Sammie Katt, Samuel Kaski | cs.AI | 2026-09-02 |
| #13 | Signal or Noise? Auditing Rotation-Induced Saliency Drift in Medical and Aerial Imaging | Khawaja Murad ul Hassan, Mehran Ebrahimi | cs.CV | 2026-09-02 |
| #14 | Do Cantonese-Adapted Language Models Better Predict Cantonese Reading? A Cross-Model Eye-Tracking Evaluation | Ziqi Zhang, Emmanuele Chersoni, Mohammad Momenian | cs.CL | 2026-09-02 |
| #15 | Beyond Context Windows: Persistent Discovery Context for Data-Centric Agents | Jalal Mahmud | cs.AI | 2026-09-02 |
| #16 | Semantic Signal-Assisted Inspection and Recovery Allocation in Reverse Logistics | Jiani He, Dingyan Shang, Yihua Xu +4 | cs.AI | 2026-09-02 |
| #17 | Predict, Don't Iterate: Efficient Adaptive-Length Infilling for Diffusion Language Models | Haobo Xu, Sirui Chen, Yuanchen Bei +5 | cs.CL | 2026-09-02 |
| #18 | CHIME: Credit-Aware Hierarchical Memory Evolution for Long-Horizon Agentic Planning | Yongshi Ye, Tian Lan, Feihu Jiang +7 | cs.AI | 2026-09-02 |
| #19 | DocHop: Benchmarking Out-of-domain Multi-hop Reasoning in Information-Dense Documents | Zhuoran Yu, Le Thien Phuc Nguyen, Jaden Park +5 | cs.AI | 2026-09-02 |
| #20 | HeadWiseKV: Budgeted Per-Head Cache Residency for Hybrid Long-Context Language Models | Renjie Xie, Juncheng Yang, Aoting Hu +4 | cs.AI | 2026-09-02 |
| #21 | InsightSeg: Reusing Correction Insights for Guideline-Consistent Segmentation | Vanshika Vats, Ashwani Rathee, James Davis | cs.CV | 2026-09-02 |
| #22 | ClaimReceipt: Verifying Evidence Sufficiency and Coverage in Agent Evaluations | Peiying Zhu, Sidi Chang | cs.AI | 2026-09-02 |
| #23 | When Agents Implement Systems: A Case Study in Defects, Detection, and Evaluation Rigor | Phanindra Reddy Madduru | cs.AI | 2026-09-02 |
| #24 | OutageDiT: A Generative Foundation Model for Power Outage Forecasting and Scenario Simulation | Yunqin Zhu, Feng Qiu, Yao Xie | cs.LG | 2026-09-01 |
| #25 | Candidate Generation and Definition-Guided Verification for Sentence-Level Depression Symptom Recognition | Weiming Li, Catarina Barata, Miguel Constante +1 | cs.CL | 2026-09-01 |
| #26 | Agents That Model Agents: Five Principles Toward a Theory of Mind for 6G Networks | Hatim Chergui, Carolina Fernández-Martínez, Mehdi Bennis +1 | cs.NI | 2026-09-01 |
| #27 | Allocate Before You Embed: Adaptive Visual Input Allocation for Video Embeddings | Song Jin, Zhongtao Jiang, Chenglei Shen +5 | cs.CV | 2026-09-01 |
| #28 | Dictionary-Guided Mutation Operators for Automated HDL Repair | Maisha Mastora, Dean Sullivan | cs.ET | 2026-09-01 |
| #29 | Pooling and Drift in Delayed Bandits | Melika Baghi | stat.ML | 2026-09-01 |
| #30 | When Can a Machine Trust a Statute? A Survival Certificate for Machine-Extracted Legal Logic | Surya Saka | cs.AI | 2026-09-01 |
| #31 | HEAT: Faster Fully Homomorphic Inference via Approximations-Weights Co-Adaptation | Alessandro Zirilli, Davide Marincione, Evgenios M. Kornaropoulos +2 | cs.CR | 2026-09-01 |
| #32 | SpatialGuard: Harness-Guided Verifiable Spatial Reasoning for Text-to-Image Generation | Ziyun Qian, Zizhi Chen, Yizhou Liu +3 | cs.CV | 2026-09-01 |
| #33 | Gradient-Update Mismatch: Rethinking Conflict-Free Training of Physics-Informed Neural Networks | Jing Xiao, Xinhai Chen, Qinglin Wang +5 | cs.LG | 2026-09-01 |
| #34 | BS: Take the Hint - Interactive Multitracer PET/CT Lesion Segmentation with a Scribble-Conditioned ResEnc U-Net | Marven Sherif, Amgad Elmasry, Youssef Ghazal +1 | cs.CV | 2026-09-01 |
| #35 | When Guardrails Look Effective: Construct Validity Failures in LLM Agent Commerce Evaluation | Peiying Zhu, Sidi Chang | cs.AI | 2026-09-01 |
| #36 | A Sensor-Adaptive Incremental Learning Framework for Artifact Detection in Satellite Precipitation Data | Andres F. Monsalve, Hernan A. Moreno, Christian D. Kummerow | physics.ao-ph | 2026-09-01 |
| #37 | Accurate Reconstruction of Gas Turbine Blade Geometry Using 3D/2D Rigid Registration and CT View Optimization | Hristo Valtchanov, Nicolas Piché, Vladimir Brailovski +3 | cs.CV | 2026-09-01 |
| #38 | Separating Syntax from Language: A Mechanistic Account of Translation in Multilingual LLMs | Mikhail Sonkin, Tanja Baeumel, Daniil Gurgurov +2 | cs.CL | 2026-09-01 |
| #39 | Explore Before Committing: Hypothesis-Guided Search for Deep Research Agents | Ruochen Zhou, Zhengyu Chen, Luan Zhang +3 | cs.CL | 2026-09-01 |
| #40 | EmbodiedSkills: A Unified Framework for Orchestrating, Training, and Deploying VLA Agents | Wei Wang, Wenqiao Zhang, Yutong Lin +14 | cs.RO | 2026-09-01 |
| #41 | Post-Training Science for Supervised Fine-Tuning | Charles O'Neill, Mudith Jayasekara, Harry Partridge | cs.LG | 2026-09-01 |
| #42 | Pre-carved Niches: The Formation Dynamics of Modular Task Partitions in Early LLM Training | Guangqi Li, Yongxin Li | cs.LG | 2026-09-01 |
| #43 | CopyShield: A Cross-Level Benchmark of Copyright Defenses in LLMs | Maryam Alshehyari, Dushyant Singh Chauhan, Samuele Poppi +3 | cs.LG | 2026-09-01 |
| #44 | Superposed Latent Autoencoder | Quanling Zhao, Jiaying Yang, Tianqi Zhang +4 | cs.LG | 2026-09-01 |
| #45 | DNC-IMM: Early Lane-Change Intention Recognition via Neural Calibration Based on Driving Context Information | Woong-Chan Byun, Seung-Hyun Kong | cs.RO | 2026-09-01 |
| #46 | ClinTraceBench: Source-Verifiable Longitudinal Clinical Reasoning over EHR-Derived Dialogues | Huimin Wang, Zhengyi Zhao, Yutian Zhao | cs.CL | 2026-09-01 |
| #47 | Lagged Coupling: Internal Representations Become Readable Before They Become Causal | Xining Xun | cs.CL | 2026-09-01 |
| #48 | QILP-0: Constructing Observational Declarative Twins of Quantum Circuits | Marina de la Cruz Echeandía, César Luis Alonso, Tony Ribeiro +1 | cs.AI | 2026-09-01 |
| #49 | Context-Grounding Gains Are Mediated by Pre-existing Machinery: Auditing GRPO, SFT, and DPO | Prakhar Gupta, Vaibhav Gupta | cs.CL | 2026-09-01 |
| #50 | FLaG: Frequency-Domain Latent-attention Gated Pooling for Token Aggregation | Kewei Li, Rongying Zhang, Xueli Wang +6 | cs.AI | 2026-09-01 |