| 1 | Discriminative World Models for Web Agents | Kelvin Li, Dhruv Pendharkar, Anish Pahilajani +6 | cs.AI | 2026-09-02 |
| 2 | Towards Trustworthy Autonomous Robots: An Explainable AI-Based Decision Framework | Cagri Temel | cs.RO | 2026-09-02 |
| 3 | Post-Training Language Models for Gold-Medal Performance in Coding Competitions | Aleksander Ficek, Sean Narenthiran, Mehrzad Samadi +2 | cs.LG | 2026-09-02 |
| #4 | AI Contextual Measurement for Recovering Individual and Group-Level Effects: Validation Against Survey Measures and an Occupational Application | Wenxin Jiang, Xuyang Wang, Yuxiao Wu | cs.AI | 2026-09-02 |
| #5 | Large Language Models (LLMs) for Telecom Root Cause Analysis (RCA): A Structured Reasoning Framework for Evidence-Grounded Diagnosis | Hao Zhou, Mandar Kulkarni, Hao Chen +3 | cs.AI | 2026-09-02 |
| #6 | frb100-40 After Two Decades: An Optimality Certificate and a Preregistered Search Study | Onur Uğurlu | cs.DM | 2026-09-02 |
| #7 | Dutch Books for Language Models | Isaiah Andrews, Suproteem Sarkar | econ.GN | 2026-09-02 |
| #8 | SafeEvolve: Harness-Policy Co-Evolution from Agent Experience for Safety Alignment | Qinghua Mao, Wanying Qu, Dadi Guo +8 | cs.AI | 2026-09-02 |
| #9 | From Reweighting to Rewriting: Unlocking the Intervention Effects of Influential Samples in Training Data Attribution | Yuzhang Luo, Chenpeng Wang, Jianhui Chen +1 | cs.CL | 2026-09-02 |
| #10 | Measurement-Driven Sub-Network Selection for On-Premise Retrieval-Augmented Factory Agents | Vasileios Rizeakos, Georgios Paisios, Alexandros Machairas +2 | cs.AI | 2026-09-02 |
| #11 | Untangling the Mechanisms of Misleading Context in Medical Question Answering | Robin Linzmayer, Noémie Elhadad | cs.CL | 2026-09-02 |
| #12 | Bilevel Coordinated Reflection: A Game-Theoretic Approach to Multi-Agent LLM Systems | Yihang Chen, Yuxiang Chen, Yuxuan Huang +3 | cs.AI | 2026-09-02 |
| #13 | Repo-To-Skill: Distilling GitHub Repositories Into AI4AI Skills | Jianlyu Chen, Yuyang Hu, Hongjin Qian +8 | cs.AI | 2026-09-02 |
| #14 | HiPoly: a hierarchical polymer-native AI framework for property prediction and generative design | Ge Sun, Gervasio Zaldivar, Yuan Tian +5 | physics.chem-ph | 2026-09-02 |
| #15 | Language Models Can Control Their Own Attention | Namgyu Ho, Huzama Ahmad, Woosung Koh +3 | cs.CL | 2026-09-02 |
| #16 | RVSD: Retrieval Vision Sparse Decoding for Mitigating Visual Hallucinations in Large Vision-Language Models | Canjie Liu, Jiawen Kang, Jinbo Wen +1 | cs.CV | 2026-09-02 |
| #17 | Door-in-the-Face Requests and Refusal Behaviour in Large Language Models | Til Jordan | cs.AI | 2026-09-02 |
| #18 | DKL: Decoupled Knowledge Learning for Instruction-Tuned Language Models | Kushagra Bhushan, Meghanadh Pulivarthi, Sai Krishna Reddy Sathi +7 | cs.CL | 2026-09-02 |
| #19 | From Tokens to Semantics: Leveraging Complementary Signals for Hallucination Detection in Black-Box LLMs | Urja Pawar, Rajitha Ramanayake, Owen O'Neill +4 | cs.CL | 2026-09-02 |
| #20 | Loom: Weaving Diagnostic Strands into Free-Text Consensus via Embedding-Space Reweighting | Ron Begleiter, Katya Egert Berg, Gilad Saban +1 | cs.AI | 2026-09-02 |
| #21 | TaRA: Training-Aware Low-Rank Adaptation Initialization | Taehyeon Kim, Eunhyeok Park | cs.CL | 2026-09-02 |
| #22 | Automated Vulnerability Injection in Smart Contracts Using Large Language Models | Luca Migliaccio, Roberto Natella, Naghmeh Ivaki +2 | cs.SE | 2026-09-02 |
| #23 | Collective creativity in hybrid societies | Mason Youngblood, Katie Mudd, Manuel Anglada-Tort +4 | cs.AI | 2026-09-02 |
| #24 | Competitive Market Behavior of LLMs | Pawel Struski, Jakub Swistak, Inez Okulska +1 | cs.MA | 2026-09-02 |
| #25 | ProbeMatchDTI: Probe-Driven Multi-Scale Biochemical Pattern Matching for Drug-Target Interaction Prediction | Quan Hao, Mengyue Fan, Zifan Dong +8 | cs.LG | 2026-09-02 |
| #26 | Learn from Whoever Is Right: Answer-Verified Multi-Teacher Distillation for Multi-Domain LLMs | Xixiang He, Xingming Li, Baiqi Wu +4 | cs.LG | 2026-09-02 |
| #27 | Fine-Grained Anomaly Perception in Wild UGC-Enhanced Images: A Comprehensive Dataset and Difference-Fusion Framework | Yan Zhong, Gefei Chen, Qiufang Ma +4 | cs.CV | 2026-09-02 |
| #28 | Spectral Initialization and Scheduled Graph Smoothness for Uncertain Knowledge Graph Completion | Md Abrar Jahin, Taufikur Rahman Fuad, Jay Pujara +1 | cs.LG | 2026-09-02 |
| #29 | Blending Concepts: Benchmarking Visual Metaphor Generation in Text-to-Image Models | Chuer Chen, Zichen Wang, Yi He +2 | cs.CV | 2026-09-02 |
| #30 | RINSE: Robust Target-Time Normality Estimation for Zero-Shot Graph Anomaly Detection | Taufikur Rahman Fuad, Md Abrar Jahin, Amir Hussain | cs.LG | 2026-09-02 |
| #31 | ViSAR: Training-Free Adaptive-$k$ Retrieval for Visual Document Question Answering | Adrien Mialland, Marc Plantevit, Julien Gallois +1 | cs.IR | 2026-09-02 |
| #32 | DeepAffinity: Long-Term Aspect Preference Prediction in eCommerce using Small Language Models | Yotam Eshel, Guy Hadad, Guy Feigenblat +3 | cs.LG | 2026-09-02 |
| #33 | CivBench: A Long-Horizon Benchmark for Tool-Mediated Agents in Civilization VI | Austin Tudor David Andrews, Liam Wilkinson, Jamie Heagerty +3 | cs.AI | 2026-09-02 |
| #34 | Addressing Trust in AI Systems through Education: A Didactic Perspective | Pierre Haritz, Hendrik Krone, Thomas Liebig | cs.CY | 2026-09-02 |
| #35 | Scalable Kronecker-Fisher Approximation: Efficient Hessian Analysis for Billion-Parameter Language Models Compression | Viacheslav Yusupov, Daria Cherniuk, Evgeny Frolov | cs.LG | 2026-09-02 |
| #36 | Towards One-for-All Robustness Across a Continuum of Threat Levels | Zhichao Hou, Xiaorui Liu | cs.LG | 2026-09-02 |
| #37 | UTP-Bench: Uncertainty-aware Travel Planning Benchmark | Etcharla Revanth Rao, Priyanshu Karmakar, Shubhojit Mallick +3 | cs.AI | 2026-09-02 |
| #38 | Coverage, Not Targeting: A Structural Regime in Multi-Turn Agent Credit Assignment | Chenyu Zhou, Qiliang Jiang, Shuning Wu +1 | cs.LG | 2026-09-02 |
| #39 | Before the Script, Set the Stage: How Worldview Simulation Amplifies Psychologically Grounded Persuasion in Multi-Turn Jailbreaking | Siyu Chen, Haoran Wang, Xiaojian Li +3 | cs.CL | 2026-09-02 |
| #40 | Evidence for Shared Routing Geometry and Dynamics in Sparse Mixture-of-Experts | Kirill Labzin, Stepan Kulibaba, Artem Dzhalilov +1 | cs.LG | 2026-09-02 |
| #41 | Contrastive Explanations in Quantitative Bipolar Argumentation Frameworks | Xiang Yin, Nico Potyka, Antonio Rago +1 | cs.AI | 2026-09-02 |
| #42 | PolERo: Studying Political Evasion in Romanian | Gabriel Stefan, Sergiu Nisioi | cs.CL | 2026-09-02 |
| #43 | MultiGhostBench: A Multilingual Benchmark for Long-Form LLM-Generated Text Attribution under Distribution Shifts | Matteo Greco, Anudeex Shetty, Andrea Tagarelli +1 | cs.CL | 2026-09-02 |
| #44 | Percolation Dynamics in Optimization : Variance Cascades and Discrete Scale Invariance | Sai Niranjan Ramachandran, Suvrit Sra | cs.LG | 2026-09-02 |
| #45 | Diagnosing with Insights: Structured Analysis of Agent Failures via Behavioral Abstractions | Jiayi Bi, Yanjie Gao, Yuanmin Xie +4 | cs.AI | 2026-09-02 |
| #46 | NE-R1: Enhancing Named Entity Recognition Model via Reinforcement Learning | Meixuan Chen, Hehan Li, Ruizhi Zhao +8 | cs.CL | 2026-09-02 |
| #47 | Towards a Foundational Ontology for Identifying and Resolving Contradictions in Dialogue-based Human-Robot Interactions | Maitreyee Tewari, Michele Persiani | cs.HC | 2026-09-02 |
| #48 | Fair Stable Matching: A Nash Social Welfare Approach | Parth Desai, Rasheed M, Ganesh Ghalme +1 | cs.GT | 2026-09-02 |
| #49 | Subcellularly Resolved Single-Cell Embedding Learning with Transcriptomic data, Protein Structure and Localization Information | Zhen Zhou, Jiachen Li, Yuan Liu +2 | q-bio.GN | 2026-09-02 |
| #50 | AGI Maze Prediction Datasets: A Compact Benchmark for Learning World Dynamics with Transformers | Alexey Potapov | cs.LG | 2026-09-02 |
| #51 | SALA: Semantic-Aware Logical Alignment for Complex Reasoning in In-Context Learning | Zhao Ji, Wenqing Chen, Zhixuan Chu +4 | cs.AI | 2026-09-02 |
| #52 | ORB-SVM : An Innovative Hybrid Framework for Efficient Brain Tumor Detection from MRI Scans | Amirhosein Azarpour | cs.CV | 2026-09-02 |
| #53 | What Is Worth Representing? Representational Empowerment for Continual Model Construction | Fei Dai, Hanqi Zhou, Alison Gopnik +1 | cs.LG | 2026-09-02 |
| #54 | DiffIE: Diffusion-based Open Information Extraction | Konstantin Fedorov, Valentin Malykh | cs.CL | 2026-09-02 |
| #55 | Improving Evaluation Realism with Inference-Time Compute and Deployment Scaffolds | Axel Ahlqvist, Richard Guan, Juan-Pablo Rivera +6 | cs.AI | 2026-09-02 |
| #56 | SEAL: Reinforcing Global Safety in Mixture-of-Experts through Shared Expert ALignment | Qingyu Meng, Yiwei Zha, Jiahuan Pei +3 | cs.LG | 2026-09-02 |
| #57 | SCX Router: Streaming Zero-Shot Model Selection with a Decoder-KV Classifier and a Real-World Task Ontology | Ihor Stepanov, Aleksandr Smechov, Mykhailo Shtopko +2 | cs.AI | 2026-09-02 |
| #58 | VoRTeC: Taming Foundation Flow for One-step Real time Video Compression | Yichong Xia, Qinhong Wu, Qinhong Wu +3 | cs.CV | 2026-09-02 |
| #59 | RouteGraph-Mona: Confusion-Aware Routing Fine-Tuning for Mineral Image Classification | Jierui Li, Zhiyuan Qi, Hao Zhu +7 | cs.CV | 2026-09-02 |
| #60 | Auditory Illusion Benchmark for Large Audio Language Models | Hayoon Kim, Eunice Hong, Kyogu Lee | cs.SD | 2026-09-02 |
| #61 | Do Large Language Models Capture the Diversity in their Training Data? | Youqi Wu, Farzan Farnia | cs.CL | 2026-09-02 |
| #62 | CoMerge: Conflict-Driven Preference Optimization for Multi-Task Model Merging | Mingjie Zheng, Zihao Chen, Wenqing Chen +4 | cs.AI | 2026-09-02 |
| #63 | PaperCompiler: Faithful Paper-to-Code Generation via Repository-Level Specification Compilation | Yunhao Liu, Hong Phuc Pham, Jaehong Yoon | cs.CL | 2026-09-02 |
| #64 | CrashDiffuser: VLM-Guided Collision Intent Reasoning for Fine-Grained Safety-Critical Traffic Scenario Generation | Shucheng Zhang, Yuang Zhang, Bingzhang Wang +3 | cs.RO | 2026-09-02 |
| #65 | Retrosynthesis of Synthetic Media for Explainable AI Provenance Forensics | Yijie Lin, Ching-Chun Chang, Isao Echizen +2 | cs.CR | 2026-09-02 |
| #66 | Codebook Agent: Amortized Topology Design for LLM Multi-Agent Systems | Jinxi Yu, Yubei Li, Eric Hanchen Jiang +6 | cs.AI | 2026-09-02 |
| #67 | APEx: Distillation of Agent Procedural Experience for Adaptive Deep Research Question Answering | Jie Ding, Rui Sun, Xinyuan Zhang +2 | cs.AI | 2026-09-02 |
| #68 | DiffuSearch: How Hybrid Trajectory Planning Benefits from Aligned Objectives in Diffusion and Action Space | Steffen Hagedorn, Aron Distelzweig, Alexandru P. Condurache | cs.RO | 2026-09-02 |
| #69 | SAUF-Net: Structure--Appearance Representation Learning with Uncertainty Feedback for Semi-Supervised Medical Image Segmentation | Qin Lu, Zheyang Jing, Yujie Yang +3 | cs.CV | 2026-09-02 |
| #70 | LLM-as-a-Judge Is Not an Oracle: Why Self-Improving Agents Need Deterministic Guardrails | Vansh Wahi | cs.AI | 2026-09-02 |
| #71 | Task-Level Natural Language Priors as Learning Signals for Low-Resource LLM Training | Jian Gao, Xiao Zhang, Xun Zhu +2 | cs.AI | 2026-09-02 |
| #72 | Propose to Learn, Learn to Propose: Evaluability-Aware Assistance under Bounded Rationality | Yifan Zhu, Sammie Katt, Samuel Kaski | cs.AI | 2026-09-02 |
| #73 | PGPO: Potential-Guided Policy Optimization for Multi-Turn Agentic Tasks | Yuyao Zheng, Haipeng Sun, Junwei Bao +4 | cs.AI | 2026-09-02 |
| #74 | InfraPatch: Cross-Task Targeted Grayscale Patch Attacks on Infrared-Adapted Vision-Language Models | Chengyin Hu, Dingyi Lu, Jiaju Han +5 | cs.CV | 2026-09-02 |
| #75 | PhoenixNest-Video: Evidence-Grounded Multimodal Agent Framework for Automated Video Interview Assessment | Fan Yuxuan, Huang Miaojun, Zhang Haimei +2 | cs.AI | 2026-09-02 |
| #76 | Signal or Noise? Auditing Rotation-Induced Saliency Drift in Medical and Aerial Imaging | Khawaja Murad ul Hassan, Mehran Ebrahimi | cs.CV | 2026-09-02 |
| #77 | SkillGLoW: Procedural-Family Skill Consolidation for Self-Improving Agents on Long-Horizon Task Streams | Ao Yan, Xin Zhang, Jiawei Du +1 | cs.AI | 2026-09-02 |
| #78 | PEARL: Path-Entity Aligned Relational Learning with Contextual Subgraphs for Inductive Knowledge Graph Completion | Yunchi Yang, Longlong Li, Cunquan Qu | cs.AI | 2026-09-02 |
| #79 | ASCII Attack: Recontextualising Harmful Requests as Artistic Critique in Large Language Models | Da Cheng Gu, Yifei Dong, Xinghao Yang +2 | cs.AI | 2026-09-02 |
| #80 | SMart: A Multi-source Multi-phase Time Series Representation Transfer Framework | Fang He, Wang-chien Lee | cs.LG | 2026-09-02 |
| #81 | Schrödinger Bridges on Lie Group Manifolds for Probabilistic Intrinsic Generation | Shizhe Zhang, Mingyang Zhao, Lei Ma | stat.ML | 2026-09-02 |
| #82 | Examining the Vulnerability of Multi-Agent Medical Systems to Human Interventions for Clinical Reasoning | Benjamin C Liu, Dillon Mehta, Rishi Malhotra +7 | cs.AI | 2026-09-02 |
| #83 | FUSE: An Evaluating Framework for Dangerous Capabilities of LLMs | Zhengyi Jin, Ru Zhang, Xiao Chen +5 | cs.AI | 2026-09-02 |
| #84 | GeoSPRINT: Geometric Redundancy-Aware Step Pruning for Inference in Diffusion Trajectories | Arpita Joshi | cs.LG | 2026-09-02 |
| #85 | OBJECTION! Lawyer Agents Mitigate Guilty Bias in Legal Judgment Prediction | Jaehoon Jeong, Jay-Yoon Lee | cs.CL | 2026-09-02 |
| #86 | Beyond Modality Harmony: Orthogonal Purification and Topology-Guided MoE for Conflict-Aware Multimodal Recommendation | Jialin Liu, Zhaorui Zhang, Ray C. C. Cheung | cs.IR | 2026-09-02 |
| #87 | OmegaUse-SOP: SOP Engineering for Professional Computer Use from Human Demonstrations | Yixiong Xiao, Lang An, Hucheng Yang +9 | cs.HC | 2026-09-02 |
| #88 | Online Non-Monotone DR-Submodular Maximization Matching the Offline $0.401$ Factor | Vaneet Aggarwal, Yiyang Lu | cs.LG | 2026-09-02 |
| #89 | A Power Law in Logarithm's Clothing: On the Scalability of Graph-Based Vector Search | Sajad Faghfoor Maghrebi, Navid Eslami, Niv Dayan | cs.DB | 2026-09-02 |
| #90 | EmoStance: Response-Side Affective-Orientation Control for Empathetic Response Generation via Emoji Weak Supervision | Ziyuan Jin, Yuxuan Ge, Zheng Tian | cs.AI | 2026-09-02 |
| #91 | C$^{3}$T: Counterfactual Causal Reasoning for Sentiment Shifts in Social-Media Conversation Trees | S M Rafiuddin, Atriya Sen | cs.CL | 2026-09-02 |
| #92 | Beyond Context Windows: Persistent Discovery Context for Data-Centric Agents | Jalal Mahmud | cs.AI | 2026-09-02 |
| #93 | Semantic Signal-Assisted Inspection and Recovery Allocation in Reverse Logistics | Jiani He, Dingyan Shang, Yihua Xu +4 | cs.AI | 2026-09-02 |
| #94 | text2ql: Multi-Target Natural Language Querying via a Language-Agnostic Intermediate Representation | Ritesh Kumar | cs.CL | 2026-09-02 |
| #95 | Disease Burden over Skin Tone: Decomposing the Dermatology-AI Generalization Gap | Nirajan Kunwor, Sanjaya Poudel, Quoc-Huy Trinh +2 | cs.CV | 2026-09-02 |
| #96 | MeanField Surrogate Modeling for Scalable Runtime Scheduling of Concurrent Heterogeneous AI Inference on Shared GPUs | Youssef Ennouri, Soonhoi Ha | cs.DC | 2026-09-02 |
| #97 | Predict, Don't Iterate: Efficient Adaptive-Length Infilling for Diffusion Language Models | Haobo Xu, Sirui Chen, Yuanchen Bei +5 | cs.CL | 2026-09-02 |
| #98 | Git4Data: Database-Native Version Control for AI Agents | Hongshen Gou, Zuyu Zhang, Yuze Sun +4 | cs.DB | 2026-09-02 |
| #99 | Federated LoRA Adaptation of BiomedCLIP Across Four International Chest X-Ray Cohorts | Sanjaya Poudel, Nirajan Kunwor, Manish Dhakal +2 | cs.LG | 2026-09-02 |
| #100 | READY or Not: Reliable Enterprise Agent Deployment | Veronica Chatrath, Bryan Zhu, Jingxuan Fan +15 | cs.AI | 2026-09-02 |
| #101 | MASkills: Continual Skills Optimization for Multi-Agent LLM Systems | Huaiyuan Yao, Xiaoou Liu, Charles Fleming +2 | cs.AI | 2026-09-02 |
| #102 | Beyond Outcome Gaps: Process-Aware Fairness Diagnosis for LLM-based Multi-Agent Decision Systems | Yiran Zhao, Lu Zhou, Liming Fang +4 | cs.AI | 2026-09-02 |
| #103 | Transfer Safety Awareness for Cross-Modal Safety Drift in Multimodal Large Language Models | Tianqi Xiao, Shiyao Cui, Minghao Zhang +2 | cs.MM | 2026-09-02 |
| #104 | CHIME: Credit-Aware Hierarchical Memory Evolution for Long-Horizon Agentic Planning | Yongshi Ye, Tian Lan, Feihu Jiang +7 | cs.AI | 2026-09-02 |
| #105 | ToolGate: An Executable Acceptance Pipeline for Tool-Dependent Scientific Benchmark Construction | Ke Zhang, Yankang Liu, Roya Zandi +1 | cs.AI | 2026-09-02 |
| #106 | MineTRACE: An Evidence-Grounded Interactive Reasoning System for Mineral Prospectivity | Yiran Zhang, Jinwen Liu, Daniel Su +7 | cs.AI | 2026-09-02 |
| #107 | DocHop: Benchmarking Out-of-domain Multi-hop Reasoning in Information-Dense Documents | Zhuoran Yu, Le Thien Phuc Nguyen, Jaden Park +5 | cs.AI | 2026-09-02 |
| #108 | Monitoring Web Agents Without Internal Signals: Observable Trajectories and Key-Step Supervision | Sitong Pan, Yipeng Shen, Yilin Lu +3 | cs.AI | 2026-09-02 |
| #109 | Modeling What Changes: Sparse, Residual World Models for Object-Centric Manipulation | Param Thakkar, Parsika Paresh Shah, Manisha Sushant Gote | cs.RO | 2026-09-02 |
| #110 | HeadWiseKV: Budgeted Per-Head Cache Residency for Hybrid Long-Context Language Models | Renjie Xie, Juncheng Yang, Aoting Hu +4 | cs.AI | 2026-09-02 |
| #111 | Seed-Anchored Budget-Bounded Graph Rendering for Question Answering on Industry-Standard Power-Grid Information and Exchange Models | Jayakumar Manoharan, Yamini Sehgal | eess.SY | 2026-09-02 |
| #112 | InstEditSeg: Instruction-Driven Image Editing for Polyp and Skin Lesion Segmentation | Ziquan Liu, Zhewei Zhu, Xuyang Shi | cs.CV | 2026-09-02 |
| #113 | InsightSeg: Reusing Correction Insights for Guideline-Consistent Segmentation | Vanshika Vats, Ashwani Rathee, James Davis | cs.CV | 2026-09-02 |
| #114 | ClaimReceipt: Verifying Evidence Sufficiency and Coverage in Agent Evaluations | Peiying Zhu, Sidi Chang | cs.AI | 2026-09-02 |
| #115 | When Agents Implement Systems: A Case Study in Defects, Detection, and Evaluation Rigor | Phanindra Reddy Madduru | cs.AI | 2026-09-02 |
| #116 | Benchmarking Language Models for Statistical Problem Formulation | Chen Wang, Junzhe Zhao, Xin Cong +2 | cs.AI | 2026-09-02 |
| #117 | Knowing Is Not Enough: Information Retrievability as a Precondition to Effective LLM Oversight | Xinyu Fu, Narayan Ramasubbu, Dennis Galletta | cs.HC | 2026-09-02 |
| #118 | Post-Training Ternarization of Qwen3-4B Capability, Effective Bit Budget, Storage Compression, and Deployment | Anirudh Malik, M Sparsh Mehra, Poojith Devan | cs.AI | 2026-09-02 |
| #119 | Convergence Theory of Knowledge Distillation in Asynchronous P2P Gossip Learning Network | Lucas Qingyang Fang, Tiyao Liu, Jinhao Jing +4 | cs.LG | 2026-09-01 |
| #120 | On-Policy Distillation Meets Off-Policy GRPO: Training Compact Instruction-Following Rerankers | Vignesh Prabhakar, Jialing Pan, Anil Babu Ankisettipalli | cs.LG | 2026-09-01 |
| #121 | Sparse Readout Prism: Explaining Logit-Lens Scores in Features Instead of Tokens | Matteo He, William F. Shen, Xinchi Qiu +1 | cs.CL | 2026-09-01 |
| #122 | Looped Transformers under the Jacobian Lens: Does the Global Workspace Survive Recurrence? | Wenlong Wang, Fergal Reid | cs.AI | 2026-09-01 |
| #123 | The Ceiling Is in the Channel: Auditing Learner Gaps and Measurement Frontiers in Clinical Prediction | Sayeed Shafayet Chowdhury, Nusrat Jahan, Snehasis Mukhopadhyay +2 | cs.AI | 2026-09-01 |
| #124 | Accurate in space, unreliable in time: how LLMs represent national cultural change | Yalda Daryani, Miranda Bogen, Madeleine I. G. Daepp | cs.CY | 2026-09-01 |
| #125 | OutageDiT: A Generative Foundation Model for Power Outage Forecasting and Scenario Simulation | Yunqin Zhu, Feng Qiu, Yao Xie | cs.LG | 2026-09-01 |
| #126 | Epistemic Sybil Resistance: Multiplying AI Agents Without Multiplying Evidence | Marc Bara | cs.AI | 2026-09-01 |
| #127 | Thinking effort aligns between humans and reasoning models in abductive reasoning | Henry Arthur | cs.CL | 2026-09-01 |
| #128 | Belief-Calibrated Optimization: An Explicit World Model for Agentic Optimization | Yuhan Chen, Zhihua Tian, Mahavir Dabas +7 | cs.AI | 2026-09-01 |
| #129 | The Memory Trust Gap: Capability-Dependent Failures in Persistent-Memory Agents | Jundong Hu, Shekar Ramachandran | cs.AI | 2026-09-01 |
| #130 | SSAKG 2.0: An Open-Source Package for Structural Associative Sequence Memory and Context-Based Retrieval | Przemysław Stokłosa, Janusz A. Starzyk, Paweł Raif | cs.AI | 2026-09-01 |
| #131 | Import What You Need: Learning When and How to Augment EHR Graphs with External Knowledge | Chen Chen, Mohsen Nayebi Kerdabadi, Dongjie Wang +2 | cs.LG | 2026-09-01 |
| #132 | Agent Memory Is a Surface for Endogenous Authorization Laundering | Tommaso Cerruti, Mika Okamoto, Ansel Kaplan Erol | cs.CR | 2026-09-01 |
| #133 | Architecting Conversational Data Systems for Stateless LLM APIs: The Hydration Proxy Pattern | Joseph Axisa | cs.AI | 2026-09-01 |
| #134 | Interpretable Symptom Vectors for Depression in a Large Language Model | Fangyi Zhu, Ajay Subramanian, Allison Constant +3 | cs.CL | 2026-09-01 |
| #135 | Zeta-Lite: A Concurrent, Branchable In-Browser SQL Database for Agentic Memory | Gene Zhang | cs.DB | 2026-09-01 |
| #136 | Induction and Inquiry via Probabilistic Reasoning over Language and Code | Wasu Top Piriyakulkij, Sam Acquaviva, Cassidy Langenfeld +2 | cs.AI | 2026-09-01 |
| #137 | When Does Information Sharing Improve Decentralized Discovery? Aggregation, Independent Rescue, and Equilibrium Selection | Yohei Nakajima | cs.AI | 2026-09-01 |
| #138 | hLLM: Single Pass Decoding for Generative Reranking | Emil Laftchiev, Prachi Agrawal, Moe Kayali +7 | cs.LG | 2026-09-01 |
| #139 | VakyArth: Evaluating Pragmatic Competence in LLMs across Indic Languages | Usneek Singh, Poorvaja Veera Balaji Kumar, Parth Nanda +4 | cs.CL | 2026-09-01 |
| #140 | Agents That Model Agents: Five Principles Toward a Theory of Mind for 6G Networks | Hatim Chergui, Carolina Fernández-Martínez, Mehdi Bennis +1 | cs.NI | 2026-09-01 |
| #141 | Dictionary-Guided Mutation Operators for Automated HDL Repair | Maisha Mastora, Dean Sullivan | cs.ET | 2026-09-01 |
| #142 | Swin Meets EfficientNet: Lightweight Architectures for GAN-Based Face Forensics | Sejuti Basu, Ashima Sood, Vijay Kumar +1 | cs.CV | 2026-09-01 |
| #143 | When Can a Machine Trust a Statute? A Survival Certificate for Machine-Extracted Legal Logic | Surya Saka | cs.AI | 2026-09-01 |
| #144 | Harness Engineering in LLM Tool Use via Agent-Native Reusable Tool Primitives | Haibo Jin, Suijin Wang, Xucheng Yu +2 | cs.SE | 2026-09-01 |
| #145 | HEAT: Faster Fully Homomorphic Inference via Approximations-Weights Co-Adaptation | Alessandro Zirilli, Davide Marincione, Evgenios M. Kornaropoulos +2 | cs.CR | 2026-09-01 |
| #146 | RecKAN: Kolmogorov-Arnold Networks with a Learnable Recursive Polynomial Basis | Amirhosein Azarpour | cs.LG | 2026-09-01 |
| #147 | Efficient SWE Agent Benchmarking via Trajectory-Aware Evaluation | Kefeng Duan, Dewu Zheng, Yanlin Wang +7 | cs.SE | 2026-09-01 |
| #148 | Adaptive Critical Token-Aware Retrieval for Repository-Level Code Generation | Kefeng Duan, Dewu Zheng, Yanlin Wang +8 | cs.SE | 2026-09-01 |
| #149 | CordisBench: Can Language Models Reason About Component Lifecycles in Dynamic Agent Harnesses? | Damien Sileo, Dimitri Kachler | cs.CL | 2026-09-01 |
| #150 | The Rise of Verbal Reinforcement Learning | Kshitij Tayal, Arun Sharma, Genta Indra Winata +2 | cs.CL | 2026-09-01 |
| #151 | Mechanism Design for Alignment and Control | Dirk Bergemann, Andrew Koh, Stephen Morris | econ.TH | 2026-09-01 |
| #152 | Designing Proactive Thought Partners for Writing | Chao Zhang, Abe Davis, Chih-Wei Chen +1 | cs.HC | 2026-09-01 |
| #153 | Scaling Near-Optimal SFT-RL Annotation Budget Allocation from Small to Large LLMs | Jingtan Wang, Arun Verma, Xiaoqiang Lin +4 | cs.CL | 2026-09-01 |
| #154 | Selective Agent Guidance via Entropy: Learning Autonomous Policies from Imperfect VLM Teachers | Giovanni Bonetta, Matteo Merler, Davide Zago +2 | cs.AI | 2026-09-01 |
| #155 | From Confusion to Clarity: Confusion-Aware Retrieval and Knowledge Injection for Text Classification | Manish Gupta, Chaitanya Giri, Jayasimha Talur | cs.CL | 2026-09-01 |
| #156 | H3-World: Turning Language Understanding into World Control | Danze Chen, Zeqing Wang, Ziyue Lin +2 | cs.CV | 2026-09-01 |
| #157 | Retrieved but not ranked: surface-form bias in structural retrieval, from mathematics to agent trajectories | Nabira Rashid, Manolis Kellis | cs.LG | 2026-09-01 |
| #158 | BS: Take the Hint - Interactive Multitracer PET/CT Lesion Segmentation with a Scribble-Conditioned ResEnc U-Net | Marven Sherif, Amgad Elmasry, Youssef Ghazal +1 | cs.CV | 2026-09-01 |
| #159 | Can LLMs Discover Scientific Laws in Real and Parallel Worlds? | Yiming Huang, Ziche Liu, Zhuohang Wu +11 | cs.AI | 2026-09-01 |
| #160 | A Mathematical Theory of Reusable Neural Bases for Network Compression | Binshuai Wang | cs.LG | 2026-09-01 |
| #161 | Can LLMs Design Video Coding Tools? A Case Study on Planar Mode | Yingwen Zhang, Meng Wang, Liqiang He +1 | cs.MM | 2026-09-01 |
| #162 | EvoSCM: Scientific Belief Revision Through Causal Model Evolution and Experimentation | Qing Zhao, Haowei Li, Weijian Deng +2 | cs.AI | 2026-09-01 |
| #163 | Relational-Core Graph Analytics Querying graphs at SQL scale, and why the node/edge model is a performance tax, not a truer picture of connected data | Gene Zhang | cs.DB | 2026-09-01 |
| #164 | When Guardrails Look Effective: Construct Validity Failures in LLM Agent Commerce Evaluation | Peiying Zhu, Sidi Chang | cs.AI | 2026-09-01 |
| #165 | TempCloze: Can Video-LLMs Identify the Missing Middle? | Wenqi Pei, Henry Hengyuan Zhao, Yilai Liu +4 | cs.CV | 2026-09-01 |
| #166 | LatentPress: Context Compression Beyond Text and Vision | Zhengze Zhou, Hejian Sang | cs.LG | 2026-09-01 |
| #167 | Public-Sharing Labels and Verbatim Field Egress in an MCP-to-A2A Agent Configuration: A Controlled Multi-Model Study | Arpan Kumar Mahapatra | cs.CR | 2026-09-01 |
| #168 | Optimizing Byzantine Node Placement in Decentralized Federated Learning | Edoardo Gabrielli, Gabriele Tolomei | cs.LG | 2026-09-01 |
| #169 | Rethinking Learnability in Offline Data-driven Optimization | Chao Qian, Chen-Guang Wang, Rong-Xi Tan +1 | cs.LG | 2026-09-01 |
| #170 | GlossoGen: Emergent Language in Complex Multi-Agent LLM Interactions | Elias Stengel-Eskin, Newton Sander, Carlos Bonetti +4 | cs.CL | 2026-09-01 |
| #171 | Defense-as-Skill: Evolving Runtime Guard Skill for Skill-Augmented Agents | Xiaofang Yang, Ziqi Miao, Dianbo Sui +2 | cs.CR | 2026-09-01 |
| #172 | Harness-of-Harness: Multi-Day Autonomous Software Development with Continual Improvement | Haoyang Yan, Min-le Su, Hangfan Zhang +6 | cs.AI | 2026-09-01 |
| #173 | Parsing the Stream: A Live Trace Model for Long-Horizon Agents and Their Observers | Egor Pakhomov, Erik Nijkamp | cs.AI | 2026-09-01 |
| #174 | When Safety Routing Breaks: Understanding Alignment Fragility under Benign Fine-Tuning | Yitong Guo, Xiaoyi Chen, Siyuan Zhang +2 | cs.CR | 2026-09-01 |
| #175 | Efficiently Estimating Optimal Hyperparameter Scaling Laws through Power-Law Entropy Search | Zhiliang Chen, Sebastian Ament, David Eriksson +3 | cs.LG | 2026-09-01 |
| #176 | Learning Sparse Decision Trees via Transformer Variational Auto-Encoders | Giacomo Fidone, Alessio Cascione, Riccardo Guidotti | cs.LG | 2026-09-01 |
| #177 | Semantic-Guided Multimodal Preprocessing for Vision Transformer-Based Clear Cell Renal Cell Carcinoma Grading | Fatemeh Javadian, Zhu Chen, Zahra Aminparast +1 | cs.CV | 2026-09-01 |
| #178 | Provably Safe Sim-to-Real Transfer | Tingting Ni, Maryam Kamgarpour | cs.LG | 2026-09-01 |
| #179 | EdiTikZ: Scientific Figure Editing from Revision Trajectories | Christian Greisinger, Zhixue Zhao, Steffen Eger | cs.AI | 2026-09-01 |
| #180 | Neuro-Symbolic Geometric Abstraction (NeuSOGA): From Observations to Symbolic Mathematical Representations | Qingde Li, Qingqi Hong, Zihan Li +1 | cs.AI | 2026-09-01 |
| #181 | Evaluating Multimodal LLMs as Generalist Vision-Language-Action Agents for Drone Control: Commanding, Approaching, Tracking and Searching | Jaewoo Park, Minyoung Lee, Sukmin Seo +11 | cs.RO | 2026-09-01 |
| #182 | Measuring consistency via ensemble margin and local prediction variability: Auditing decision systems in the presence of predictive multiplicity | Sinjini Banerjee, Tim Marrinan, Anand D. Sarwate | stat.ML | 2026-09-01 |
| #183 | EDGE: Error Dependency Graph-Guided Multi-Error Attribution in Multi-Agent LLM Systems | Jun Hou, Priya Pitre, Yi Fang +1 | cs.AI | 2026-09-01 |
| #184 | PopPert: Population-level Joint-Distribution Modeling for Single-Cell Perturbation Prediction | Handong Wang, Jiaxin Qi, Haochen Feng +1 | q-bio.GN | 2026-09-01 |
| #185 | SymFold: Synergizing Evolutionary and Structural Priors for Accurate Protein Inverse Folding | Handong Wang, Jiaxin Qi, Baisheng Lai +1 | cs.AI | 2026-09-01 |
| #186 | CHARM: Character Hallucination for Multicultural Role Play Benchmark | Sunkyung Han, Nahyeon Park, Gaeun Seo +2 | cs.CL | 2026-09-01 |
| #187 | Scalable Rao-Blackwellized Online Planning for High-Dimensional POMDPs | Jiho Lee, Nisar Ahmed, Kyle Hollins Wray +1 | cs.RO | 2026-09-01 |
| #188 | Meta-ethics and AI: exploring the novel meta-ethical questions in the era of AI | Shang Lu | cs.AI | 2026-09-01 |
| #189 | Cheap Verifiers, Large Blind Spots: Measuring the Reliability Cost of Cost-Saving Cascades | Dushyant Rajput | cs.AI | 2026-09-01 |
| #190 | Probing Factual Knowledge Transfer with Training Data Interventions | Romina Oji, Marc Braun, Marcel Bollmann +2 | cs.CL | 2026-09-01 |
| #191 | LEAP: Likelihood Elicitation and Aggregation for LLM-based Probabilistic Forecasting | Yufei Chen, Yiran Zhao, Xiaogang Xu +3 | cs.AI | 2026-09-01 |
| #192 | Bandits in Prod: Hyperparameter Optimization at Inference Time | Louis Abraham, Tuan-Anh Nguyen, Nicolas Devatine | cs.LG | 2026-09-01 |
| #193 | Automated Event Log Generation from Unstructured Text Using Finetuned LLMs | Maximilian Seeth, Gabriel Marques Tavares, Daniel Schuster | cs.AI | 2026-09-01 |
| #194 | MIDR: Enrichment-Augmented Indexing for Multimodal Document Retrieval | Debanjan Mahata, Atharva Tendle, Daniel Preotiuc-Pietro +2 | cs.IR | 2026-09-01 |
| #195 | A Composable Evaluation System for Reproducible Omni-Modal Foundation Model Evaluation | Hodong Lee, Sanghee Park, Dohoon Ryu +4 | cs.AI | 2026-09-01 |
| #196 | GazeRefine: Expert Gaze as a Test-Time Prompt for Training-Free Medical Image Segmentation | Mohammed Oussama Benyahia, Marouane Tliba, Mohamed Amine Kerkouri +10 | eess.IV | 2026-09-01 |
| #197 | Analog-DB: An Agent-First Analog Integrated Circuit Database, From Blocks to Systems | Danial Noori Zadeh, Mohamed B. Elamien | cs.AI | 2026-09-01 |
| #198 | HiLRP: Toward One Trustworthy Explanation for Vision Transformer: Conservation-Valid Attribution via Attention Primitives | Sathiyamohan Nishankar, Pubudu Sanjeewani, Asanka Perera +1 | cs.CV | 2026-09-01 |
| #199 | EmbodiedSkills: A Unified Framework for Orchestrating, Training, and Deploying VLA Agents | Wei Wang, Wenqiao Zhang, Yutong Lin +14 | cs.RO | 2026-09-01 |
| #200 | Some Emotions Run Deeper: Layer-wise Probing and Causal Intervention in Large Language Models | Tian Fang, Gaël Guibon, Davide Buscaldi | cs.CL | 2026-09-01 |