| 1 | Necessary or Sufficient? Evaluating LLM Explanations With Behavioural Evidence | Urja Pawar, Rajitha Ramanayake, Nabeel Kemal +4 | cs.AI | 2026-09-04 |
| 2 | AI for Computational Design Science: A Responsible Human-AI Framework and Case Study on Short-Form Video Safety Surveillance | Wenli Zhang, Jiaheng Xie, Zhihe Pan +3 | cs.AI | 2026-09-04 |
| 3 | CABAL: Multi-Agent Simulacra for Tracing the Effects of Collusive Bidding in Peer Review | Jicheng Zhou, Kemou Li, Kahim Wong +5 | cs.AI | 2026-09-04 |
| #4 | A Hybrid Predictive Ensemble of Machine Learning and Deep Neural Networks for Early Cardiovascular Disease Risk Assessment | Balaji Venkateswaran | cs.AI | 2026-09-04 |
| #5 | TIER: Threat Implicitness Benchmark for Evaluating LLM Safety Behaviors | Thu-Hien Trinh-Thi, Hai-Yen Vong, Thanh-Ha Ung-Dung +1 | cs.CR | 2026-09-04 |
| #6 | Artificial Intelligence in Equity and Crypto Markets: Progress, Profitability Evidence, and the Limits of Automated Investing | Linsen Zhu, Mengqing Cai | cs.AI | 2026-09-04 |
| #7 | Adaptation Interfaces for In-Context Tabular Foundation Models in Time-to-Event Prediction | Minh-Khoi Pham, Luca Cotugno, Dan Cernei +10 | cs.LG | 2026-09-04 |
| #8 | ReCAST: Restoration-aware Cascaded Stage-wise Training for Obfuscated SMS Risk Classification | Jieyun Huang, Yi Shen, Kaikai Zhao +7 | cs.CR | 2026-09-04 |
| #9 | PAPT++: Risk-Aware Adversarial Tuning and Generation for Single Domain Generalization | Zhipeng Xu, De Cheng, Xinyang Jiang +5 | cs.CV | 2026-09-04 |
| #10 | PACE: Propagation-Aware Collaborative Correction for One-Shot Personalized Federated Graph Learning | Ruizhe Huang, Chengran Li, Xiaochuan Shi | cs.LG | 2026-09-04 |
| #11 | When Financial Fine-tuning Fails: A Three-Level Detectability Analysis of Numerical Hallucination in Domain-Adapted Language Models | Xiaodong Li, Peiwei Liu | cs.AI | 2026-09-04 |
| #12 | Refuse without Refusal: A Structural Analysis of Safety-Tuning Responses for Reducing False Refusals in Language Models | Minji Kim, Hyounghun Kim | cs.CL | 2026-09-04 |
| #13 | Model Retirement Creates Reproducibility Risk in Biomedical AI Publications | Nathan Wolfrath, Meghan Conroy, Thomas Kosten +7 | cs.AI | 2026-09-04 |
| #14 | Beyond Code Generation: Reliability, Verification, and Cost Economics in the Agentic Software Development Lifecycle | Happy Bhati | cs.SE | 2026-09-04 |
| #15 | Harness-agnostic detection and immunization of reward hacking in self-evolving language models | Rongxin Yang, Yang Liu, Shang Luo +10 | cs.AI | 2026-09-04 |
| #16 | Choosing the Right Language Mode at Inference Time for Multilingual Reliability | Ekata Mitra, Ameeta Agrawal | cs.CL | 2026-09-04 |
| #17 | Leveraging Imperfect Restoration for Data Availability Attack | Yi Huang, Jeremy Styborski, Mingzhi Lyu +2 | cs.AI | 2026-09-04 |
| #18 | Hidden In Plain Gaze: Gaze Representations as Privacy Controls for Utility and Re-identification Risk in XR | Cory Ilo, Brendan-David John, Doug A. Bowman | cs.CV | 2026-09-04 |
| #19 | PatchBench: Evaluating AI Agents for Vulnerability Patching | Chihao Shen, Jiacheng Li, Aastha Mahajan +3 | cs.CR | 2026-09-03 |
| #20 | More Criticism Does Not Make a Better Review: EquiReview-R | Zexing Zhang, Jichao Li, Tianyang Lei +2 | cs.AI | 2026-09-03 |
| #21 | IndicSafeEval: Safety Robustness of Large Language Models under Multilingual Persuasive Jailbreak Attacks | Saikat Mondal, Mamta, Deeksha Varshney +2 | cs.CL | 2026-09-03 |
| #22 | Rethinking World Models for Safety-Critical Embodied Systems | Kailang Ma, Heye Huang, Inhi Kim +1 | cs.AI | 2026-09-03 |
| #23 | Proactive Service Agents: A Unified Decision Framework, Methods, and Evaluation | Yan Tang, Tingyu Cao, Yuanbo Tang +2 | cs.AI | 2026-09-03 |
| #24 | Genetic Algorithms for Tractable Bayesian Network Fusion via Pre-Fusion Edge Pruning | Pablo Torrijos, José A. Gámez, José M. Puerta +1 | cs.NE | 2026-09-03 |
| #25 | Resolution-Aware Experimental Design under Partial Identifiability | Sofianos Panagiotis Fotias | cs.LG | 2026-09-03 |
| #26 | Towards a Statistical Understanding of Mixture-of-Experts | Siyuan He, Bokai Yang, Jie Hu +2 | stat.ML | 2026-09-03 |
| #27 | SafeRestore: Detector-Relative Risk Certificates for Selective Industrial Image Restoration | Shaoliang Yang, Jun Wang | cs.CV | 2026-09-03 |
| #28 | Mind the Gap: Robustness Risks in PII Detection Systems | Adeel Zafar, Slawomir Nowaczyk | cs.LG | 2026-09-03 |
| #29 | SurgeGen: A Hybrid Generative Diffusion Framework for Storm Surge Scenario Synthesis | Shunan Zheng, John J. Hasenbein | math.DS | 2026-09-03 |
| #30 | Risk and Anomaly Identification for Distribution Network Optimal Operation Based on Reinforcement Learning and Uncertainty Quantification | Ziqi Zhang | cs.LG | 2026-09-03 |
| #31 | Reducing Catastrophic Risk from AI with Systematic Monitoring and Evaluation of Rogue AI Progression | T. Bauer, W. P. Kegelmeyer, E. Begoli +15 | cs.CY | 2026-09-02 |
| #32 | RACE-AIMC: Selective Inference for Heterogeneous Analog In-Memory Accelerators at the Edge | Osama Yousuf, Martin Lueker-Boden | cs.ET | 2026-09-02 |
| #33 | Occupancy-based Quantile Risk Control | Zihao Shi, Huajun Xi, Bingyi Jing +1 | stat.ML | 2026-09-02 |
| #34 | Differentially private federated learning with Byzantine-robust aggregation: A cross-domain framework for secure model training in banking and healthcare systems | Srikumar Nayak | cs.CR | 2026-09-02 |
| #35 | Towards Trustworthy Autonomous Robots: An Explainable AI-Based Decision Framework | Cagri Temel | cs.RO | 2026-09-02 |
| #36 | Evaluating Graph Neural Networks for Change-Criticality Classification in Maritime Navigation Charts | Abhishek Potnis, Jacob Arndt | cs.LG | 2026-09-02 |
| #37 | SafeEvolve: Harness-Policy Co-Evolution from Agent Experience for Safety Alignment | Qinghua Mao, Wanying Qu, Dadi Guo +8 | cs.AI | 2026-09-02 |
| #38 | Bilevel Coordinated Reflection: A Game-Theoretic Approach to Multi-Agent LLM Systems | Yihang Chen, Yuxiang Chen, Yuxuan Huang +3 | cs.AI | 2026-09-02 |
| #39 | Momentum in large-batch training: Polyak enlarges the critical batch size, Nesterov improves data efficiency | Jia-Nan Wang, Zixun Huang, Kairui Li +1 | stat.ML | 2026-09-02 |
| #40 | Eliciting ESG Preferences for Reinforcement Learning-Based Portfolio Optimization | Giovanni Dispoto, Marcello Restelli, Carmine Ventre | q-fin.PM | 2026-09-02 |
| #41 | Predictors of Loneliness in Older Adults Using Multimodal Analysis of Speech and Language | Vinmay Khandode, Sai Karthik Kosuri, Neil K. R. Sehgal +6 | cs.CL | 2026-09-02 |
| #42 | Training seeds and model-selection stability in recommender-system evaluation | Juan Manuel Rodriguez, Oleg Lesota, Antonela Tommasel | cs.IR | 2026-09-02 |
| #43 | Improving Health Literacy through Lay Summarization of Radiological Reports: An Evaluation of BioNER and Retrieval-Augmented Generation | Egecan Çelik Evgin, İlknur Karadeniz, Olcay Taner Yıldız | cs.CL | 2026-09-02 |
| #44 | Privacy Leakage in Federated Learning: Gradient-Based Client Identity Inference and Defenses for Inertial Sensing in Vehicular Edge Networks | Ali Akarma, Toqeer Ali Syed, Muhammad Khan +2 | cs.CR | 2026-09-02 |
| #45 | Privacy-Preserving Topology-Guided Safety for LLM-Based Multi-Agent Systems via Federated Graph Learning | Jinxi Yu, Eric Hanchen Jiang, Levina Li +6 | cs.CR | 2026-09-02 |
| #46 | LeakageBench: Document-Level Leakage Risk for Redacting Personally Identifiable Information in Document Images | Vishnu Prasad Vijaya Kumar, Santhosh Venkatesh, Ivan P. Yamshchikov | cs.CV | 2026-09-02 |
| #47 | DMRL: Document-Mediated Reinforcement Learning for Skill Optimization in Advertising Recommendation | Wei Zhang, Hongji Li, Song Sun +4 | cs.LG | 2026-09-02 |
| #48 | GenCAR: Generative Counterfactual Alignment with Risk-Controlled Selection for Out-of-Distribution Recommendation | Qianqian Wang, Yunshan Li, Jiawen Zeng +2 | cs.IR | 2026-09-02 |
| #49 | SoK: Where Do Flow Labels Come From? Auditing Label Provenance in Encrypted Traffic Benchmarks | Sizhe Huang, Shujie Yang | cs.NI | 2026-09-02 |
| #50 | Semantic Signal-Assisted Inspection and Recovery Allocation in Reverse Logistics | Jiani He, Dingyan Shang, Yihua Xu +4 | cs.AI | 2026-09-02 |