| 1 | PlannerForge: LLM Agents for Scenario-Based Testing of Motion Planners in Autonomous Driving | Yuan Gao, Sebastian Müller, Mattia Piccinini +5 | cs.AI | 2026-09-08 |
| 2 | The Rater Ising-Potts Model with LLM-Derived Weights: An Application to Multi-Category Scoring Reliability | Matthias von Davier | stat.AP | 2026-09-08 |
| 3 | Supervised Cross-Modal Feature Alignment for Zero-Wearable Freezing of Gait Detection in Parkinsonism | Aryan Singh, Chandan Biswas | cs.CV | 2026-09-08 |
| #4 | From Glance to Scrutiny: Progressive Distortion Reasoning for Fine-Grained Image Quality Assessment | Aoting Zhang, Mingze Gao, Dongbao Yang +5 | cs.CV | 2026-09-08 |
| #5 | SAM3-O2D2: Zero-Shot Object Out-of-Distribution Detection by Object Class Prompting of the SAM3-Image Model | Lucas Görnhardt, Timo Bartels, Tim Fingscheidt | cs.CV | 2026-09-08 |
| #6 | CALIPER: Clean Scenes Cannot Rank Physical Inference in Pretrained Visual Representations | Aman Mehta, Riya Baviskar | cs.RO | 2026-09-08 |
| #7 | SynthGait-19K: A Physically Grounded Synthetic Video Dataset for Gait Parameter Estimation | Soroush Mehraban, Xin Lei Lin, Vida Adeli +5 | cs.CV | 2026-09-08 |
| #8 | VEX-Bench: Benchmarking LLM Agents for Assessing Exploitability of Software Supply Chain Vulnerabilities | Jiahao Shi, Edward Tsien, Yifeng Di +10 | cs.CR | 2026-09-07 |
| #9 | SAFER-Activities: A Dataset for Smart Assessment of Fall Events and Routine Activities | Diwas Lamsal, Pramod Wickramatilake, Jednipat Moonrinta +2 | cs.CV | 2026-09-07 |
| #10 | Solving the Elastic Wave Equation with Physics-Informed Neural Networks: A Robust and Critical Assessment | Davide Staub, Ben Moseley | cs.LG | 2026-09-07 |
| #11 | The Role of Uncertainty in Assessing the Fairness of Machine Learning Models | Francesca Panero, Ernst C. Wit, Marco Scutari | stat.ML | 2026-09-07 |
| #12 | Do Large Language Models Know What They Don't Know II? A Fully Behavioral, Non-Cognitive Measure of Epistemic Honesty | Ali Şenol, H. Russell Bernard, Huan Liu | cs.AI | 2026-09-07 |
| #13 | CodeTD: Topology of Attention Detects Hallucinations in Code LLMs | Daria Voronkova, Ilya Trofimov, Anton Dmitriev +3 | cs.SE | 2026-09-07 |
| #14 | Bag of Tricks or Bag of Myths? Reducing Modeling Complexity with Task Knowledge in Explainable Suicide Risk Assessment | Shlok Shelat, Shrey Salvi, Souvik Roy +2 | cs.CL | 2026-09-07 |
| #15 | The Profit Alignment Problem: How Profit Mandates Induce Alignment Failures in LLMs | Eric So | cs.AI | 2026-09-07 |
| #16 | A radiographic world model for clinical reasoning and evidence generation | Suyang Xi, Songtao Hu, Shansong Wang +8 | cs.AI | 2026-09-07 |
| #17 | Privacy Leakage from a Thousand Words: Millipixel Location Recovery from Dot Maps | Yuntao Du, Tanishq Pauskar, Hao Wang +2 | cs.CR | 2026-09-07 |
| #18 | FinCUABuild: Can Agents Build Reliable Benchmarks for Dynamic Financial Computer Use? | Jingpu Yang, Fengxian Ji, Jinri Guo +7 | cs.AI | 2026-09-07 |
| #19 | AFID: A Unified Open Framework for Automated Fingermark Identification, Quality Assessment and Feature Extraction | Tim Oblak, Rudolf Haraksim, Peter Peer | cs.CV | 2026-09-07 |
| #20 | BlueprintAgent: Constraint-Triggered Targeted Revisits for Simulation-Ready Generation from Scanned Structural Blueprints | Zhouyuan Xu, Chen Yang, Linhao Wang +2 | cs.CL | 2026-09-07 |
| #21 | From Explicit References to Scene Manifolds: Distributional Fidelity and Realism for Radiance Field Quality Assessment | Saeed Mahmoudpour, Gi-Mun Um, Hyon-Gon Choo +1 | cs.CV | 2026-09-07 |
| #22 | CRISP: Corneal Confocal Microscopy Real-Time Image Stitching Pipeline | Qincheng Qiao, Puli Zhang, Jian Zhou +1 | cs.CV | 2026-09-07 |
| #23 | LANTERN: Language Model Assessment on Noisy and Transformed Tasks for Understanding Error and Robustness Nuances | Vamsi Krishna Kodavali, Rituraj Singh | cs.CL | 2026-09-07 |
| #24 | Probing the Structure and Dynamics of LLM Value Expression through Value Conflicts | Kaicheng Zhang, Jingyi Xiao, Renjun Hu +3 | cs.CL | 2026-09-07 |
| #25 | EmoMed: An Emotionally-Aware Agent for Multimodal Medical Support with Real-Time Information Retrieval | Ivan Nasonov, Nikita Glazkov, Ivan Makovetskiy +4 | cs.AI | 2026-09-07 |
| #26 | EEG-Driven Decoding Framework for Passenger Hazard Perception in Highly Automated Vehicles | Yingkai Yang, Ashton Yu Xuan Tan, Bowen Li +10 | cs.AI | 2026-09-07 |
| #27 | PTCG: Persona-guided Tree-based Counterargument Generation | Eunbeen Son, Yohan Jo, Joonsuk Park +1 | cs.CL | 2026-09-07 |
| #28 | AF-Mamba: Efficient Long-Term Signal Modeling for Early Prediction of Atrial Fibrillation Onset | Yongbin Lee, Ki H. Chon | cs.LG | 2026-09-07 |
| #29 | PCSDiff: Diffusion-Based Bias Correction and Super Resolution Toward Practical Operational Medium-Term Precipitation Forecast | Yuze Sun, Shiyi Wang, Jiancheng Pan +7 | cs.LG | 2026-09-07 |
| #30 | CARDEA: Auditable Reasoning Grounded in Spatial Evidence for End-to-End Coronary Angiography Interpretation | Jia-Jen Lee, Shih-Yen Hou, Kee Koon Ng +2 | cs.CV | 2026-09-07 |
| #31 | When Speech Meets Lips: Interpretable Audio-Visual Synchronization for L2 Pronunciation Assessment | Bowen Yu, Mingyu Huang, Yishen Liu +1 | cs.CV | 2026-09-06 |
| #32 | Simulating the Marginal Green Contribution of AI Modules in a Smart-Agriculture Platform: Evidence from Two Monte Carlo Experiments | Zhaoyang Li, Ruijie Zhang, Zhaoji Sun +1 | cs.AI | 2026-09-06 |
| #33 | Monte Carlo-Based Ex-Ante Assessment of the Green Benefits of an AI-Driven Smart Agriculture Platform in Hainan | Zhaoyang Li, Ruijie Zhang, Zhaoji Sun +1 | cs.AI | 2026-09-06 |
| #34 | SAGE: A Hierarchical Framework for Evaluating Interpretive Literary Quality in Narratives | Tianyu Wang, Nianjun Zhou | cs.CL | 2026-09-06 |
| #35 | Reading Decoder Trajectories: Training-Free Counterfactual Query-Trajectory Reliability for Small-Object Detection | Zhaoning Shi, Bo Ma | cs.CV | 2026-09-06 |
| #36 | SRD-GUARD: A Defense Framework of LLMs via Semantic Rewriting and Joint Multi-Model Scoring for Latent Intent Exposure | Qi Wang, Chengcheng Wan, Jiangtao Wang | cs.CR | 2026-09-06 |
| #37 | Large-Scale Pretraining for Improving Deep Learning-Based Geometric Distortion Correction of Diffusion-Weighted Imaging | Saroj Khanal, Yashawant Kumar Yadav, Kritam Bhattarai +12 | cs.CV | 2026-09-06 |
| #38 | Separating Capability from Confidence: Grounded Dual-State Calibration for GRPO-Trained Medical Vision-Language Models | Yangyang Xie, Ke Hao, Jiaqi Liu +2 | cs.CV | 2026-09-06 |
| #39 | Reliability, validity, and diagnostic evidence for multi-model LLM short-answer scoring | Chunyi Zhao, Chao Li | cs.CL | 2026-09-06 |
| #40 | Who Should Grade My Work? Student Perspectives on Transparent AI-Assisted Writing Assessment in Higher Education | Rayed AlGhamdi | cs.AI | 2026-09-04 |
| #41 | Can Large Language Models Anticipate Behavioral Responses to Social Policies? A Case of Pension Enrollment Prediction among China's Flexible Workers | Yumiao Li, Peixin Liu, Donglin Di +2 | cs.CL | 2026-09-04 |
| #42 | A Hybrid Predictive Ensemble of Machine Learning and Deep Neural Networks for Early Cardiovascular Disease Risk Assessment | Balaji Venkateswaran | cs.AI | 2026-09-04 |
| #43 | A Human-in-the-Loop Framework for AI-Assisted Scoring in Large-Scale Writing Assessment | María Eugenia Curi, Germán Capdehourat, Isabel Amigo +4 | cs.CL | 2026-09-04 |
| #44 | How do LLMs Evaluate Perceived Moral Agency? Investigating Moral Decision-Making in Human-Artificial Agents Interactions | Fernanda Mansilla, Aloysius Tok, Bahia Guellaï +2 | cs.CL | 2026-09-04 |
| #45 | ARIA - An Agentic Framework for Autonomous Testing of Infotainment Systems | António Azevedo, Bruno Lima, João Pascoal Faria | cs.SE | 2026-09-04 |
| #46 | Attention-guided super-resolution of 4D flow MRI in carotid arteries | Ali Mokhtari, Dominik Obrist | physics.med-ph | 2026-09-04 |
| #47 | PRISM-Bench: An Audio-Centric Diagnostic Benchmark for Text-to-Audio-Video Generation | Yuchen Sun, Qian Yang, Jun Wang +4 | cs.MM | 2026-09-04 |
| #48 | Building a research-software catalog with a coding agent: from hackathon prototype to public deployment | Kazuyoshi Yoshimi, Satoshi Terasaki, Gotai Yamada | cs.SE | 2026-09-04 |
| #49 | Retinal OCTA Phenotyping with LLM Reporting for Alzheimer's Disease | Progga Paromita Dutta, Jeba Maliha, Md Rafiul Kabir | cs.CV | 2026-09-04 |
| #50 | Last Translation Benchmark | Vilém Zouhar, Niyati Bafna, Mukund Choudhary +241 | cs.CL | 2026-09-03 |