| 1 | Multiplicative Optimism for Constant Regret in Games | Ashkan Soleymani, Georgios Piliouras | cs.GT | 2026-09-18 |
| 2 | Guiding Agents of Quantum Games to Equilibrium using Matrix Exponential Fixed-Point Iteration | Alireza Habibi, Luis F. Abanto Leon, Setareh Maghsudi | quant-ph | 2026-09-18 |
| 3 | Single-Loop Stochastic Projected Damped Extragradient Methods for Stochastic Nonconvex--(Strongly) Concave Minimax Optimization | Huiling Zhang, Minhao Zhang, Zi Xu | math.OC | 2026-09-18 |
| #4 | GameLogicBench: Evaluating Coding Agents on Runtime Game Logic with Tick-Level State Assertions | Xinyu Che, Yunfei Ge, Shihao Li +9 | cs.SE | 2026-09-18 |
| #5 | MIRAGE: Multi-Perspective Creative Language Model Reasoning with Reinforcement Learning Guidance | Arash Lagzian, Srinivas Anumasa, Dianbo Liu | cs.CL | 2026-09-18 |
| #6 | GameASG-Bench: Benchmarking Autonomous Software Generation for Game Development | Xiuhui Zhang, Yi Chen, Shusheng Xu +4 | cs.AI | 2026-09-18 |
| #7 | SplashSplat: Reconstructing Splashing Liquids from Real-World Multi-View Videos | Peiyu Liu, Dingxi Zhang, Federico Tombari +3 | cs.CV | 2026-09-17 |
| #8 | The First-Order Oracle Complexity of Lipschitz Convex Optimization in Nondual Settings | David Martínez-Rubio, Brian Bullins, Cristóbal Guzmán +1 | math.OC | 2026-09-17 |
| #9 | Mitigating Retaliatory Algorithmic Collusion in Repeated Games | Karthik Sivachandran, Rohan Paleja | cs.LG | 2026-09-17 |
| #10 | A Qualitative Model for Reasoning about Path and Support | Abhishek Jaiswal, Zoe Falomir | cs.AI | 2026-09-17 |
| #11 | Steering Equilibrium Selection in Regularized Self-Play via the Reference Policy | Luis Leal | cs.AI | 2026-09-17 |
| #12 | SIMLIFE: Pattern Understanding for Long-Horizon Human-Agent Partnership | Run Peng, Zinnia Nie, Jing Ding +7 | cs.AI | 2026-09-17 |
| #13 | Agentic AI Networking for Heterogeneous Unmanned Aerial Systems in Low-Altitude Wireless Networks | Nguyen Duc Minh Quang, Chang Liu, Shuangyang Li +1 | cs.AI | 2026-09-17 |
| #14 | For Your Eyes Only: Evaluating Coordination Between Isolated Language Model Instances | Alexander Shirnin, Aleksey Kudelya | cs.CL | 2026-09-16 |
| #15 | Efficient Nash Equilibrium Computation for Cybersecurity Games | Michael Lanier, David Farmer, Yevgeniy Vorobeychik | cs.GT | 2026-09-16 |
| #16 | Flag Game: A Toy Model for Mechanistic Swarm Interpretability | Elizabeth Pavlova, Hidenori Tanaka | cs.AI | 2026-09-16 |
| #17 | Playing log(N)-Questions over Wikipedia Abstracts: Communication Efficiency Between Paired Frontier Models | Peter Potash | cs.CL | 2026-09-16 |
| #18 | Long-Lived Characters, Local Inference: Incremental Memory Maintenance for Game NPCs | Zimu Xu | cs.CL | 2026-09-16 |
| #19 | Clueing up LLMs with Tool-Augmented Deductive Reasoning | Rebecca Ansell, Autumn Toney-Wails | cs.AI | 2026-09-16 |
| #20 | Which LLM is Best for Translating Natural Language Goals to PDDL | Tomas Balyo, Lukas Chrpa, G. Michael Youngblood | cs.AI | 2026-09-16 |
| #21 | PACT: Can Enterprise AI Assistants Be Trusted Under Pressure? | Mika Okamoto, Ansel Kaplan Erol | cs.CL | 2026-09-16 |
| #22 | Online Robust Reinforcement Learning Through Monte-Carlo Planning | Tuan Dam, Kishan Panaganti, Brahim Driss +1 | cs.LG | 2026-09-16 |
| #23 | Recursive Reasoning or Statistical Extrapolation? In-Context Learning in Multi-Agent Interdependent Decision-Making | Yu Liu, Wenwen Li, Yifan Dou +1 | cs.AI | 2026-09-16 |
| #24 | Matching Multi-Loop Complexities with a Single Loop: Optimal Optimization Stationarity and Best-Known Game Stationarity in Nonconvex--Concave Minimax Optimization | Minghao Zhang, Zi Xu | math.OC | 2026-09-16 |
| #25 | Easy to Catch a Liar, Hard to Clear an Honest One: Language Models Diagnosing a Corrupted Reward Channel from a Verified Record | Arman Nik Khah | cs.LG | 2026-09-15 |
| #26 | Constant Swap Regret in General-Sum Games via Optimistic Transition Matrices | Tung Mai | cs.GT | 2026-09-15 |
| #27 | AI for Games in the Foundation Model Era | Meng Luo, Yanlin Li, Hao Li +7 | cs.AI | 2026-09-15 |
| #28 | Symmetric solution of the Bellman optimality equation for repeated harmony game | Hisato Komatsu | cs.GT | 2026-09-14 |
| #29 | LynnReal-Omni: Native multi-modal Video Generation for Agentic Visual Workflows | Xiaofeng Mao, Peijia Lin, Shaohao Rui +3 | cs.CV | 2026-09-14 |
| #30 | Learning to Coach for Experiential Learning | Guanheng Chen, Tianzhu Ye, Li Dong +3 | cs.CL | 2026-09-14 |
| #31 | A Game-Theoretic Framework for Incentive-Compatible AI training Under Renewable-Energy Constraints | Konstantinos Varsos, Ramin Khalili, Adamantia Stamou +2 | cs.ET | 2026-09-14 |
| #32 | Math for AI safety: an invitation for mathematicians | Lionel Levine | math.HO | 2026-09-14 |
| #33 | Improving the Last-Iterate Guarantees of Anytime Algorithms for Stochastic Monotone Variational Inequalities | Jun-Hyun Kim, Ahmet Alacaoglu | math.OC | 2026-09-14 |
| #34 | High-Probability Nash Regret for Decentralized Learning in Markov $α$-Potential Games: Episodic and Fully Online Asynchronous Algorithms with Applications to Markov Congestion Games | S. Rasoul Etesami | cs.LG | 2026-09-14 |
| #35 | Shapley Value Estimation for Multi-Site Data with Blockwise-Missing Features | Siqi Li, Wangxuan Fan, Yiming Li +2 | stat.ML | 2026-09-14 |
| #36 | Pathwise Individual Rationality in Federated Learning: A Mechanism-Architecture Co-Design | Amin Meghrazi, Srinivasan Parthasarathy, Andrew Perrault | cs.LG | 2026-09-13 |
| #37 | TATK: Triple-Aware Top-K Learning with Knowledge-Grounded Verification for LLM-based Sequential Recommendation | Yuchen Guan, Jiaye Liu, Yifei Han +3 | cs.CL | 2026-09-13 |
| #38 | Oops, Not Now: PEARL, a RAG-Based Support Agent for Gameplay and What Players Want from AI Help | Jiahong Li, Sai Siddartha Maram, Atieh Kashani +6 | cs.HC | 2026-09-12 |
| #39 | GPU-CFR: 80x Faster Counterfactual Regret Minimization by Compiling the Game to Static Dataflow and CUDA Graph Replay | Boning Li, Longbo Huang | cs.DC | 2026-09-10 |
| #40 | The Convention Gap: Towards Measuring Implicit Communication in Cooperative AI Evaluation | Makoto Fukushima, Hua-Dong Xiong, Ehsan Moradi Pari | cs.AI | 2026-09-10 |
| #41 | Tapes Together Strong: The Co-evolution of Computation and Cooperation | Kunal Jha, Francesco Cicala, Blaise Agüera y Arcas +4 | cs.MA | 2026-09-09 |
| #42 | The Truth Was Never Gone: Perfect Aliasing in Compliant-Context Truth Probes | Dylan Jayabahu | cs.LG | 2026-09-09 |
| #43 | Programmable World Model | Zheng-Hui Huang, Guixu Lin, Jiacheng Lin +8 | cs.CV | 2026-09-09 |
| #44 | Subgroup Membership Inference Audits of Differentially Private Synthetic Text | Yidan Sun, Viktor Schlegel, Srinivasan Nandakumar +2 | cs.CR | 2026-09-09 |
| #45 | Exact-Form Regret for Gradient Descent, Mirror Descent and Follow-the-Regularized-Leader | Ashkan Soleymani, Gabriele Farina, Patrick Jaillet | cs.LG | 2026-09-08 |
| #46 | Valerant: An Automatic Navigable Game Map Generator via Action-Conditioned World Model Exploration | Yiran Qiao, Feng Wang, Jing Ma | cs.AI | 2026-09-08 |
| #47 | MeClear: Cooperative Game-Theoretic Attribution and Risk-Aware Memory Clearance for Long-Horizon LLM Agents | Boyu Yang, Jiazheng Sun, Zilong Lu +3 | cs.AI | 2026-09-08 |
| #48 | The Surprising Effectiveness of Approximate Value Iteration in Self-Play | Raphael Boige, Amine Boumaza, Bruno Scherrer | cs.AI | 2026-09-08 |
| #49 | PlayTrain: An Efficient Reinforcement Learning Framework for LLM-Generated Adaptable JavaScript Games | Ryan Truong, Lance Ying, Samuel J. Gershman +1 | cs.LG | 2026-09-08 |
| #50 | Deposon: An Auditable, Conservation-Guaranteed, Game-Theoretically Tested Scattering Layer over LLM Reasoning Paths | Qihao Yuan | cs.AI | 2026-09-08 |