PaperScope
LIVE · 2026-09-04 05:40 UTC

Reducing Catastrophic Risk from AI with Systematic Monitoring and Evaluation of Rogue AI Progression

T. Bauer, W. P. Kegelmeyer, E. Begoli, A. Sadovnik, T. Emerson, C. Corley, N. Generous, J. Moore, B. Bartoldson, R. Goldhan, M. Goldman, M. Greaves, M. J. D. Vermeer, B. MacLennan, D. Schulker, N. VanHoudnos, J. Bansemer, Y. Bengio

Latestcs.CLcs.LGcs.AIcs.CV
arXiv ID
2609.03189 v1
Category
Submitted
2026-09-02

Abstract

This article presents a structured framework of behavioral indicators that may signal progression toward potentially catastrophic threats from artificial intelligence systems. We adopt a pragmatic approach, inspired by established methodologies in cybersecurity and national security. By establishing clear metrics, indicators, and thresholds across multiple dimensions of AI capability and behavior, this framework enables researchers and policymakers to implement evidence-based monitoring protocols.

arXiv abs page · PDF · same-day batch