PaperScope
LIVE · 2026-10-08 05:40 UTC

AI Safety Considerations for Agents With Limited Time to Act

Leo Zeitler, Jack Richings, Victoria Nockles

Latestcs.CLcs.LGcs.AIcs.CV
arXiv ID
2610.10285 v1
Category
Submitted
2026-10-07

Abstract

In the wake of the increasingly public discussion about AI alignment, recent work has tried to propose specific AI architectures that behave safely. However, the proposed arguments that seemingly demonstrate proved alignment mostly neglect the environment the agent needs to act in. We discuss theoretical bounds for agent-agnostic safety guarantees in environments that can only be partially observed and within which an action is required within limited time. We introduce two realistic scenarios, one with an infinite state space and one with signal mixture. In these scenarios, we prove that even a perfect agent cannot guarantee safe behaviour. It will be argued that for any proof of AI safety or alignment, the environment and associated safe actions need to be specifically considered together with the agent.

arXiv abs page · PDF · same-day batch