PaperScope
LIVE · 2026-09-07 05:40 UTC

The Mirror Agent Model: a Bayesian Architecture for Interpretable Agent Behavior

Michele Persiani, Thomas Hellström

Latestcs.CLcs.LGcs.AIcs.CV
arXiv ID
2609.05190 v1
Category
Submitted
2026-09-04

Abstract

In this paper we illustrate a novel architecture generating interpretable behavior and explanations. We refer to this architecture as the Mirror Agent Model because it defines the observer model, that is the target of explicit and implicit communications, as a mirror of the agent's. With the goal of providing a general understanding of this work, we firstly show prior relevant results addressing the informative communication of agents intentions and the production of legible behavior. In the second part of the paper we furnish the architecture with novel capabilities for explanations through off-the-shelf saliency methods, followed by preliminary qualitative results.

Comment: Accepted at the International Workshop on Explainable, Transparent Autonomous Agents and Multi-Agent Systems, 2022

arXiv abs page · PDF · same-day batch