Skip to results
MLSift

Titles, abstracts, or an arXiv ID

← Back to results
routineAgents & LLM SystemsMirror Agent Model2609.05190

The Mirror Agent Model: a Bayesian Architecture for Interpretable Agent Behavior

Michele Persiani, Thomas Hellström

cs.AI

Abstract

In this paper we illustrate a novel architecture generating interpretable behavior and explanations. We refer to this architecture as the Mirror Agent Model because it defines the observer model, that is the target of explicit and implicit communications, as a mirror of the agent's. With the goal of providing a general understanding of this work, we firstly show prior relevant results addressing the informative communication of agents intentions and the production of legible behavior. In the second part of the paper we furnish the architecture with novel capabilities for explanations through off-the-shelf saliency methods, followed by preliminary qualitative results.

Topics

Classified with taxonomy v2 on Mon, 7 Sept 2026.

Report a classification error

Loading the PDF downloads the document. Open it in your browser's viewer, or load it here.

Open PDF