Counterfactual Reasoning for Responsibility Attribution in Probabilistic Multi-Agent Systems
Counterfactual Reasoning for Responsibility Attribution in Probabilistic Multi-Agent Systems
Chunyan Mu, Muhammad Najib
Proceedings of the Thirty-Fifth International Joint Conference on Artificial Intelligence
Main Track. Pages 288-296.
https://doi.org/10.24963/ijcai.2026/33
Responsibility allocation---determining the extent to which agents are accountable for outcomes---is a fundamental challenge in the design and analysis of multi-agent systems. In this work, we model such systems as concurrent stochastic multi-player games and introduce a notion of retrospective (backward) counterfactual responsibility, which quantifies an agent's accountability for outcomes resulting from a given strategy profile. To allocate responsibility among agents, we utilise the Shapley value and formally show that this method satisfies key desirable properties, including fairness and consistency. Building on this foundation, we propose a formal framework that supports both verification and strategic reasoning in responsibility-aware multi-agent systems. Furthermore, by adopting Nash equilibrium as the solution concept, we demonstrate how to compute stable strategy profiles in which agents trade off responsibility against expected reward.
Keywords:
Agent-based and Multi-agent Systems: Formal verification, validation and synthesis
AI: Agent-based and Multi-agent Systems
