Skip to content
Conference Open access

Counterfactual Reasoning for Responsibility Attribution in Probabilistic Multi-Agent Systems

Sep 2026 · Proceedings of the Thirty-Fifth International Joint Conference on Artificial Intelligence · 0 citations · 48 references

TL;DR

This work introduces a notion of retrospective (backward) counterfactual responsibility, which quantifies an agent's accountability for outcomes resulting from a given strategy profile, and demonstrates how to compute stable strategy profiles in which agents trade off responsibility against expected reward.

Abstract

Responsibility allocation---determining the extent to which agents are accountable for outcomes---is a fundamental challenge in the design and analysis of multi-agent systems. In this work, we model such systems as concurrent stochastic multi-player games and introduce a notion of retrospective (backward) counterfactual responsibility, which quantifies an agent's accountability for outcomes resulting from a given strategy profile. To allocate responsibility among agents, we utilise the Shapley value and formally show that this method satisfies key desirable properties, including fairness and consistency. Building on this foundation, we propose a formal framework that supports both verification and strategic reasoning in responsibility-aware multi-agent systems. Furthermore, by adopting Nash equilibrium as the solution concept, we demonstrate how to compute stable strategy profiles in which agents trade off responsibility against expected reward.

Read PDF

Similar papers

#artificial intelligence Preprint Sep 2026

Social Laws for Multi-agent Coordination in Stochastic Environments

In multi-agent environments, coordinating agents to prevent interference and ensure robust individual performance is a critical challenge. Previous research on social laws for multi-agent systems has primarily focused on deterministic, goal-based settings. This paper extends the concept of social laws to stochastic, re...

Rolando Fernandez, Caleb Probine, Tyler Lee et al. · 0 citations
#artificial intelligence Preprint Sep 2026

Multi-agent Scaling Across Disjunctive and Compensatory Tasks

Multi-agent LLM systems are often expected to improve as team size increases, yet the scaling behavior may depend on task structure. Our central contribution is to introduce Steiner's taxonomy of group tasks as a framework for analyzing multi-agent LLM scaling and focusing the analysis on disjunctive and compensatory t...

Carolina Fortuna, Blaž Bertalanič · 0 citations

“Who Is to Blame?” and “What Is to Be Done?”

AI systems are increasingly deployed in safety-critical domains, where they behave non-deterministically, act under uncertainty, and form part of larger multi-agent systems without central control. When such systems cause harm, two questions arise: who is to blame, and what could have been done to prevent it? This thes...

M. Gladyshev · 4 citations
Preprint Aug 2026

Multi-Agent AI Safety as an Institutional Design Problem

This is the first paper from POLIS, an ongoing research programme studying algorithmic institutions for multi-agent systems, and asks which parts of an AI institution produce safety and how they do it.

X. Abdullah · 1 citation

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.