Aug 2026· 1 citation· ⚡ 1 influential· 48 references
Computer Science
TL;DR
A large-scale survey is conducted to elicit human judgments of responsibility in multi-agent sequential decision-making scenarios, using a modified version of the card game Goofspiel to assess multiple responsibility attribution methods and highlight key factors that influence human responsibility judgments.
Abstract
With the growing adoption of artificial intelligence in high-stakes decision-making, identifying the causes of outcomes--particularly failures--and determining who is responsible has become a critical concern. In this work, we examine how well formal definitions of \textit{responsibility attribution}, grounded in the framework of \textit{actual causality}, align with human judgments of responsibility. To this end, we conduct a large-scale survey to elicit human judgments of responsibility in multi-agent sequential decision-making scenarios, using a modified version of the card game Goofspiel. We evaluate multiple responsibility attribution methods, assess their alignment with human judgments about responsibility, and identify factors that significantly shape responsibility judgments. While no single responsibility attribution method consistently aligns with human responses, our findings highlight key factors that influence human responsibility judgments, including agent-specific biases and amount of information available to agents during decision-making.
This study designs, prototypes, and evaluates Servi.AI, a personality-aware multi-agent intelligent decision support system for enterprise strategy, and finds that selective rather than default deployment is supported.
Xu Zhou, Zhong-Yi Jiang· Applied System Innovation· 0 citations
This work introduces a notion of retrospective (backward) counterfactual responsibility, which quantifies an agent's accountability for outcomes resulting from a given strategy profile, and demonstrates how to compute stable strategy profiles in which agents trade off responsibility against expected reward.
Chunyan Mu, Muhammad Najib· Proceedings of the Thirty-Fi...· 0 citations
It is demonstrated that multi-agent consensus can enforce artificial agreement at the expense of true human alignment at the expense of true human alignment, revealing a structural limitation in consensus-style, role-specialized MAD protocols for subjective scoring.
Mi-Ra Song, Chanwoo Kim, Sugyeong Eo et al.· 0 citations
This paper argues that a more adequate account of responsibility can be developed by drawing on Gary Watson’s distinction between attributability and accountability, together with his later elaboration of responsibility in terms of attributability, accountability, answerability, and culpability, to provide a structured...
This work introduces XstrAI, an audience-aware multi-agent framework that treats local explanations as fixed evidence and structures how it is communicated to each target reader, and evaluates XstrAI on diabetes and stroke risk prediction against 11 baselines.
F. Musicco, Danilo Danese, Giuseppe Fasano et al.· 0 citations
As LLMs take on roles requiring moral advice, understanding how they attribute moral agency becomes critical. Humans possess moral agency, the capacity to make ethically guided decisions and bear responsibility for their consequences, a well-established construct in moral psychology. Yet as artificial agents (AAs) such...
F. Mansilla, Aloysius Y. F. Tok, B. Guellaï et al.· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.