ADAE is introduced, a proposed AI engineering subdiscipline concerned with establishing measurable, continuous, and actionable accountability for deployed AI systems and treats accountability as a deployment-layer property rather than solely as a property of an individual model.
Abstract
Artificial intelligence systems are rapidly becoming critical components in healthcare, finance, public services, and other safety-critical domains. Yet the engineering practices used to evaluate these systems remain predominantly model-centric, emphasizing properties such as accuracy, robustness, fairness, and interpretability before deployment. These properties are necessary but insufficient once an AI system operates within an ever changing socio-technical environment characterized by distribution shifts, institutional constraints, human feedback loops, privacy requirements, and interactions among multiple AI agents. This vision paper introduces AI Deployment Accountability Engineering (ADAE), a proposed AI engineering subdiscipline concerned with establishing measurable, continuous, and actionable accountability for deployed AI systems. ADAE treats accountability as a deployment-layer property rather than solely as a property of an individual model. It seeks to determine whether an AI-enabled system continues to operate within acceptable risk limits, identify the contexts in which failures emerge, attribute failures across interacting technical and human components, translate technical failures into downstream consequences, and support timely intervention. We articulate a research agenda built around four interconnected pillars: structured discovery of context-dependent failure modes, privacy-preserving accountability measurement, system-level risk analysis for agentic AI, and translation of technical failures into operational, and institutional risks. The broader goal is to establish foundational principles, mathematical tools, and system architectures for accountable AI deployment across safety-critical applications.
The integration of Artificial Intelligence (AI)/Machine Learning (ML) into high-impact domains such as finance and autonomous systems offers significant benefits, but also introduces complex risks and regulatory challenges. These systems exhibit properties including non-determinism, data dependence, and evolving vuln...
Anita Khadka, C. Maple· Artificial Intelligence Revi...· 0 citations
The proliferation of agentic artificial intelligence (AI) systems autonomous, goal-directed agents capable of acting with minimal human oversight, has exposed fundamental inadequacies in existing civil liability frameworks. Traditional tort, product liability, and custodian regimes struggle to accommodate systems whose...
Emma Lunardi· 2026 IEEE 34th International...· 0 citations
Artificial intelligence (AI) system failures are often silent. Applications may remain operational and continue to generate polished outputs even as recommendations degrade, manipulated inputs alter behavior, or automated agents exceed their intended authority. This lack of transparency creates a governance vacuum, obs...
Jess Montgomery, J. Copeland· The Pinnacle: A Journal by S...· 0 citations
AI reliability concerns whether an AI system performs its intended function dependably over a stated period and under stated operating conditions, with stated evidence. As these systems become more autonomous, that function includes more than a correct output. Retrieval, memory, tool use, permissions, human oversight,...
We argue that a recurring failure in the evaluation of deployed AI systems occurs when data collected for operational monitoring or regulatory compliance are interpreted as if they were designed for comparative evaluation. Automated driving provides a concrete example of this problem. U.S. disengagement and crash-repor...
Hung-Yu Lin, Xing-Ran Huang, Qi-Ming Guo et al.· 0 citations
With $2.1 million funding from Google.org, the open-source Public Transit Intelligence Hub will unify public transit monitoring, operations, and passenger communication.