Skip to content

Category

human-computer interaction

1,682 papers

#robotics Preprint Sep 2026

Toward Human-in-the-Loop Robot Failure Recovery: Bridging Communication Gaps in Human-Robot Collaboration

Robots can recover from failures by asking bystanders for help, but effective human-in-the-loop recovery requires communication that accounts for differences in people's knowledge. Prior inverse-semantics work generates requests using a single listener model, leaving differences in listener knowledge untested. We intro...

Promise Ekpo, T. Vijay, Dhruv Mandalik et al. · 0 citations
#artificial intelligence Preprint Sep 2026

Do Not Trust the Benchmark: Limitations of General LLM Rankings and a Case for Task-Specific Evaluation

Benchmark scores increasingly influence the development, marketing, and selection of large language models (LLMs). Yet an overall score is interpretable only in relation to the system tested, the questions included, and the conditions of evaluation. This perspective examines five connected limitations of general LLM ra...

Danial Amin · 0 citations
#artificial intelligence Review Sep 2026

The Moral Check: Strategic AI Governance for the Pacing Problem

Technology cannot steer itself. Strategy provides that steering, establishing the rule that purpose and judgment must precede compute capital. As the frontier artificial intelligence (AI) ecosystem accelerates exponentially, the pacing problem induces severe cognitive tunneling in engineering teams, prioritizing scalar...

Zaid Amin, Rahma Santhi Zinaida, Nazlena Mohamad Ali · 0 citations
#artificial intelligence Preprint Open access Sep 2026

Toward Auditable and Calibrated AI for Dementia-Related Crash Severity Prediction: A Selective Deferral Framework to Support Human Review

Public crash databases increasingly support automated safety analysis, but crash severity prediction remains difficult to translate into public-sector decision workflows when models are evaluated primarily as ordinary classifiers. This study reframes dementia-related crash severity modeling as a decision-aware triage p...

Gaurab Chhetri, Anika Baitullah, Subasish Das · 0 citations
#human-computer interacti... Preprint Open access Sep 2026

Passthrough Rigidity: The Behavioral and Visuomotor Costs of Mediated Perception

Broad public adoption of head-mounted displays using video passthrough remains elusive despite significant market investment. A precise understanding of why users experience persistent discomfort even as hardware factors such as resolution and latency have dramatically improved remains an open issue. This paper investi...

Markus D. Solbach, Mohit Goyal, Sakar Khattar et al. · 0 citations
#human-computer interacti... Preprint Open access Sep 2026

Who Does What in AI Auditing? Designing Human-AI Collaboration for Auditing Generative AI

AI auditing increasingly incorporates AI agents to expand the scale and breadth of audit coverage, yet little is known about how auditing work should be divided without displacing human judgment. We introduce Human-Agent Audit Collaboration (HAAC), a workflow and system for structuring human-AI collaboration in AI audi...

Eunkyu Park, Markelle Roesti, Wesley Hanwen Deng et al. · 0 citations
#artificial intelligence Review Sep 2026

Generative Tutorial: Towards Live Contextualized Visual Instructions for Physical Tasks

Visual instructions for physical tasks are typically authored in one context and followed in another, requiring users to translate demonstrated tools, materials, and spatial relationships into their own environment. We introduce Generative Tutorial, a conceptual framework for live visual instruction that depicts intend...

Mu-Zhe Wu, Zuojun Li, Xu Wang et al. · 0 citations
#human-computer interacti... Preprint Open access Sep 2026

Whose Facts Count? A Culturally Responsive Audit of LLM Evaluation Benchmarks

LLM benchmarks function as evaluation instruments, informing decisions that affect education, labor, and public services worldwide. Drawing on Hood, Kirkhart, and Hopson's culturally responsive evaluation (CRE) frameworks, this paper applies a six-dimension CR rubric to audit OpenAI's SimpleQA (N = 4,326 items) and the...

Fatima Tuz Zahra, Md. Sajeebul Islam Sk., Rachel Chung · 0 citations
#human-computer interacti... Preprint Open access Sep 2026

EMooly: Supporting Autistic Children in Collaborative Social-Emotional Learning with Caregiver Participation through Interactive AI-infused and AR Activities

Children with autism spectrum disorder (ASD) have social-emotional deficits that lead to difficulties in recognizing emotions as well as understanding and responding to social interactions. This study presents EMooly, a tablet game that actively involves caregivers and leverages augmented reality (AR) and generative AI...

Yue Lyu, Di Liu, Pengcheng An et al. · 0 citations
#human-computer interacti... Preprint Open access Sep 2026

ATCion: Exploring the Design of Icon-based Visual Aids for Enhancing In-cockpit Air Traffic Control Communication

Effective communication between pilots and air traffic control (ATC) is essential for aviation safety, but verbal exchanges over radios are prone to miscommunication, especially under high workload conditions. While cockpit-embedded visual aids offer the potential to enhance ATC communication, little is known about how...

Yue Lyu, Xizi Wang, Hanlu Ma et al. · 0 citations
#artificial intelligence Preprint Open access Sep 2026

Small-world Networks of Agents Brainstorm AI Risks to Support Ideation

The ideation phase of participatory AI risk assessment often starts with a blank slate or a limited list of predefined risks, making it difficult to surface indirect or systemic harms. To address this limitation, we propose a three-stage ideation support tool. The tool complements participatory AI, rather than replacin...

Ke Zhou, Edyta Bogucka, Daniele Quercia · 0 citations
#human-computer interacti... Preprint Sep 2026

CRiDiT: Instantiating a run-time testbed for trust calibration in AI-infused systems

The integration of AI into larger technical infrastructures has made the alignment of human trust with system trustworthiness, known as trust calibration, a critical engineering concern, since misplaced trust in either direction leads to operational and safety risks. While conceptual frameworks provide a strong foundat...

Yun-Tian Ding, N. Herbaut, C. Salinesi · 0 citations

From tech blogs

See all →
MIT News · Artificial Intelligence Oct 7, 2026

Discovering the value of humanistic inquiry

Students in MIT’s Concourse program delve deeply into the human condition, debate challenging questions, and learn to develop judgment about issues that can’t be quantified.

Microsoft Research Blog Oct 6, 2026

What AI gets wrong and what failure teaches us

Jennifer Neville did not want to go into computer science—but that’s exactly where she landed. Neville discusses the starts and stops that led to her professional sweet spot and her work identifying “surprising failures” making it hard for AI to handle complexity.  The post What AI gets wrong and what failure teaches us appeared first on Microsoft Research.

MIT News · Artificial Intelligence Sep 30, 2026

This game-playing AI is the new champ at Stratego

Able to defeat top-ranked human players and more efficient than other models, the new system could help decision-makers in military maneuvers or business negotiations.

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.