Skip to content

Category

human-computer interaction

1,660 papers

#robotics Preprint Open access Oct 2026

TriDeliver: Cooperative Air-Ground Instant Delivery with UAVs, Couriers, and Crowdsourced Ground Vehicles

Instant delivery, shipping items before critical deadlines, is essential in daily life. While multiple delivery agents, such as couriers, Unmanned Aerial Vehicles (UAVs), and crowdsourced agents, have been widely employed, each of them faces inherent limitations (e.g., low efficiency/labor shortages, flight control, an...

Junhui Gao, Yan Pan, Qianru Wang et al. · 0 citations
#computer vision Preprint Open access Oct 2026

CADReasoner: Iterative Program Editing for CAD Reverse Engineering

Computer-Aided Design (CAD) powers modern engineering, yet producing high-quality parts still demands substantial expert effort. Many AI systems tackle CAD reverse engineering, but most are single-pass and miss fine geometric details. In contrast, human engineers compare the input shape with the reconstruction and iter...

Soslan Kabisov, Vsevolod Kirichuk, Andrey Volkov et al. · 0 citations
#human-computer interacti... Preprint Open access Oct 2026

Multimodal Data Comprehension: Understanding How Visual-Textual Chains of Information Influence Data Interpretation

Visualizations and text often work together to support effective data communication. Despite this common paradigm, we know little about how the interplay of these modalities affects people's data comprehension. We present a novel experimental paradigm to investigate multimodal data comprehension---the process of people...

Arran Zeyu Wang, Fuling Sun, Danielle Albers Szafir · 0 citations
#human-computer interacti... Preprint Open access Oct 2026

What to Distinguish and How? Opportunities and Challenges of Augmenting Multiple, Cluttered Objects in Complex Scenes for People with Low Vision

People with low vision (PLV) struggle to perceive complex scenes like busy kitchens and crowded streets, which contain many objects, visual clutter, and dynamic elements. Prior AR systems for low vision either enhance low-level visual features or augment task-relevant objects for single tasks in simple settings, leavin...

Yuheng Wu, Ruijia Chen, Jaewook Lee et al. · 0 citations
#human-computer interacti... Preprint Open access Oct 2026

Tutor, Not Solver: Designing a Guardrailed AI Assistant for Learning in Higher Education: A Design Case of PeteChat

Generative artificial intelligence (AI) tutors hold significant promise for higher education, yet designing systems that scaffold learning without undermining academic integrity remains an open design challenge. This paper presents PeteChat, a course-aligned AI tutor developed and piloted at a large U.S. research unive...

Belle Li, Lily Tan, Wei Zakharov et al. · 0 citations
#computer vision Preprint Open access Oct 2026

EduGage: A Multimodal Dataset and Benchmark for Sensor-Based Momentary Assessment of Engagement in Self-Guided Video Learning

Engagement, which links to attentional, emotional, and cognitive dimensions, plays an important role in learning. In online and video-based learning environments, learners often need to regulate their own interactions with instructional materials. Measuring and reflecting on engagement can therefore support both learne...

Zikang Leng, Edan Eyal, Yingtian Shi et al. · 0 citations
#human-computer interacti... Preprint Open access Oct 2026

Harnessing the Power of AI in Qualitative Research: Role Assignment, Engagement, and User Perceptions of AI-Generated Follow-Up Questions in Semi-Structured Interviews

Semi-structured interviews highly rely on the quality of follow-up questions, yet interviewers' knowledge and skills may limit their depth and potentially affect outcomes. While many studies have shown the usefulness of large language models (LLMs) for qualitative analysis, their possibility in the data collection proc...

He Zhang, Yueyan Liu, Xin Guan et al. · 0 citations
#human-computer interacti... Preprint Open access Oct 2026

MATE: Diagnosing Empathy Calibration Failures in Multi-Turn Human-LLM Interaction

Large language models are increasingly used in emotionally consequential interactions. Response-level evaluation, however, struggles to diagnose how empathy fails across turns. We introduce MATE (Multi-turn Assessment of calibraTed Empathy), a framework that evaluates empathy as a multi-turn, perception-centered proces...

Yeseon Hong, Junhyuk Choi, Minju Kim et al. · 0 citations
#computer vision Preprint Open access Oct 2026

MemoCare: An Interactive Multimodal Mobile System for Automated Cognitive Screening

MemoCare is an interactive mobile system for automated multimodal cognitive screening. A React Native application combines spoken responses, temporal and spatial orientation, touchscreen actions, and visuoconstruction in complete English and Vietnamese workflows. Speech is transcribed by Google Speech-to-Text and score...

Duy-Cat Can, Mau Minh Phuc Le, Tuan-Khoa Hoang et al. · 0 citations
#human-computer interacti... Preprint Open access Oct 2026

CrossWeave: Bridging Perspectives Across Online Communities with a Dual-Pane Design

Social media systems typically display conversations among already familiar contributors, which can be predictable and one-sided. In civic discourse, this design narrows discussion, reinforces divides, and distorts the perception of public opinion. To encourage cross-community engagement, we present CrossWeave, an AI-p...

Fei Fang, Reva Hirave, William Jurayj et al. · 0 citations
#human-computer interacti... Preprint Open access Oct 2026

Reading Position Is the Baseline to Beat: A Time-Ordered Evaluation of Personalised Highlight Prediction

A reader's first highlights on a page are the cheapest personal signal a reading product has. The natural plan is to suggest what similar earlier readers marked, and to judge the result against popularity. We argue that the baseline to beat is reading position. In a time-ordered evaluation on one social highlighting pl...

Kazuki Nakayashiki, Keisuke Watanabe · 0 citations
#human-computer interacti... Preprint Open access Oct 2026

Intonation Perception in Real and Synthetic Speech across Varying Familiarity Levels: A Pilot Study of Equivalence Assessment

Language training relies on a corpus constructed by a large number linguistic materials. AI-powered voice clones provide a way to construct the corpus with relatively low cost. Singing voice conversion (SVC) model is used to generate synthetic voices. This study compares participants' performances on natural and synthe...

Hanrui Zhou, Gaoyuan Zhang, Yixiang Chen et al. · 0 citations

From tech blogs

See all →
MIT News · Artificial Intelligence Oct 7, 2026

Discovering the value of humanistic inquiry

Students in MIT’s Concourse program delve deeply into the human condition, debate challenging questions, and learn to develop judgment about issues that can’t be quantified.

Microsoft Research Blog Oct 6, 2026

What AI gets wrong and what failure teaches us

Jennifer Neville did not want to go into computer science—but that’s exactly where she landed. Neville discusses the starts and stops that led to her professional sweet spot and her work identifying “surprising failures” making it hard for AI to handle complexity.  The post What AI gets wrong and what failure teaches us appeared first on Microsoft Research.

MIT News · Artificial Intelligence Sep 30, 2026

This game-playing AI is the new champ at Stratego

Able to defeat top-ranked human players and more efficient than other models, the new system could help decision-makers in military maneuvers or business negotiations.

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.