Skip to content

Category

human-computer interaction

1,682 papers

#artificial intelligence Preprint Sep 2026

E-CONAN (Entailment, CONtradition And Neutral) Diagnostics Dataset Investigating Linguistic Phenomena in Arabic Natural Language Understanding

Natural Language Understanding (NLU) plays a crucial role in various applications, yet its performance suffers from weaknesses in handling the complexities of human languages, ranging from lexical ambiguity to high-level reasoning difficulties. Analyzing errors across diverse linguistic phenomena is crucial for NLU imp...

Khloud Al Jallad, Nada Ghneim, Ghaida Rebdawi · 0 citations
#artificial intelligence Preprint Sep 2026

DataMagic: Authoring Data Videos through Declarative Multi-Agent Orchestration

Data videos communicate data insights through dynamic charts, voice narration, and synchronized animations, and have become a widely adopted form of data storytelling. However, producing them requires expertise in data analysis, narrative design, and video editing. Static visualization tools lack narrative and animatio...

Yu-Peng Xie, Zhen-Yang Wang, Liang-Wei Wang et al. · 2 citations
#artificial intelligence Preprint Sep 2026

Beyond Tasks: A Vision for Reproducing an Animal-like Behavioral Substrate Using Modern Robot Learning Techniques

This work argues that their continual coordination under competing demands constitutes an important and underexplored target for modern robot learning and proposes the ethological behavioral substrate as a conceptual lens for studying this form of competence in artificial agents.

Samiyuru Menik, H. Jayalath · 0 citations
#artificial intelligence Preprint Sep 2026

ParallelPilot: Supporting Coordination and Monitoring in Parallel AI Coding

As coding assistants become increasingly autonomous, developers run multiple sessions in parallel, shifting the challenge from code generation alone to coordinating and monitoring concurrent agent work. Through a formative study (N=14), we identified PILOT: five supervisory practices for Planning, Isolating, Logging, O...

Tao Long, Wei Shi, Hussein Mozannar et al. · 0 citations
#artificial intelligence Preprint Sep 2026

DashAct: A Progressive Diagnostic Benchmark for GUI Agents in Interactive Dashboard Analysis

Interactive dashboards require users to reveal and connect evidence across stateful interactions. Although graphical user interface (GUI) agents could automate this process, existing dashboard benchmarks primarily report final answers or task success. They provide limited insight into whether failures arise from mainta...

Chu-Han Zhang, Qinghongbing Xie, Zi-Yue Wang et al. · 0 citations
#artificial intelligence Open access Sep 2026

Vibe Analysis: Exploring LLM Adoption by Data Visualization Practitioners.

These findings show that Vis designers actively use LLMs for both creative and technical aspects of the visualization process, and opens up opportunities for research combining LLM-mediated work with Vis tools that incorporate data visualization guidance, constraints, and best practices.

S. C. Spivak, Aditi Krishna, Mahsan Nourani et al. · 0 citations
#artificial intelligence Review Sep 2026

CLAIRE: A Schema-Grounded Hybrid Workflow for Healthcare Administrative Form Completion

Healthcare administrative staff transfer structured information from electronic health records, referrals, claims systems, provider rosters, and work queues into dynamic forms. We developed and evaluated CLAIRE (Clinical Language and Agentic Intelligence for Reasoning and Entry), a hybrid workflow that separates field-...

Garapati Keerthana, Manik Gupta · 0 citations
#artificial intelligence Preprint Open access Sep 2026

Artificial intelligences and human scientists exhibit complementary strengths in theory building

We investigate the effectiveness of artificial intelligences (AI)-specifically large language models (LLMs)-relative to human scientists at high-level cognitive tasks in social science such as theory formulation, predictions of novel empirical results, and theory revision in response to new evidence. The research domai...

Ke Li, Spyros I. Zoumpoulis, Phanish Puranam et al. · 0 citations
#artificial intelligence Review Sep 2026

A Benchmark for LLM's Understanding of Middle School and High School Science Topics

Large language models (LLMs) are increasingly integrated into educational settings, yet educators lack robust, standards-aligned tools to evaluate their effectiveness in K-12 science contexts. Existing benchmarks predominantly assess general language or advanced scientific reasoning, leaving a critical gap in understan...

Noah L. Schroeder, Yessy Eka Ambarwati, Yu-Ji Zhang et al. · 0 citations
#robotics Preprint Open access Sep 2026

Measurement and Potential Field-Based Patient Modeling for Model-Mediated Tele-ultrasound

Teleoperated ultrasound can improve diagnostic medical imaging access for remote communities. Having accurate force feedback is important for enabling sonographers to apply the appropriate probe contact force to optimize ultrasound image quality. However, large time delays in communication make direct force feedback im...

Ryan S. Yeung, David G. Black, Septimiu E. Salcudean · 0 citations

From tech blogs

See all →
MIT News · Artificial Intelligence Oct 7, 2026

Discovering the value of humanistic inquiry

Students in MIT’s Concourse program delve deeply into the human condition, debate challenging questions, and learn to develop judgment about issues that can’t be quantified.

Microsoft Research Blog Oct 6, 2026

What AI gets wrong and what failure teaches us

Jennifer Neville did not want to go into computer science—but that’s exactly where she landed. Neville discusses the starts and stops that led to her professional sweet spot and her work identifying “surprising failures” making it hard for AI to handle complexity.  The post What AI gets wrong and what failure teaches us appeared first on Microsoft Research.

MIT News · Artificial Intelligence Sep 30, 2026

This game-playing AI is the new champ at Stratego

Able to defeat top-ranked human players and more efficient than other models, the new system could help decision-makers in military maneuvers or business negotiations.

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.