To meet the ever-increasing demands of the cybersecurity workforce, AI tutors have been proposed for personalized, scalable education. But, while AI tutors have shown promise in introductory programming courses, no work has evaluated their use in hands-on exploration and exploitation exercises (e.g., "Capture the Flag"...
Michael Tompkins, Nihaarika Agarwal, Ananta Soneji et al.· 0 citations
Artificial Intelligence (AI) and large language models (LLMs) are increasingly used in social and psychological research. Among potential applications, LLMs can be used to generate, customise, or adapt measurement instruments. This study presents a preliminary investigation of AI-generated questionnaires by comparing t...
Mario Angelelli, Morena Oliva, Serena Arima et al.· 0 citations
Rel relational authority is developed: authorization is distributed across people, records, providers, and audiences, and must remain traceable as those relations change, and must remain traceable as those relations change.
Voice agents must complete users' tasks despite noise, reverberation, and competing speech. Evaluating agents' robustness therefore requires following how acoustic conditions affect the conversation and the actions taken on the user's behalf. This overview examines what existing benchmarks reveal about agents' ability...
Amir Ivry, Kai-Wei Chang, Lin Zhang et al.· 0 citations
Reach audiences
Advertise in front of researchers, engineers, and readers.
Priva-See is built, an LLM-based inference system for app-collected user data that reflects the best understanding of how real-life adtech companies would leverage machine learning to build user profiles, and suggests changes to how smartphone OSes should gather user consent for data access, to better inform users abou...
Sarah Radway, Zoe Robert, Matthew Soto et al.· 0 citations
This paper explores the potential of integrating e-textiles as part of the approach to delivering computing in UK secondary schools. As one of the few UK-based exploratory studies of teachers experiences, it investigates how e-textile platforms such as the SewSimple maker kit and the BBC micro:bit can be incorporated i...
Yifan Feng, Hanlin Zhang, Yishan Du et al.· 0 citations
This paper introduces experience-centered design (ECD) as an approach to investigate potential NDRAs within FAVs by examining the intricate connections between individuals'daily routines, specific travel contexts, and the activities they might undertake in transit.
Ke-Qi Chen, Xiao Xue, Xin-Yi Liu et al.· 0 citations
Teachers are increasingly using generative AI to support instruction, yet it remains unclear how pedagogical intentions are translated into chatbot configurations and reflected in chatbot behavior. We studied a teacher-facing chatbot authoring tool in professional development workshops with 27 middle school teachers, a...
Bahare Riahi, Deniz Ozturk, Alice Guth et al.· 0 citations
External human-machine interfaces (eHMIs) are evolving from predefined displays toward adaptive communication strategies that respond to changing traffic and road-user states. This transition requires experimental infrastructure that supports human-in-the-loop (HIL) interaction, software-in-the-loop (SIL) algorithm exe...
Yun Ye, Zexuan Li, Haoyang Liang et al.· 0 citations
National cybersecurity and digital-governance capacity frameworks, most prominently the Cybersecurity Capacity Maturity Model for Nations (CMM), shape hundreds of millions of dollars in donor-funded governance investments across the Global South. Yet a persistent, under-theorised gap remains between state-level institu...
Wael Albayaydh (University of Oxford), Ivan Flechais (University of Oxford)· 0 citations
This work evaluates state-of-the-art LLMs as pointwise and pairwise judges of conversational success on CANDOR, finding pointwise scoring correlates moderately with human ratings, while pairwise comparison suffers from long transcripts and positional bias.
Maike Zufle, Patrícia Schmidtová, Vilém Zouhar et al.· 0 citations
A novel definition of pedestrian-vehicle interaction as a partially observable Markov decision process (POMDP) with theory-grounded perceptual, cognitive, and motor constraints is introduced to establish a blueprint for simulator-ready pedestrian models that can support the development and evaluation of automated drivi...
Ruo-Feng Wang, Patrick Ebel, Philipp Wintersberger et al.· 0 citations
Students in MIT’s Concourse program delve deeply into the human condition, debate challenging questions, and learn to develop judgment about issues that can’t be quantified.
Jennifer Neville did not want to go into computer science—but that’s exactly where she landed. Neville discusses the starts and stops that led to her professional sweet spot and her work identifying “surprising failures” making it hard for AI to handle complexity. The post What AI gets wrong and what failure teaches us appeared first on Microsoft Research.
MIT News · Artificial Intelligence· news.mit.eduSep 30, 2026
Able to defeat top-ranked human players and more efficient than other models, the new system could help decision-makers in military maneuvers or business negotiations.
Computer scientist, entrepreneur, and philanthropist will collaborate with the MIT Schwarzman College of Computing to advance AI and scientific discovery.
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.