Skip to content
← All posts

How MIT students are helping to prevent cyberattacks

MIT News · Artificial Intelligence · news.mit.edu · By Nicole Estvanik Taylor | Department of Urban Studies and Planning · July 13, 2026

Students from the MIT Cybersecurity Clinic help local governments and other vulnerable organizations defend against digital threats.

Read on MIT News · Artificial Intelligence → Opens the original article in a new tab.

More from the blog

MIT News · Artificial Intelligence Oct 7, 2026

Discovering the value of humanistic inquiry

Students in MIT’s Concourse program delve deeply into the human condition, debate challenging questions, and learn to develop judgment about issues that can’t be quantified.

Related papers

BadRAG: Identifying Vulnerabilities in Retrieval Augmented Generation of Large Language Models

A novel threat is unveiled in which attackers steer the RAG system's response by injecting malicious passages into its knowledge base, enabling the attacker to steer the response without altering the user input or modifying the RAG weights.

Jiaqi Xue, Meng Zheng, Yebowen Hu et al. · 109 citations · ⚡8

OverThink: Slowdown Attacks on Reasoning LLMs

This work evaluates Overthink on proprietary and open-source reasoning models across the FreshQA, SQuAD, and MuSR datasets, and shows that newer generations of RLMs, while showing a drastic increase in per-token cost, also exhibit up to a 2.3x increase in reasoning tokens, leaving them more vulnerable to Overthink atta...

Abhinav Kumar, Jaechul Roh, Ali Naseh et al. · 92 citations · ⚡9

Learning diverse attacks on large language models for robust red-teaming and safety tuning

This work proposes to use GFlowNet fine-tuning followed by a secondary smoothing phase, to train the attacker model to generate diverse and effective attack prompts, and finds that the attacks generated by the method are effective against a wide range of target LLMs, both with and without safety tuning, and transfer we...

Seanie Lee, Minsu Kim, Lynn Cherif et al. · 62 citations · ⚡8

WAInjectBench: Benchmarking Prompt Injection Detections for Web Agents

The key findings show that while some detectors can identify attacks that rely on explicit textual instructions or visible image perturbations with moderate to high accuracy, they largely fail against attacks that omit explicit instructions or employ imperceptible perturbations.

Yinuo Liu, Ruohan Xu, Xilong Wang et al. · 22 citations · ⚡1

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.