Skip to content

Category

software testing

2,517 papers

#software testing Open access Oct 2026

Fuzzing the Boundary between Models and Code in Hybrid AI-Enabled Systems

Deep learning (DL) techniques are increasingly integrated into traditional software systems, giving rise to hybrid AI-enabled systems that combine neural models with program logic. While these systems exhibit remarkable capabilities, their complex and heterogeneous architectures pose significant challenges for reliabil...

Xin-Yu Gao, Yang Feng, Yu-Chen Lu et al. · 0 citations
#software testing Book Open access Oct 2026

(This Is Not a Paper about) Mutation Driven Development

Test driven development (TDD) is a controversial and interesting approach to software development; while many think of better tests as a primary purpose of TDD, in practice the goal is as much to use tests to encourage continued progress in coding. That goal however rests on the notion that TDD ensures tests are good e...

Alex Groce · 0 citations
#software testing Book Open access Oct 2026

Curated Semantic Mutants: Multi-purpose Artifacts for Grading and Hinting Student Test Suites

A human-LLM workflow that pairs each curated semantic mutant with an instructor-approved seed phrase for an on-demand LLM expansion, which shows that the curated semantic-mutant set contains a fraction of the mutants a traditional mutation engine produces.

Rebecca Williams Earle, Jonathan Bell · 0 citations
#software testing Book Open access Oct 2026

How Far Does Replication Pay Off? A Throughput-Area Study of Brute-Force SAT on an ECP5 FPGA

SAT solving is a core computational challenge across hardware verification, electronic design, and planning; an FPGA can evaluate SAT candidates at hardware clock rates without software overhead, enabling sub-microsecond solving on small instances. Yet, the throughput-area tradeoffs of parallel FPGA SAT datapaths remai...

Andrew Bonilla · 0 citations
#software testing Book Open access Oct 2026

The Mechanisms and Practice of Mutation Testing in Improving Software Quality

This proposed dissertation investigates mutation testing from three complementary perspectives: how mutations interact with program executions and test oracles to produce mutant kills and reveal opportunities for test improvement, and how mutation testing is adopted, configured, sustained, and acted upon in popular ope...

Hang Du · 0 citations
#software testing Book Open access Oct 2026

Beyond Next-Token Prediction: Autonomous Verification and Visual Testing in the Agent-First Era (Keynote)

This talk explores how agent-first development platforms like Google Antigravity shift the paradigm from manual test authoring to autonomous software verification, and describes recent advances in autonomous visual testing.

Paige Bailey · 0 citations
#software testing Book Open access Oct 2026

PyMut4SE: Comprehensive Mutation Testing for Python

PyMut4SE is a novel mutation tool for Python that focuses on a comprehensive set of mutations for any Python project, providing a rich and extensible set of mutation operators, access to mutated source code and its characteristics, detailed execution and behavioral observations, and support for both selective mutation...

Laura Plein, Matthieu Jimenez, Mike Papadakis · 0 citations
#software testing Book Open access Oct 2026

Towards Scalable and Verifiable Automated Translation from C to Safe Rust

Memory-safety bugs are one of the oldest and most common sources of security vulnerabilities, and their modern-day prevalence is a consequence of the widespread use of non-memory-safe languages such as C and C++. Transitioning away from C and C++ to memory-safe languages, namely Rust, to eliminate memory-safety bugs is...

Victor Chen · 0 citations
#software testing Book Open access Oct 2026

Typed Template Fuzzing

Fuzzing finds bugs by testing software behavior with high volumes of input. To create semantically valid inputs for different targets, traditional fuzzers require significant engineering effort. LLM-based fuzzing approaches can be adapted to create inputs for a wide range of software but introduce new issues: every tes...

Lu Maltsis · 0 citations
#software testing Book Open access Oct 2026

SPINACH: Inferring Properties of Web Applications for Property-Based Testing

The paper asks whether high-level conceptual specifications improve LLM-generated PBT quality and whether they help developers extend and maintain AI-generated systems (RQ2), and reports preliminary results applying Spinach to two open-source applications.

Savitha Ravi, Michael Coblenz · 0 citations
#software testing Book Open access Oct 2026

Decomposing LLM-Based Testing with Agent Skills: A Case Study on Numerical Inconsistencies

LLM-based testing can expose subtle reliability issues in numerical software, but monolithic prompts make it hard to inspect which guidance drives effectiveness. We study whether Agent Skills can make such workflows more explainable by factoring procedural knowledge into composable testing components. Using numerical i...

Yu-Tong Wang, Cindy Rubio-González · 0 citations
#software testing Book Open access Oct 2026

CoVerif: An Automated Contract Verifier for Java using Symbolic Execution

Specifying contracts as program behavior using Hoare triples is an established practice in software engineering. Compared to traditional testing approaches, formal verification of these contracts provides stronger guarantees of functional correctness. However, its practical adoption has remained limited since significa...

Aryan Kumar, Alex Toppo, Sandip Ghosal · 0 citations

From tech blogs

See all →
MIT News · Artificial Intelligence Oct 2, 2026

Documenting the tech worker movement

Writing as a participant and researcher, PhD student JS Tan SM ’22 has co-authored a new book about the rise of tech worker protests and the employer backlash that followed.

GPT-Lab Sep 23, 2026

Requirements Don’t Live in Isolation: What We’re Exploring with Req-Space

Requirements in large systems rarely exist in isolation. Their meaning depends on the wider project context - other requirements, policies, decisions, tests, and implementation details. That becomes especially important when AI is used for review, because spotting a possible conflict or gap is only the beginning. ReqSpace explores how AI, visualisation, and connected project context can help reviewers understand those findings, trace the relationships behind them, and focus on the questions that…

GPT-Lab Sep 17, 2026

Beyond Prompt Engineering: The Role of Tacit Knowledge in Software Engineering

AI is making software generation faster, but speed does not remove the need for expertise. As more work is delegated to AI, tacit knowledge may become one of the most important human advantages in software engineering. The post Beyond Prompt Engineering: The Role of Tacit Knowledge in Software Engineering appeared first on GPT-Lab.

MIT News · Artificial Intelligence Aug 17, 2026

Q&A: Rethinking how innovation happens

In his latest book, Professor Eugene Fitzgerald examines the forces that turn breakthroughs into value — and why innovation resists simple formulas.

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.