Skip to content

Category

software testing

2,461 papers

Enhancing test data generation via path-grouped reusable prioritized DQN and PSO for mutation testing

This work pioneers integrating Deep Q-Network (DQN) into test data generation optimization, and models iterative test generation as a sequential decision-making process, capturing complex input-testing objective relationships and learning effective action strategies to guide PSO-based evolutionary search.

Jia-Le Wang, Xiang-Ying Dang, Dun-Wei Gong et al. · 0 citations
#software testing Editorial Open access Oct 2026

Editorial: Responsible and robust evaluation for real-world recommendation and search systems

Both search and recommender systems are software tools that are used on a daily basis and influence a large number of decisions, affecting users, different stakeholders (companies, governments, organizations, etc.), and society as a whole (Ricci et al., 2022; Alonso and Baeza-Yates, 2024). Hence, as the evaluation of t...

Alejandro Bellogín, Pablo Sánchez, Laura Sebastiá · 0 citations
#diffusion models Open access Oct 2026

Supplementary Software S1: Three-dimensional model of coupled heat and moisture transfer, grain respiration and insect population dynamics in aerated wheat storage

Python/NumPy/SciPy implementation of the three-dimensional cell-centred finite-volume model described in the article "Three-Dimensional Modelling of Coupled Heat and Moisture Transfer, Grain Respiration and Insect Population Dynamics in Aerated Wheat Storage: Continuous versus Controlled Aeration" (Kurbonov, Shadmanov,...

Abdinabi Mukhamadiyev, Nozim Kurbonov, Istam Shadmanov et al. · 0 citations
#software testing Dataset Open access Oct 2026

Artifact for "Memorised, Not Generated: Verbatim Recall of Published Fixtures in LLM-Generated Test Data, Measured Across Five Models and Eight Public Schemas"

Complete artifact for the empirical study "Memorised, Not Generated: Verbatim Recall of Published Fixtures in LLM-Generated Test Data, Measured Across Five Models and Eight Public Schemas" (submitted to Empirical Software Engineering). Version 1.2.1 adds experiment E1c: the copy-rate measurement (exact-text match, near...

Vijay Prasad Javvadi · 0 citations
#large language models Open access Oct 2026

Causal Encyclopedia

Causal Encyclopedia — Cumulative PASS295 A Canonical Knowledge Infrastructure for LLM-Assisted Reasoning The Causal Encyclopedia is a cumulative, machine-oriented knowledge infrastructure designed primarily to support large language models, reasoning systems, and human–AI research workflows. Rather than functioning as...

Son David Bolduc · 0 citations
#software testing Open access Oct 2026

SSX360/matrixscroll: Matrix Scroll 0.11.0: protocol v2

Matrix Scroll 0.11.0 introduces protocol v2 with mandatory pure, hedged ML-DSA-87 signatures, attached COSE_Sign1 and RFC 8785 canonical JSON. SHA-384 ledgers use RFC 9162 trees and authenticated checkpoint coverage. Python, Rust, Node WebCrypto and the bundled browser verifier share CONSISTENT / INDETERMINATE / INCONS...

SSX360 · 0 citations
#software testing Open access Oct 2026

The recoverable resolution of virtual-cell prediction — Analysis code

This record contains the analysis code and figure-level source tables for “The recoverable resolution of virtual-cell prediction.” The study shows that molecular measurement detail can exceed the predictive detail supported by the information available at prediction time. It establishes recoverable resolution as a desi...

Y. B. Huang · 0 citations
#software testing Dataset Open access Oct 2026

Replication materials for "Spatial autocorrelation of suicide mortality worldwide at three time points spanning the COVID-19 pandemic period: 2019, 2021, and 2023"

Replication materials (data and code) for the manuscript "Spatial autocorrelation of suicide mortality worldwide at three time points spanning the COVID-19 pandemic period: 2019, 2021, and 2023", submitted to PLOS Global Public Health (PGPH-D-26-02463). The package contains: (1) the analytical dataset for 203 countries...

José Eduardo Chiriboga Varea · 0 citations
#software testing Open access Oct 2026

xfeeds: an independence-aware public threat intelligence feed

xfeeds is a self-updating, open threat intelligence pipeline that collects, normalizes, enriches, scores, filters, and publishes malicious-IP indicators in CSV, JSON, STIX 2.1, MISP, and firewall rule sets. A GitHub Actions pipeline targets six-hour refreshes and publishes a dashboard through GitHub Pages; scheduling i...

Neil Weitzel · 0 citations
#software testing Open access Oct 2026

SuRT-GeoHarmonizer: auditable administrative-scale Earth-data harmonization and provenance labelling

SuRT-GeoHarmonizer is an open R and Python command-line workflow for converting heterogeneous environmental rasters into consistent, provenance-labelled administrative-unit GeoJSON layers. It provides a provider-agnostic raster and polygon interface, Rwanda reference builders for CHIRPS rainfall, ERA5-Land temperature,...

TUYISHIME AUDRE PRINCE · 0 citations
#software testing Dataset Open access Oct 2026

Data and scripts for "Δ-HGPR: A Hierarchical Gaussian Process Framework for Active-Learning Δ-Potentials"

Data and scripts underlying "Δ-HGPR: A Hierarchical Gaussian Process Framework for Active-Learning Δ-Potentials" (T. Nakajima, submitted to J. Chem. Theory Comput.). This version corresponds to the revised manuscript and supersedes version 1. The archive contains the inputs, labeled data, trained difference potentials,...

Takahito Nakajima · 0 citations
#large language models Open access Oct 2026

AutoQA-MAS: An LLM-Based Multi-agent System for Autonomous Website End-to-End Testing

Manual execution of website end-to-end (E2E) test cases is a repetitive, labour-intensive process that creates bottlenecks in modern software development workflows. We investigate the feasibility of automating this process using AutoQA-MAS, a large language model (LLM)-based multi-agent system (MAS) comprising four spe...

Adriana‐Simona Mihăiţă, Kai Sor · 0 citations

From tech blogs

See all →
MIT News · Artificial Intelligence Oct 2, 2026

Documenting the tech worker movement

Writing as a participant and researcher, PhD student JS Tan SM ’22 has co-authored a new book about the rise of tech worker protests and the employer backlash that followed.

GPT-Lab Sep 23, 2026

Requirements Don’t Live in Isolation: What We’re Exploring with Req-Space

Requirements in large systems rarely exist in isolation. Their meaning depends on the wider project context - other requirements, policies, decisions, tests, and implementation details. That becomes especially important when AI is used for review, because spotting a possible conflict or gap is only the beginning. ReqSpace explores how AI, visualisation, and connected project context can help reviewers understand those findings, trace the relationships behind them, and focus on the questions that…

GPT-Lab Sep 17, 2026

Beyond Prompt Engineering: The Role of Tacit Knowledge in Software Engineering

AI is making software generation faster, but speed does not remove the need for expertise. As more work is delegated to AI, tacit knowledge may become one of the most important human advantages in software engineering. The post Beyond Prompt Engineering: The Role of Tacit Knowledge in Software Engineering appeared first on GPT-Lab.

MIT News · Artificial Intelligence Aug 17, 2026

Q&A: Rethinking how innovation happens

In his latest book, Professor Eugene Fitzgerald examines the forces that turn breakthroughs into value — and why innovation resists simple formulas.

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.