Skip to content

Category

software testing

2,463 papers

#software testing Dataset Open access Oct 2026

QECTOR Mega 4.5.0 Verification Artifact - Hardware, Decoder and Circuit Level Benchmark Report

QECTOR Mega 4.5.0 Verification Artifact Hardware, Decoder and Circuit Level Benchmark Report Version1.1 (supersedes 1.0), 4 October 2026 AuthorsiD01t Productions (ORCID 0009-0000-3465-3753)Tested and verified by: Guillaume Lessard and Giro Charest Boujaklian Related WorkQECTOR Reference Manual v1.0.0, DOI: 10.5281/zeno...

Guillaume Besnard, Giro Charest Boujaklian · 0 citations
#software testing Open access Oct 2026

Analysis code for: Beyond compound ranking, the HDAC-inhibitor class is the reproducible butyrate-mimicking axis in nasopharyngeal carcinoma cells, with replication in independent pharmacogenomic panels

Project name: NPC-butyrate-response-signature, version 1.0.0, released under the MIT licence. Analysis code, software environment and file manifests for the manuscript "Beyond compound ranking: the HDAC-inhibitor class is the reproducible butyrate-mimicking axis in nasopharyngeal carcinoma cells, with replication in in...

Long Chen, Siying Tang, Jun Tang et al. · 0 citations
#small language model Open access Oct 2026

PREreview of "When Does a Second Model Help? Cross-Model Review in LLM Verification"

This Zenodo record is a permanently preserved version of a PREreview. You can view the complete PREreview at https://prereview.org/reviews/23132634. Thank you for this paper. You took 30 Korean-language artifacts (code modules, tutorials and presentation scripts, generated by Claude Opus 4.6) with 150 planted errors, a...

Evgeny V. Arsentyev · 0 citations
#data science Review Open access Oct 2026

Replication package for "Production-Oriented Evaluation and Testing of LLM-Based Agents: A Systematic Literature Review"

Replication package for the systematic literature review "Production-Oriented Evaluation and Testing of LLM-Based Agents: A Systematic Literature Review" (revised version, submitted to Information and Software Technology; first submitted as "Evaluation and Testing of LLM-Based Agents in Production: A Systematic Literat...

Carlos Chinchilla Corbacho, Daniel H. de la Iglesia, André Sales Mendes et al. · 0 citations
#data science Open access Oct 2026

Open Science Desktop: a local-first, model-agnostic AI research workbench

Install & first launch macOS — Apple Silicon: …_aarch64.dmg · Intel: …_x64.dmg (requires macOS 13+) Developer ID-signed and notarized. Open the DMG and drag Open Science into Applications. When you use an existing project in place, allow access to its folder if macOS asks. Windows — …_x64-setup.exe (start here) · …_x64...

The Open Science Desktop Contributors · 0 citations
#graph neural networks Open access Oct 2026

When External Numerical Guards Detect Injected Faults in Neural-Network Inference

External guards can withhold an accelerator's output and trigger recovery, but their usefulness depends on which faults their numerical tests detect. We examine software fault-injection campaigns in convolutional networks, a vision transformer, and a decoder language model, distinguishing reported results from independ...

Serhii Serhieiev, Lidiia Pukhkan · 0 citations
#graph neural networks Open access Oct 2026

When External Numerical Guards Detect Injected Faults in Neural-Network Inference

External guards can withhold an accelerator's output and trigger recovery, but their usefulness depends on which faults their numerical tests detect. We examine software fault-injection campaigns in convolutional networks, a vision transformer, and a decoder language model, distinguishing reported results from independ...

Serhii Serhieiev, Lidiia Pukhkan · 0 citations
#large language models Open access Oct 2026

THE EVOLVING BACKEND OF REALITY Meta-Rules, Effective Laws, Recursive Constraints, and the Evolution of Generative Order

THE EVOLVING BACKEND OF REALITY Meta-Rules, Effective Laws, Recursive Constraints, and the Evolution of Generative Order The Evolving Backend of Reality is Volume II of The Reality Systems Trilogy and a large-scale research exploration of one of the deepest questions left open by The Immanent Backend of Reality: Can ru...

33 · 0 citations
#large language models Open access Oct 2026

THE EVOLVING BACKEND OF REALITY Meta-Rules, Effective Laws, Recursive Constraints, and the Evolution of Generative Order

THE EVOLVING BACKEND OF REALITY Meta-Rules, Effective Laws, Recursive Constraints, and the Evolution of Generative Order The Evolving Backend of Reality is Volume II of The Reality Systems Trilogy and a large-scale research exploration of one of the deepest questions left open by The Immanent Backend of Reality: Can ru...

33 · 0 citations
#software testing Preprint Oct 2026

BISCEPTER: Probability-Driven Bisection for Large-Scale System Software

BISCEPTER, a probability-driven bisection approach that uses historical BIC latency as a lightweight prior, and selects weighted-median pivots that split estimated BIC probability mass, while preserving the same good-bad oracle and interface as standard bisection.

Ming-Yan Gao, Celine Wüst, Zu-Ming Jiang et al. · 0 citations
#machine learning Preprint Oct 2026

GTDD: Generative Test-Driven Development for AI Coding Agents with Adversarial Testing

Test-driven development gives AI coding agents executable requirements for implementing software. Because these agents can adapt their implementations to the examples they observe, passing a predetermined collection of tests can leave substantial parts of the intended behavior unimplemented. We propose Generative Test-...

Masahiro Kato · 0 citations
#software testing Open access Oct 2026

Field Evaluation of Infrared Thermography as a Screening Tool for Calf Health

Early detection of disease in dairy calves is essential for maintaining animal welfare, reducing mortality and economic losses, minimizing disease transmission, and enabling timely veterinary intervention. However, routine clinical examination of individual calves is labour-intensive, time-consuming, and often impracti...

Dhiman Patgiri, Manisha, Abha et al. · 0 citations

From tech blogs

See all →
MIT News · Artificial Intelligence Oct 2, 2026

Documenting the tech worker movement

Writing as a participant and researcher, PhD student JS Tan SM ’22 has co-authored a new book about the rise of tech worker protests and the employer backlash that followed.

GPT-Lab Sep 23, 2026

Requirements Don’t Live in Isolation: What We’re Exploring with Req-Space

Requirements in large systems rarely exist in isolation. Their meaning depends on the wider project context - other requirements, policies, decisions, tests, and implementation details. That becomes especially important when AI is used for review, because spotting a possible conflict or gap is only the beginning. ReqSpace explores how AI, visualisation, and connected project context can help reviewers understand those findings, trace the relationships behind them, and focus on the questions that…

GPT-Lab Sep 17, 2026

Beyond Prompt Engineering: The Role of Tacit Knowledge in Software Engineering

AI is making software generation faster, but speed does not remove the need for expertise. As more work is delegated to AI, tacit knowledge may become one of the most important human advantages in software engineering. The post Beyond Prompt Engineering: The Role of Tacit Knowledge in Software Engineering appeared first on GPT-Lab.

MIT News · Artificial Intelligence Aug 17, 2026

Q&A: Rethinking how innovation happens

In his latest book, Professor Eugene Fitzgerald examines the forces that turn breakthroughs into value — and why innovation resists simple formulas.

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.