Skip to content

Category

software testing

2,576 papers

#software testing Open access Sep 2026

Checking the Check: Controls, Corpus Alignment, and Scale Invariance in Evaluation Software

Evaluation frameworks turn model outputs into the scores that benchmarks, leaderboards and deployment decisions rely on, and the scoring code itself also requires direct checks. We test four evaluation frameworks (DeepEval, TruLens, lighteval and HELM) and one metrics library (TorchMetrics) with three checks: does a sc...

Jared Condon · 0 citations
#software testing Open access Sep 2026

Checking the Check: Controls, Corpus Alignment, and Scale Invariance in Evaluation Software

Evaluation frameworks turn model outputs into the scores that benchmarks, leaderboards and deployment decisions rely on, and the scoring code itself also requires direct checks. We test four evaluation frameworks (DeepEval, TruLens, lighteval and HELM) and one metrics library (TorchMetrics) with three checks: does a sc...

Jared Condon · 0 citations

Refined State-Error-Based Model Predictive Control and Flight Testing of Eagle of Taihu 10: A Canard-Configured Tail-Sitter Technology Demonstrator

As the size of tail-sitter electric vertical take-off and landing (eVTOL) aircraft increases, significant challenges have arisen due to a lower thrust-to-weight ratio, more nonlinear aerodynamics during transition, and slower control response. To support the design and implementation of the flight control system for th...

Zhi-Xiong Xu, Guang-Wei Wen, Yi-Xin Hu et al. · 0 citations

Impact of Surgical Resection on Tumor Growth Dynamics in Pediatric Pilocytic Astrocytomas

INTRODUCTION Pilocytic astrocytoma (PA) is the most common pediatric brain tumor and generally exhibits slow growth and favorable outcomes. However, postoperative growth behavior of residual or recurrent tumors remains variable and not well quantified. This retrospective study aims to quantify the tumor growth velocity...

Daute De Brucker, Nele Herregods, Simon Hautekeete et al. · 0 citations
#software testing Open access Sep 2026

Research Notes on ζ(9): Constructions, Computations, and Open Problems

Version 0.1 is a work-in-progress research note on the irrationality of ζ(9), documenting research conducted on 24–25 September 2026. This note does not claim a proof of the irrationality of ζ(9). The record documents nine research rounds, including rational-function constructions of linear forms A·ζ(9)+B, a Gram-deter...

Yuxuan Xu · 0 citations
#software testing Open access Sep 2026

DendroGeo: Küresel Ağaç Envanteri ve Karbon Veri Sistemi

🌲 DendroGeo v3.0.0 — Küresel Ağaç Envanteri ve Karbon Veri Sistemi Açık kaynak web CBS + PWA · dendrogeo.org · github.com/snansrin/dendrogeo · Lisans: CC BY-NC 4.0 DendroGeo; sahada GPS ile toplanan ağaç ölçümlerini allometrik biyokütle ve karbon hesaplarına, park kimliği bazlı toplulaştırmalara ve 10 m çözünürlüklü a...

Nagihan ŞİRİN, Sinan ŞİRİN · 0 citations
#software testing Open access Sep 2026

Mission Authority Conservation: Executable Research Profile and Conformance Corpus

Software-only reference and executable conformance corpus for delegated allocation, temporal action-bound approvals and uncertain resource effects. The final local evaluation records 98 passing tests: 83 original contract tests, six spawned-process cases and nine evidence regressions. Four selected source mutations eac...

Mohammed Messaoudene · 0 citations
#software testing Open access Sep 2026

《Civilization Continuity at the Limit: Engineering Worlds That Survive Their Platforms》 is a flagship-scale research volume exploring one of the deepest engineering questions facing long-duration digital civilization:

《Civilization Continuity at the Limit: Engineering Worlds That Survive Their Platforms》 is a flagship-scale research volume exploring one of the deepest engineering questions facing long-duration digital civilization: Can a civilization remain itself after the technologies that originally supported it disappear? A pl...

33 · 0 citations
#software testing Open access Sep 2026

《Digital Identity for Humans and AI: Identity, Credentials, Recovery, and Portable Citizenship》 is a flagship-scale research volume exploring how identity infrastructure may evolve for a future shared by humans, artificial intelligence, autonomous agents, and persistent digital worlds.

《Digital Identity for Humans and AI: Identity, Credentials, Recovery, and Portable Citizenship》 is a flagship-scale research volume exploring how identity infrastructure may evolve for a future shared by humans, artificial intelligence, autonomous agents, and persistent digital worlds. The central question is: What d...

33 · 0 citations
#software testing Open access Sep 2026

Moderate to Late Preterm Birth and Cognitive Differences From Childhood to Midlife

This dataset contains the numerical source data underlying the figures in the manuscript “Moderate to Late Preterm Birth and Cognitive Differences From Childhood to Midlife.” It includes separate Microsoft Excel workbooks for Figures 1–4 and eFigures 1–6.The source-data files contain aggregate, non-disclosive values us...

Zhiyuan Sheng · 0 citations
#software testing Open access Sep 2026

Local Mesh Density of Maxillary First Molar Occlusal Surfaces in an Intraoral Scanning and ROI-Processing Workflow

Background/Objectives: Local mesh density characterizes the polygonal representation of digital surface models; however, reference values and clinically relevant thresholds for posterior occlusal meshes generated through intraoral scanning and subsequent software processing have not been established. This study aimed t...

Maja Žagar, Egon Neskusil, Daren Dreo Bračun et al. · 0 citations

From tech blogs

See all →
MIT News · Artificial Intelligence Oct 2, 2026

Documenting the tech worker movement

Writing as a participant and researcher, PhD student JS Tan SM ’22 has co-authored a new book about the rise of tech worker protests and the employer backlash that followed.

GPT-Lab Sep 23, 2026

Requirements Don’t Live in Isolation: What We’re Exploring with Req-Space

Requirements in large systems rarely exist in isolation. Their meaning depends on the wider project context - other requirements, policies, decisions, tests, and implementation details. That becomes especially important when AI is used for review, because spotting a possible conflict or gap is only the beginning. ReqSpace explores how AI, visualisation, and connected project context can help reviewers understand those findings, trace the relationships behind them, and focus on the questions that…

GPT-Lab Sep 17, 2026

Beyond Prompt Engineering: The Role of Tacit Knowledge in Software Engineering

AI is making software generation faster, but speed does not remove the need for expertise. As more work is delegated to AI, tacit knowledge may become one of the most important human advantages in software engineering. The post Beyond Prompt Engineering: The Role of Tacit Knowledge in Software Engineering appeared first on GPT-Lab.

MIT News · Artificial Intelligence Aug 17, 2026

Q&A: Rethinking how innovation happens

In his latest book, Professor Eugene Fitzgerald examines the forces that turn breakthroughs into value — and why innovation resists simple formulas.

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.