Decomposing LLM-Based Testing with Agent Skills: A Case Study on Numerical Inconsistencies
LLM-based testing can expose subtle reliability issues in numerical software, but monolithic prompts make it hard to inspect which guidance drives effectiveness. We study whether Agent Skills can make such workflows more explainable by factoring procedural knowledge into composable testing components. Using numerical i...