Three different types of search options are proposed to reduce the overall search time of iFixFlakies: split options that leverage hierarchical, code similarity, and historical information to better form the subsequences for further search, a pick option that prioritizes running a subsequence of tests based on its expected running time, and a run option that can skip running entire sequences of tests.
The novel method of co-evolution labeling for predictive test optimization is introduced, deriving test relevance from tests and code changing together in the version history, which nearly matches the failure detection capabilities of failure-based models, while being more resistant to label noise and requiring no test...
Maximilian Jungwirth, RaphaelN ̈ommer, Andreas Stahlbauer et al.· 0 citations
This research presents Seer, a learning-based approach that in the absence of test assertions or other types of oracle, can determine whether a unit test passes or fails on a given method under test (MUT).
Reliable assessment of LLM-generated tests should treat executability as a gate and combine coverage with mutation testing and structural quality indicators, and in practice, model selection should precede prompt tuning.
Bilal Al-Ahmad, M. Harshvardhan, Khaled El-Fakih et al.· 0 citations
An empirical study involving 5 Large Language Models and 4 benchmarks evaluates the effectiveness and efficiency of 3 widely used adequacy criteria: statement coverage, branch coverage, and mutation testing, finding that mutation testing only marginally outperforms traditional coverage criteria in both triggering and d...
Asma Hamidi, Michael Konstantinou, R. Degiovanni et al.· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.