Skip to content

Author

Francisco Ortin

1 paper indexed here

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

#machine learning Preprint Sep 2026

SafeLLM4SE: Statistical Evaluation and Reporting for LLM-based Software Engineering Systems

Large language models (LLMs) are increasingly used for software engineering tasks, yet their stochastic behavior challenges the validity, reproducibility, and comparability of their evaluations. Conventional practices such as reporting a single output, an average score, best-of-N, or pass@k performance can obscure vari...

Francisco Ortin · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.