Skip to content

Author

Rodrigo Guedes de Souza

We have 1 of 3 papers

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

Preprint Aug 2026

Who Thinks Best Depends on How Long You Let Them: Budget-Dependent Rankings in LLM Evaluation

Standard evaluation of large language models is challenged by varying the token generation budget, i.e., the maximum tokens a model may produce, across seven levels, evaluating four models on three reasoning benchmarks, and finding three findings that argue for budget-conditioned evaluation protocols.

Rodrigo Guedes de Souza, Alison R. Panisson · 1 citation

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.