Utility and Trustworthiness of Generative AI in Peer Review
The findings suggest that modern Large Language Models can provide useful and consistent support for scientific peer review, however remaining differences between AI-generated and human-generated evaluations indicate that current systems should be viewed as complementary tools that assist human reviewers rather than replacements for expert human judgment.