seek, Self-Evaluative Exploration for Knowledge Retrieval, a training-free framework that addresses this limitation through iterative corpus interaction at test time through iterative corpus interaction at test time.
QueryRoute is introduced, a benchmark that freezes the expensive artifacts needed to study this inference-time decision problem reproducibly: original queries, generated variants, ranked lists under multiple retrievers, retrieval scores, and per-query oracle labels.
Hai-Son Le, Negar Arabzadeh, Amin Bigdeli et al.· 0 citations
Peer-review evaluation is increasingly being automated with LLM-as-a-judge metrics, but this creates a measurement risk. A review may receive a high score because it is fluent, organized, and polished, rather than because it provides a strong evaluation of the paper. This risk is especially important in AI-assisted rev...
Shakiba Amirshahi, Sajad Ebrahimi, Hai-Son Le et al.· 0 citations
Reviewerly's retrieval-centered infrastructure for auditing peer review at scale is presented and three deployed systems are described, including three deployed systems that analyzes hybrid human--AI authorship by disentangling the origin of ideas from the origin of text in peer reviews, enabling measurement of how hum...
Negar Arabzadeh, Sajad Ebrahimi, Alireza Daghighfarsoodeh et al.· Annual International ACM SIG...· 1 citation
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.