Boosting AI-based speech performance assessment through pairwise comparative scoring of crowd data
Open-ended performance assessments can capture intellectual status that is difficult to measure using highly constrained response formats. However, their practical use is limited by the difficulty of obtaining reliable, scalable, and interpretable scores. Recent advances in large language models provide new opportuniti...