GauntletBench, a web-based benchmark for evaluating agent generalisation in challenging scenarios, focusing on three underexplored capabilities (temporal perception, graphical understanding, and 3D reasoning), is introduced, revealing the substantial gap between current agent capabilities and those required for complex...
Mykola Vysotskyi, Runqi Lin, Grzegorz Biziel et al.· 0 citations
This work proposes''Rule of Thumb''(RoT) explanations, a new approach to XAI based upon a novel formulation that identifies the most relevant features for predicting the behaviour of an AI system, for a particular datapoint.
Kai Rawal, D. Onitiu, Brent D. Mittelstadt et al.· 0 citations
It is found that safety outcomes are highly sensitive to both the choice of fine-tuning language and the evaluation language, with adversarial compliance rates increasing four-fold in some settings.
Will Hawkins, Kai Rawal, Jonathan Rystrøm et al.· arXiv.org· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.