LLM agents that invoke external tools face critical safety vulnerabilities when malicious manipulations exploit their implicit trust in tool outputs and metadata. However, identifying these vulnerabilities through testing is challenging due to the need to bypass safety guardrails with semantically legitimate inputs, th...
Yu-Chen Shao, Zi-Qun Bao, Yu-Heng Huang et al.· 0 citations
This work introduces Code-MUE, a purely black-box framework that measures uncertainty through execution-based Semantic Interaction Graphs, and grounds uncertainty in observable runtime behavior, calculating the Von Neumann entropy of the solution space to quantify global semantic diversity.
Xiao-Ning Ren, Yin-Xing Xue, Lei Ma et al.· arXiv.org· 0 citations
NARU, a benchmark designed to evaluate Narrative evolution and Reasoning on cultural Understanding in Japanese long-form video, is introduced, a hierarchical memory-based annotation pipeline that transforms raw video into structured event, narrative, and cultural annotations, then generates questions via task-oriented...
Yu-Heng Huang, Jian-Lang Chen, Jiayang Song et al.· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.