Agentic systems are now being widely used to orchestrate tools and reason over long contexts. However, the improving capabilities of the large language models powering these agents also create new attack surfaces for indirect prompt injection. In particular, an attacker may not need to place a complete malicious instru...
Michael E. M. Lee, Zhi-Peng Wei, Yue Dong et al.· 0 citations
This work establishes a foundation for trustworthy natural language interfaces by enabling AI systems to recognize when generated specifications may not be reliable, and demonstrates improved translation reliability, robustness under the evaluated cross-tier shifts, and effective uncertainty-aware abstention.
Yixuan Wang, Licheng Luo, Yu Fu et al.· 1 citation
Student-Aware CoT Optimization for Recommendation Distillation (SCOReD), a CoT optimization framework tailored to recommendation that first parses each teacher trace into typed segments and uses the student LLM's attention to score the importance of each segment.
H. S. Shahgir, Yufei Li, Xiaohan Wei et al.· 0 citations
This work proposes a proof-of-concept pipeline that delivers on building an AI system that can ingest multi-modal data for railway crossings and provide safety assessment and scores that align with expert opinion and with safety scoring used by the Federal Railroad Administration.
Paimon Goulart, Chansong Lim, Nícolas Roque dos Santos et al.· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.