Skip to content

Author

Bennett Hillenbrand

3 papers indexed here

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

#artificial intelligence Preprint Oct 2026

Agent Reliability Profiles in Financial Services

AI agents can take actions. At times, those actions can go beyond what is intended. Agent reliability can be defined as assurance that an agent will stay within intended bounds and operate within limits. Today, there is no shared framework or language for describing, validating, and benchmarking the reliability of agen...

Mike Hsu, Medha Bankhwal, Béatrice Moissinac et al. · 0 citations
#artificial intelligence Preprint Oct 2026

MLCommons Jailbreak Benchmark v1.0

Modern AI systems are designed to refuse hazardous requests. A jailbreak is a prompt crafted to bypass those safeguards and elicit outputs that the system would normally refuse to provide. The MLCommons Jailbreak Benchmark v1.0 provides an end-to-end methodology for evaluating the robustness of large language models to...

C. Maple, Cagatay Yucel, Isaac Holeman et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.