Skip to content

1 paper indexed here

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

#artificial intelligence Review May 2026

BenGER: Benchmarking LLM Systems on Subsumption-Based Legal Reasoning in German Law

BenGER (Benchmark for German Law), a benchmark and dataset for evaluating LLM systems on subsumption-based legal reasoning in German law, is introduced and 12 contemporary LLM systems are evaluated with a rubric-aligned LLM-as-a-Judge cross-validated against a multi-rater human-grading layer.

Sebastian Nagl, A. Mayrhofer, Martin Heidebach et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.