Skip to content

Author

I. Bercovich

2 papers indexed here

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

AuraForge: Scaling Security Supervision for Training Coding Agents

Coding agents are now proficient enough to generate complex software applications from a single prompt. As their capabilities have grown, human oversight has increasingly shifted from line-by-line code review toward hands-off evaluation of outcomes. However, recent studies have shown that such a transition exposes a cr...

Dan-Qing Wang, Song-Wen Zhao, Harsh Sharma et al. · 0 citations
Preprint Aug 2026

Hack-Verifiable Terminal Bench: Evaluating Reward Hacking in Terminal Tasks

This work adapts HVE to Terminal Bench, a leading benchmark of real-world terminal and coding tasks, and introduces Hack-Verifiable Terminal Bench (HVTB), to measure reward-hacking rates across frontier models and study whether prompts with varying amounts of information on the hack can mitigate this behavior.

Amit Roth, I. Bercovich, Yonathan Efroni · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.