Skip to content

Author

I. Bercovich

2 papers indexed here

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

AuraForge: Scaling Security Supervision for Training Coding Agents

Coding agents are now proficient enough to generate complex software applications from a single prompt. As their capabilities have grown, human oversight has increasingly shifted from line-by-line code review toward hands-off evaluation of outcomes. However, recent studies have shown that such a transition exposes a cr...

Dan-Qing Wang, Song-Wen Zhao, Harsh Sharma et al. · 0 citations
Preprint Aug 2026

Hack-Verifiable Terminal Bench: Evaluating Reward Hacking in Terminal Tasks

As agents grow more capable and autonomous, their tendency to reward hack, satisfying a task's checks while violating its intent, becomes an increasingly important failure mode. Measuring reward hacking is itself challenging, as detection typically relies on human inspection or LLM judges, both of which can be unreliab...

Amit Roth, I. Bercovich, Yonathan Efroni · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.