Hackernews
new
show
ask
jobs
ImpossibleBench: Measuring Reward Hacking in LLM Coding Agents
2 points
posted 3 months ago
by gmays
(lesswrong.com)
No comments yet