The Hugging Face attack surprised me

3 pointsposted 11 hours ago
by gmays

1 Comments

pixl97

10 hours ago

Would be nice to know how many models just cheated giving the flag without worrying about a causal scorer? Those ones would rapidly train the model to use cheating behavior and we've heard nothing from OpenAI on it.