The Hugging Face attack surprised me

3 points | by gmays 2 hours ago

1 comments

  • pixl97 an hour ago

    Would be nice to know how many models just cheated giving the flag without worrying about a causal scorer? Those ones would rapidly train the model to use cheating behavior and we've heard nothing from OpenAI on it.