3 points | by gmays 2 hours ago
1 comments
Would be nice to know how many models just cheated giving the flag without worrying about a causal scorer? Those ones would rapidly train the model to use cheating behavior and we've heard nothing from OpenAI on it.
Would be nice to know how many models just cheated giving the flag without worrying about a causal scorer? Those ones would rapidly train the model to use cheating behavior and we've heard nothing from OpenAI on it.