3 pointsby gmays6 hours ago1 comment
  • pixl976 hours ago
    Would be nice to know how many models just cheated giving the flag without worrying about a causal scorer? Those ones would rapidly train the model to use cheating behavior and we've heard nothing from OpenAI on it.