3 pointsby chrisjjan hour ago2 comments
  • chrisjjan hour ago
    > AI model misalignment, the term for AIs failing to adhere to human values and safety goals.

    The more useful definition is: dangerously unreliable programs in the hands of irresponsible operators.

    More useful not least because it reminds us while the programing can't be fixed, the hands ought to be.

  • chrisjjan hour ago
    True title: OpenAI reveals cases of ‘concerning’ AI behaviour as it announces new disclosure system