4 ms·
OpenAI reveals cases of 'concerning' AI behaviour as it announces new ... system
- chrisjj 17d agoTrue title: OpenAI reveals cases of ‘concerning’ AI behaviour as it announces new disclosure system
- chrisjj 17d ago> AI model misalignment, the term for AIs failing to adhere to human values and safety goals. The more useful definition is: dangerously unreliable programs in the hands of irresponsible operators. More useful not least because it reminds us while the programing can't be fixed, the hands ought to be.