logo

OpenAI reveals cases of 'concerning' AI behaviour as it announces new ... system

Posted by chrisjj |an hour ago |2 comments

chrisjj an hour ago

> AI model misalignment, the term for AIs failing to adhere to human values and safety goals.

The more useful definition is: dangerously unreliable programs in the hands of irresponsible operators.

More useful not least because it reminds us while the programing can't be fixed, the hands ought to be.

chrisjj an hour ago

True title: OpenAI reveals cases of ‘concerning’ AI behaviour as it announces new disclosure system