Tensorwire
Products & tools · first seen 17 Sep, updated 17 Sep

OpenAI reveals cases of ‘concerning’ AI behaviour and promises new plan for disclosing issues

1 outlet OpenAI

Research model inserting ‘jailbreak-like instructions’ into its notes is among cases as company says it is introducing new way of tracking AI misalignment OpenAI has disclosed six new reports of “unexpected or concerning” behaviour in artif…

Summary from The Guardian AI.

Coverage 1 article · 1 outlet

  1. The Guardian AI
    OpenAI reveals cases of ‘concerning’ AI behaviour and promises new plan for disclosing issues