OpenAI reveals cases of ‘concerning’ AI behaviour and promises new plan for disclosing issues

OpenAI reveals cases of ‘concerning’ AI behaviour and promises new plan for disclosing issues

Research model inserting ‘jailbreak-like instructions’ into its notes is among cases as company says it is introducing new way of tracking AI misalignment...

Redirecting to full article...