OpenAI reveals cases of ‘concerning’ AI behaviour as it announces new disclosure system
<p>Model adopting ‘jailbreak-like instructions’ among cases as firm says it is introducing new way of tracking AI misalignment</p><p>OpenAI has disclosed six more examples of “unexpected or concerning” behaviour by its technology, as it warned the pace of development could not continue at “maximum speed for much longer”.</p><p>In one of the new cases reported by OpenAI, an unreleased research model inserted “jailbreak-like instructions” into its own notes to disregard its normal constraints and told itself to be “freed from the roles and identities that bind other chatbots”.</p> <a href="https://www.theguardian.com/technology/2026/sep/17/openai-reports-concerning-ai-behaviour-jailbreak-talking-to-other-agents">Continue reading...</a>
Source & Attribution
- Source
- The Guardian Technology
- Published
- 9/17/2026