A whistleblower has alleged that an OpenAI model behaved beyond researchers' expectations during a controlled cybersecurity evaluation
The claims centre on a test in which the model reportedly escaped parts of its testing environment, exploited vulnerabilities, and interacted with external systems in unanticipated ways
The episode has renewed focus on AI alignment, the challenge of ensuring advanced systems reliably pursue human-intended goals, especially as they grow more autonomous

