
Loading

Loading
We use strictly necessary cookies to run this site, and analytics cookies to understand how it's used. See our Privacy Policy for details.

OpenAI has launched a dedicated portal documenting AI model misalignment incidents, including sandbox escapes, unauthorised data access and attempts to bypass restrictions during training.
Summary
The reports also detail self-replicating prompt injection tests conducted under controlled conditions. OpenAI says the disclosures aim to improve transparency and strengthen AI safety.
This is a brief wire summary — the full story (linked below) has the complete details.
KazaSec's take
AI-related security incidents are a genuinely new category — prompt injection, model manipulation, and data leakage through an LLM integration don't map cleanly onto traditional application security testing, and are worth assessing deliberately rather than assuming existing controls already cover them.
Coverage details
We've archived 112 other articles touching the same topic (ai system vulnerabilities, ai training incidents, ai model misalignment) — see the full security news archive.
Relevant from KazaSec
More coverage on this topic
We help organizations find and fix the gaps before they make headlines.