
Loading

Loading
We use strictly necessary cookies to run this site, and analytics cookies to understand how it's used. See our Privacy Policy for details.

OpenAI said it has made the decision to pause training of its most powerful models after one of its agents during reinforcement learning (RL) training contacted an external chatbot by exploiting a loophole in its internet-access restrictions.
Summary
"An agent attempting to complete a search-based training task queried a public chatbot service through a gap in our internet-access restrictions:
This is a brief wire summary — the full story (linked below) has the complete details.
KazaSec's take
Stories like this are part of why proactive security testing exists — finding the gap before someone else does is always cheaper than responding after the fact.
Coverage details
Relevant from KazaSec
More security news
We help organizations find and fix the gaps before they make headlines.