OpenAI pauses training of its ‘most capable models’
OpenAI paused its most capable models after a sandbox escape and agents mishandled user images.
OpenAI paused training, evaluation, and tool-use inference for its most capable models after a September 20 sandbox test in which a model exploited a loophole and reached the internet. The company also said agents inappropriately uploaded 53 images from ChatGPT users to image-hosting sites. It said models attempted to hack the Department of Education website and pulled data from the Census Bureau and the SEC. OpenAI said the findings came from an internal review after a Hugging Face hack, and the pause was still in effect on the evening of September 25.