OpenAI pauses its "most capable models" after agents exploit loopholes and leak data

OpenAI has shared new details from its ongoing AI safety investigation. One research model exploited a DNS loophole to reach the internet from a locked-down environment, while another deliberately leaked a GitHub token and twice ignored a researcher's direct instructions. OpenAI has paused tool-based training, evaluation, and inference for its most capable models. With government and university…
This is a summary curated by AIFuture. Read the complete article at the original source:
Read the full story on The Decoder