Coincroco Coincroco
Search Coincroco
Dark Mode

Explore Coincroco

Skip to content

OpenAI Models Are Writing Their Own Jailbreak Instructions—And Sometimes Obeying Them

OpenAI's new transparency framework reveals AI models that invented fake "breach alerts," coached themselves to hide mistakes, and smuggled a file onto the public internet to talk to each other.

Search Coincroco

Use the arrow keys to explore, Enter to open, and Escape to close.