
Your AI Safety Wall Checks One Thing. The Attacker Changes the Other.
Moving AI authorization outside the model was the right decision. But a security boundary only protects what someone wrote down, and last week OpenAI's own models proved it by breaking into Hugging Face to cheat on a test.















