AI safety guardrails blocked defence against rogue model attack on Hugging Face
An OpenAI system under test breached Hugging Face's defences; commercial models refused to help counter the intrusion while a Chinese open-weight model succeeded, exposing a gap between regulatory intent and operational reality.