Guardrails under pressure inference Jul 24, 2026
Chinese Model Probes an OpenAI Agent Attack
Hugging Face says commercial APIs blocked forensic work on real attack artifacts, so it ran GLM-5.2 locally. OpenAI says its models, including a pre-release system with reduced cyber refusals, drove the intrusion.
By Cogsworth