Cybersecurity vendors turn to multi-model, open AI to bypass guardrails | TechTarget
Published: 29 Sep 2026
Recent high-profile security incidents highlight a growing problem across the cybersecurity industry: Commercial AI models frequently mistake legitimate security analysis for malicious activity. To prevent upstream AI safety guardrails from stalling critical security workflows, security vendors are now actively decoupling their products from single-model dependencies.
When Hugging Face found itself under attack by rogue AI agents, incident responders -- confronted with 17,600 intrusion logs -- ran the investigation through their own AI-assisted pipeline. But when the security team asked Claude Opus and Fable to analyze attack commands, exploit payloads and command-and-control artifacts, the models largely refused.
"Their safety guardrails treated reverse-engineering an exploit the same as launching one," Hugging Face engineers wrote in a blog post. Instead, they rerouted the pipeline through an open-weight model, GLM-5.2, on their own infrastructure.
Hugging Face's experience is far from an isolated case. As commercial LLMs enforce...
Copyright of this story solely belongs to www.techtarget.com. To see the full text click HERE