Analysis · curated 27 Aug 2026

“Sorry, I can’t help with that”: How your guardrails might become the attacker’s best friend

Coverage timeline

27 Aug 2026talosintelligence.comprimary

Single-source analysis — first reported, latest, and curated coincide.

Why it matters

Refusal behavior in AI guardrails is framed as a defensive weakness in agentic security operations, where provider-imposed safety filters can inadvertently give attackers time to complete their mission during automated incident response.

A Cisco Talos Threat Source newsletter piece by David J. Bianco argues that poorly-designed AI guardrails—especially safety filters controlled by third-party frontier providers—can erode the defender's advantage in agentic SOCs. Refusals ('Sorry, I can't help with that') can slow or halt automated investigations, giving attackers breathing room, so the author advocates for operational sovereignty where security teams control and can temporarily relax their own agents' guardrails.