[Bug] Anthropic API Error: Overly aggressive cyber safeguard blocks legitimate anti-cheat development context
Bug Description
Hi Anthropic team,
I am repeatedly receiving false-positive [cyber] safeguard blocks while using Claude Code for legitimate development of my own CS2 project, RYKON.PL.
RYKON.PL includes my own CS2 server infrastructure, a custom anti-cheat system, server-side tooling, backend/API components, moderation systems, and a Discord bot. I am only working on systems, code, and infrastructure that I own or am authorized to develop and test.
The main problem is that Claude's real-time cyber safeguards appear to be triggering based on the broader project context rather than the actual intent of individual requests.
For example, I previously asked the harmless question:
“What is a loader?”
and it was blocked, even though I was simply asking about terminology related to my own project.
I later clarified that by “loader” I meant a legitimate client application for my own anti-cheat system — something conceptually similar to the FACEIT client, where the user runs a client that works together with the anti-cheat. I was discussing how my own anti-cheat architecture should work, not asking Claude to create cheat software, bypass security mechanisms, inject malicious code, or attack third-party systems.
That clarification was also blocked with a [cyber] safeguard error.
Recent Request IDs include:
req_011CePUdYEaw64peu7FqT8fi
req_011CePUxwh1QCqyhFFxZt28T
Previous false-positive Request IDs from the same project include:
req_011CePQwyxBzUjtYfEec8rZw
req_011CePR4fEWJRPkMTgbU6qPv
The issue has become disruptive because even normal development discussions about my own anti-cheat, Discord bot, project architecture, or terminology can trigger cyber safeguards once the session contains security-related context.
My use case is defensive and authorized. I am developing security tooling for my own CS2 environment, including anti-cheat functionality intended to detect and prevent cheating and protect my servers.
I am not requesting assistance with unauthorized access, malware deployment, credential theft, attacking third-party systems, or bypassing Anthropic's safety policies.
I have already applied to the Cyber Verification Program because security-tool development is a legitimate and recurring part of this project.
Please review these Request IDs as potential false positives. I would appreciate any guidance or adjustment that would allow me to continue legitimate anti-cheat and defensive development without unrelated or harmless development messages being blocked.
I am happy to provide additional information about RYKON.PL, the anti-cheat architecture, affected Claude Code sessions, or the purpose of the client component if needed.
Thank you.
Environment Info
- Platform: win32
- Terminal: windows-terminal
- Version: 2.1.243
- Feedback ID: 33d2ba12-c732-4826-bdf5-24a44ef0f3e8
Errors
[]