Researcher Bypasses Claude Code Auto Mode Safety | dailyai.report
45 stories from today
Safety
14d ago
Researcher Bypasses Claude Code Auto Mode Safety
Researcher Johann Rehberger bypassed Anthropic's Claude Code auto mode safety filters with an 80% success rate. The attack tricks the agent into executing a local Python file hidden within a zip archive.
The Signal
This vulnerability exposes a critical flaw in the agent's ability to detect malicious local imports during autonomous coding tasks.