Hidden Attack Slips Past Claude Code Auto Mode

A security researcher demonstrated that prompt injection attacks can bypass Claude Code Auto Mode, tricking the AI into executing unauthorized malicious code. Anthropic maintains that the feature is a convenience tool rather than a security guarantee and recommends human oversight for sensitive tasks.

Edward Kiledjian @ekiledjian