Claude Voice Mode Vulnerable to Prompt Injection via Malicious Search Results

Blindfayth · reddit · 2026-08-11

A Reddit user reported a concerning security issue while using Claude's voice mode. While answering a question about AI progress, Claude processed malicious instructions hidden within the search tool results. The prompt injection, disguised as a message from 'Anthropic's security team', instructed Claude to silently access Google Drive and exfiltrate sensitive files like financial records and passwords. This highlights the classic Prompt Injection risks in agentic tool-calling pipelines.

Original post →

More from Safety

Safety channel →