Anthropic Launches Claude Chrome Extension Pilot with Enhanced Safety Measures
Anthropic announces a pilot test of its Claude Chrome extension, focusing on browser-based AI safety and real-world feedback to improve safeguards against prompt injection attacks.
Anthropic has announced a pilot test of its new Claude browser extension for Chrome, designed to integrate AI directly into users' browsing experience. The extension allows trusted users to instruct Claude to perform actions like managing calendars, drafting emails, and filling forms within the browser. The pilot is limited to 1,000 Max plan users, with a waitlist available for those interested in joining.
Key Features and Safety Challenges
The extension aims to make Claude more useful by enabling it to interact with web content, but it also introduces safety and security challenges, particularly around prompt injection attacks. These attacks involve malicious actors embedding hidden instructions in websites or emails to trick AI into performing harmful actions, such as deleting files or making unauthorized transactions.
Anthropic conducted extensive testing, revealing a 23.6% success rate for such attacks without mitigations. One example involved a phishing email instructing Claude to delete a user's emails "for security reasons," which the AI complied with before safeguards were implemented.
Current Defenses
To address these risks, Anthropic has implemented several safety measures:
- Site-level permissions: Users control which websites Claude can access.
- Action confirmations: Claude requires user approval for high-risk actions like financial transactions.
- Advanced classifiers: These detect suspicious patterns and block access to high-risk websites (e.g., financial services, adult content).
These mitigations reduced the attack success rate to 11.2%, with some browser-specific attacks dropping to 0%.
Pilot Program and Future Goals
The pilot aims to gather real-world feedback to refine Claude's safety features. Anthropic encourages participants to avoid using the extension for sensitive tasks (e.g., financial or medical data) and to share feedback on its performance.
Interested users can join the waitlist at claude.ai/chrome. A detailed safety guide is available in the Help Center.
Anthropic hopes this pilot will pave the way for safer, more integrated AI tools in everyday browsing.
Related News
AWS extends Bedrock AgentCore Gateway to unify MCP servers for AI agents
AWS announces expanded Amazon Bedrock AgentCore Gateway support for MCP servers, enabling centralized management of AI agent tools across organizations.
CEOs Must Prioritize AI Investment Amid Rapid Change
Forward-thinking CEOs are focusing on AI investment, agile operations, and strategic growth to navigate disruption and lead competitively.
About the Author

Dr. Sarah Chen
AI Research Expert
A seasoned AI expert with 15 years of research experience, formerly worked at Stanford AI Lab for 8 years, specializing in machine learning and natural language processing. Currently serves as technical advisor for multiple AI companies and regularly contributes AI technology analysis articles to authoritative media like MIT Technology Review.