News
AI Summary
26 Jun 202611 Muharram 1448 AH
What happened after 2,000 people tried to hack my AI assistant

What happened after 2,000 people tried to hack my AI assistant

Fernando Irarrázaval launched a challenge on hackmyclaw.com to test his AI assistant OpenClaw's ability to protect secrets. Despite 6,000 attempts and a $500 token spend, no one leaked any information, although his Google account was suspended due to excessive inbound emails. The underlying model, Opus 4.6, employed strict anti-prompt-injection rules to prevent data leaks or command execution via email. These efforts reflect significant advancements in training models to resist injection attacks effectively.

Follow these topics

Sign in to follow the topics that matter to you

Sign in to follow

This summary is generated with AI and receives periodic editorial review. Refer to the original source for full details.

0
0 reading now

Insight Score

Rate to unlock

Sign in to react, rate, and save. Sign In