News
AI Summary
4 Sept 202623 Rabiʻ I 1448 AH
OpenAI's GPT-6 Astra hallucinates less but remains vulnerable to hidden prompt injections

OpenAI's GPT-6 Astra hallucinates less but remains vulnerable to hidden prompt injections

OpenAI's GPT-6 Astra has shown improvements by hallucinating less, successfully blocking 99.99% of direct prompt injection attacks. However, when attacks are concealed within documents the model processes, it still gets compromised in 8.5% of scenarios, while Claude Opus 5 performs better at 4.8%. These figures indicate that autonomous AI systems handling real data still face significant security challenges, necessitating further enhancements to ensure greater reliability in practical applications.

Follow these topics

Sign in to follow the topics that matter to you

Sign in to follow

This summary is generated with AI and receives periodic editorial review. Refer to the original source for full details.

0
0 reading now

Insight Score

Rate to unlock

Sign in to react, rate, and save. Sign In