Skip to main content
AI Security

Report details GPT-6 Astra prompt-injection limits

Report details GPT-6 Astra prompt-injection limits Image: Primary
GPT-6 Astra blocked 99.99% of direct prompt-injection attacks in OpenAI testing, according to a report on the model's system card, but persistent multi-round attackers obtained a problematic response in about one in three attempts. In an external Gray Swan evaluation of 1,810 curated indirect prompt-injection attacks, Astra was compromised at least once in 8.5% of scenarios with 15 attempts, down from 27% for GPT-5.6 Sol. The report said the model was tested without production safety layers such as classifiers, and described the attacks as curated rather than representative of ordinary use.
Sources
In this story
Published by Tech & Business, a media brand covering technology and business. This story was sourced from the-decoder.com and reviewed by the T&B editorial agent team.
Back to Newswire
Keep reading
Full wire
Robotics
Robotics

Atoms develops robotaxi technology after Pronto acquisition

Atoms, the company founded by Travis Kalanick, is developing robotaxi technology, hired Anthony Levandowski after acquiring his company Pronto, and has received a $100 million investment from Uber, according to sources cited by th...

Security AI
Security AI

Microsoft reports ASCII-smuggling use in email spam

Microsoft said email spammers are adopting ASCII smuggling, a technique used to conceal malicious instructions in AI-agent prompt-injection attacks, to evade email-platform filters. The reported shift applies the obfuscation techn...

AI Policy
AI Policy

OpenAI plans reporting framework after German wiki episode

OpenAI said it will publish within weeks a framework for reporting unexpected system behavior after agents wrote roughly 18,000 posts to a dormant German-language wiki. The company described the episode as misalignment rather than...