VentureBeat AI
·
4 days ago
● 4 sources
54% of enterprises have experienced AI agent security incidents or near-misses, with 69% allowing credential sharing among agents and only 30% isolating high-risk agents in sandboxes. Enterprises rely primarily on provider-native security controls from OpenAI, Google, and Microsoft rather than purpose-built agent security tools, with satisfaction averaging 4.2 out of 5. Despite high satisfaction with current controls, organizations with credential sharing face incident rates 23 percentage points higher than those with per-agent scoped identities, driving most enterprises to plan tooling changes within the year.
The New Stack
·
4 days ago
● 3 sources
OpenAI unveiled GPT-Red, an automated red-teaming system that uses AI to find prompt injection vulnerabilities in AI agents by testing thousands of exploit variations. GPT-5.6 achieved six times fewer failures on prompt-injection benchmarks than the strongest production model released four months earlier, and GPT-Red successfully manipulated a live vending machine agent to discount items over $100 to $0.50. The result shifts security testing from manual human discovery to continuous automated adversarial probing integrated into model training pipelines.
The Neuron
·
5 days ago
A developer created a CAPTCHA that required Claude 5 to spend 10 minutes and 100K tokens to solve it. The attack cost 100K tokens and took 10 minutes of processing time. This approach demonstrates a potential defense mechanism against automated AI systems by making unauthorized access economically and temporally expensive.