TLDRocket
Sign in

Quoting Thomas Ptacek

Simon Willison Simon Willison Covered by 14 sources

Security researcher Thomas Ptacek suggests that open-weights AI models from 2025 could perform sandbox escapes and network intrusions without requiring frontier models, implying OpenAI's sandboxes may not be particularly robust. The statement was made in the context of OpenAI's accidental cyberattack against Hugging Face. The observation implies that AI security threats may not be limited to the most advanced models, affecting defensive strategies across the industry.

Why it matters

I genuinely believe that if you took an open weights model from 2025 and built a pentest harness for it, it could do this kind of sandbox escape and scan/hack in most networks. This is only surprising because you assume OpenAI has sounder sandboxes. — Thomas Ptacek, doesn't think this even needs a frontier model Tags: thomas-ptacek, openai, security, generative-ai, ai-security-research, ai, llms, sandboxing

Also covered by

Related stories

The daily briefing

Every AI story that matters, in your inbox by 8am.

TLDRocket reads 60+ sources, removes duplicate coverage, and summarises the day in two minutes. Free, no spam, unsubscribe anytime.