TLDRocket
Sign in

Safety & Ethics

499 summarised stories in Safety & Ethics, each linking back to the original source. Browse all topics →

Wednesday, 22 January 2025

Trading inference-time compute for adversarial robustness

OpenAI 1 year ago 34

OpenAI researchers showed that applying more computational resources during inference—specifically through multiple verification passes—improves a model's resilience against adversarial attacks. The technique increased robustness from 16% to 51% accuracy on adversarial examples by using five verification passes instead of one. This approach suggests a path toward more reliable AI systems without requiring retraining, though it introduces a latency-accuracy tradeoff.

The daily briefing

Every AI story that matters, in your inbox by 8am.

TLDRocket reads all relevant sources, removes duplicate coverage, and summarises the day in two minutes. Follow companies and topics for alerts, or get the briefing in Slack. Free, no spam, unsubscribe anytime.