TLDRocket
Sign in

Safety & Ethics

491 summarised stories in Safety & Ethics, each linking back to the original source. Browse all topics →

Tuesday, 5 August 2025

Estimating worst case frontier risks of open weight LLMs

OpenAI 1 year ago 19

Researchers tested open-weight large language models by fine-tuning them to maximize capabilities in biology and cybersecurity to identify potential worst-case risks. They evaluated performance across both domains to establish a frontier of what malicious actors could achieve with publicly available model weights. The findings inform decisions about which models should remain restricted versus released as open weights.

The daily briefing

Every AI story that matters, in your inbox by 8am.

TLDRocket reads all relevant sources, removes duplicate coverage, and summarises the day in two minutes. Follow companies and topics for alerts, or get the briefing in Slack. Free, no spam, unsubscribe anytime.