TLDRocket
Sign in

AI Alignment & Behavior

17 summarised stories about AI Alignment & Behavior, each linking back to the original source. Browse all topics →

+ Follow this topic

Wednesday, 11 January 2023

Forecasting potential misuses of language models for disinformation campaigns and how to reduce risk

OpenAI 3 years ago 48

OpenAI researchers partnered with Georgetown University and Stanford's Internet Observatory to study how language models could be weaponized for spreading disinformation. The project involved a workshop with 30 experts in October 2021 and produced a report analyzing threats and mitigation strategies. The findings provide a framework for identifying and reducing risks from language models being used to amplify false information at scale.

The daily briefing

Every AI story that matters, in your inbox by 8am.

TLDRocket reads all relevant sources, removes duplicate coverage, and summarises the day in two minutes. Follow companies and topics for alerts, or get the briefing in Slack. Free, no spam, unsubscribe anytime.