TLDRocket
Sign in

AI Safety

107 summarised stories about AI Safety, each linking back to the original source. Browse all topics →

Wednesday, 24 August 2022

Our approach to alignment research

OpenAI Blog 3 years ago

Anthropic is developing techniques for AI systems to learn from human feedback and help humans evaluate AI behavior. The company aims to create an AI system well-aligned enough to assist in solving remaining alignment challenges. This approach assumes that a sufficiently aligned AI could accelerate progress on broader AI safety problems.

The daily briefing

Every AI story that matters, in your inbox by 8am.

TLDRocket reads 60+ sources, removes duplicate coverage, and summarises the day in two minutes. Free, no spam, unsubscribe anytime.