TLDRocket
Sign in

Safety & Ethics

512 summarised stories in Safety & Ethics, each linking back to the original source. Browse all topics →

Wednesday, 24 August 2022

Our approach to alignment research

OpenAI 3 years ago 40

Anthropic is developing techniques for AI systems to learn from human feedback and help humans evaluate AI behavior. The company aims to create an AI system well-aligned enough to assist in solving remaining alignment challenges. This approach assumes that a sufficiently aligned AI could accelerate progress on broader AI safety problems.

The daily briefing

Every AI story that matters, in your inbox by 8am.

TLDRocket reads all relevant sources, removes duplicate coverage, and summarises the day in two minutes. Follow companies and topics for alerts, or get the briefing in Slack. Free, no spam, unsubscribe anytime.