TLDRocket
Sign in

Advanced Reasoning

52 summarised stories about Advanced Reasoning, each linking back to the original source. Browse all topics →

Wednesday, 16 August 2017

More on Dota 2

OpenAI Blog 8 years ago

A machine learning system trained through self-play improved from matching high-ranked Dota 2 players to defeating top professionals in approximately one month. The system achieved superhuman performance through continuous self-improvement, where the training data quality automatically increased as the agent's skill level rose. Self-play training eliminates the performance ceiling imposed by supervised learning approaches, which are limited to the quality of their static training datasets.

The daily briefing

Every AI story that matters, in your inbox by 8am.

TLDRocket reads 60+ sources, removes duplicate coverage, and summarises the day in two minutes. Free, no spam, unsubscribe anytime.