TLDRocket
Sign in

Reinforcement Learning

64 summarised stories about Reinforcement Learning, each linking back to the original source. Browse all topics →

Friday, 18 August 2017

OpenAI Baselines: ACKTR & A2C

OpenAI Blog 8 years ago

OpenAI released two new reinforcement learning algorithm implementations: A2C, a synchronous version of A3C that performs equally well, and ACKTR, which improves sample efficiency over both TRPO and A2C. ACKTR requires only marginally more computation per update than A2C while achieving better sample efficiency. These implementations are now available in OpenAI Baselines, giving researchers easier access to these algorithms for training reinforcement learning agents.

The daily briefing

Every AI story that matters, in your inbox by 8am.

TLDRocket reads 60+ sources, removes duplicate coverage, and summarises the day in two minutes. Free, no spam, unsubscribe anytime.