article

Reinforcement Learning Progress

by Sam Altman

Jun 25, 2018·~1 minutes·1 chapter

Chapters

1 total

Reinforcement Learning Progress

1:21

Description

Today, OpenAI released a new result. We used PPO (Proximal Policy Optimization), a general reinforcement learning algorithm invented by OpenAI, to train a team of 5 agents to play Dota and beat semi-pros.

We are not affiliated with Sam Altman. This unofficial audio edition is available to listen to and share for free. Read the original post on Sam Altman's blog.

Details

Duration

~1 minutes (1K characters)

Release date

Jun 25, 2018