overfeed.news

Variance reduction for policy gradient with action-dependent factorized baselines

OpenAI News

8y

Age

Published
Collected
Read the full article at openai.com

overfeed.news indexes and links. We publish a short excerpt — the full article stays at OpenAI News.

More from OpenAI News

Log in to follow this source
Variance reduction for policy gradient with action-dependent…