Variance reduction for policy gradient with action-dependent factorized baselines
IgnoreOpenAI Blog · 2018-03-20 07:00 UTC
Not analyzed yet
Eligible for automatic cleanup in 1 day(s) unless marked Must Read.
Content
No content snippet available in the feed.