Loading timeline…
Loading timeline…
Every recorded change for this company, newest first. Back to the entity page →
49 events
Events per month · 12 months shown
OpenAI: Variance reduction for policy gradient with action-dependent factorized baselines
OpenAI: Multi-Goal Reinforcement Learning: Challenging robotics environments and request for research
Showing up to 400 events. Narrow by year or category, or use the changes feed for cursor-paged history. Times are UTC.
OpenAI: Plan online, learn offline: Efficient learning and exploration via model-based control
OpenAI: FFJORD: Free-form continuous dynamics for scalable reversible generative models
OpenAI: Some considerations on learning to explore via meta-reinforcement learning