BL 6/17/2020

Image GPT

Image GPT: transformer trained on pixel sequences for image completion and generation with competitive unsupervised classification.

BL 12/3/2019

Procgen Benchmark

Procgen Benchmark: 16 procedurally-generated environments to measure reinforcement learning agent generalization and learning speed.

BL 11/21/2019

Safety Gym

Safety Gym: open-source suite of environments and tools for measuring reinforcement learning agent progress with safety constraints.

BL 11/5/2019

GPT-2: 1.5B release

Final GPT-2 1.5B model release with code/weights and staged release process documentation as test case for future powerful models.

BL 8/20/2019

GPT-2: 6-month follow-up

Release of 774M parameter GPT-2 language model with staged release strategy, open-source legal agreements, and research on misuse/benefit.

BL 12/14/2018

How AI training scales

Analysis showing gradient noise scale predicts neural network training parallelizability across diverse tasks.

BL 11/8/2018

Spinning Up in Deep RL

Educational resource with code examples, exercises, and tutorials for learning deep reinforcement learning.

BL 8/6/2018

OpenAI Five Benchmark: Results

OpenAI Five defeats 99.95th percentile Dota 2 players in best-of-three match with live audience and 100k viewers.

BL 7/30/2018

Learning dexterity

Robot hand trained to manipulate physical objects with unprecedented dexterity using machine learning methods.

BL 7/18/2018

OpenAI Five Benchmark

OpenAI Five benchmark match announcement with removed gameplay restrictions for competition against professional Dota players.

BL 6/22/2018

Retro Contest: Results

Results from Retro Contest showing top performers used tuning/extensions of PPO and Rainbow algorithms on Sonic benchmark.

BL 5/30/2018

OpenAI Fellows Fall 2018

OpenAI Fellows program accepting applications for 6-month AI research apprenticeship. Targets individuals without formal background in AI with mentorship and team placement.

BL 5/16/2018

AI and compute

Analysis showing AI compute in largest training runs increased 300,000x since 2012 with 3.4-month doubling time versus 2-year Moore's Law period.

BL 5/3/2018

AI safety via debate

AI safety technique training agents to debate topics with human judge. Proposes approach to align advanced AI systems with human preferences with proof-of-concept experiments.

BL 4/5/2018

Retro Contest

Transfer learning RL competition measuring generalization from previous experience on unseen video game levels. Uses Gym Retro platform with new benchmark.

BL 3/8/2018

On first-order meta-learning algorithms

Analysis of first-order meta-learning algorithms for learning parameter initializations. Generalizes first-order MAML using only first-order derivatives for meta-learning updates.

BL 3/7/2018

Reptile: A scalable meta-learning algorithm

Reptile: scalable meta-learning algorithm using repeated task sampling and SGD updates. Mathematically similar to first-order MAML requiring only black-box optimizer access.

BL 2/15/2018

Interpretable machine learning through teaching

Method for training AIs to teach each other using human-interpretable examples. Automatically selects informative examples to teach concepts effectively to both AI and human learners.