We’ve trained a large-scale unsupervised language model which generates coherent paragraphs of text, achieves state-of-the-art performanc...
Our first cohort of OpenAI Fellows has concluded, with each Fellow going from a machine learning beginner to core OpenAI contributor in t...
We’ve discovered that the gradient noise scale, a simple statistical metric, predicts the parallelizability of neural network training on...
We’re releasing CoinRun, a training environment which provides a metric for an agent’s ability to transfer its experience to novel situat...
We’re releasing Spinning Up in Deep RL, an educational resource designed to let anyone learn to become a skilled practitioner in deep rei...
We’ve developed an energy-based model that can quickly learn to identify and generate instances of concepts, such as near, above, between...
We’ve developed Random Network Distillation (RND), a prediction-based method for encouraging reinforcement learning agents to explore the...
We’re proposing an AI safety technique called iterated amplification that lets us specify complicated behaviors and goals that are beyond...
We are now accepting applications for our second cohort of OpenAI Scholars, a program where we provide 6–10 stipends and mentorship to in...
We are now accepting applications for OpenAI Fellows and Interns for 2019.
Our first cohort of OpenAI Scholars has now completed the program.
OpenAI Five lost two games against top Dota 2 players at The International in Vancouver this week, maintaining a good chance of winning f...
Yesterday, OpenAI Five won a best-of-three against a team of 99.95th percentile Dota players: Blitz, Cap, Fogged, Merlini, and MoonMeande...
We’ve trained a human-like robot hand to manipulate physical objects with unprecedented dexterity.
Our first class of OpenAI Scholars is underway, and you can now follow along as this group of experienced software developers becomes mac...
The OpenAI Five Benchmark match is now over!
We introduce Glow, a reversible generative model which uses invertible 1x1 convolutions. It extends previous work on reversible generativ...
We’ve trained an agent to achieve a high score of 74,500 on Montezuma’s Revenge from a single human demonstration, better than any previo...
Our team of five neural networks, OpenAI Five, has started to defeat amateur human teams at Dota 2.