
Better exploration with parameter noise
We’ve found that adding adaptive noise to the parameters of reinforcement learning algorithms frequently boosts…
Read at source →Source
Updates from OpenAI
Technology · 1167 articles · page 32 of 33 · fetched less than a minute ago

We’ve found that adding adaptive noise to the parameters of reinforcement learning algorithms frequently boosts…
Read at source →
We’re releasing a new class of reinforcement learning algorithms, Proximal Policy Optimization (PPO), which perform…
Read at source →
We’ve created images that reliably fool neural network classifiers when viewed from varied scales and perspectives.…
Read at source →
AdSee who you follow, newest first. Free to join. Posts that stay visible — not buried by an algo.
Join siguimi.com →


We’re open-sourcing a high-performance Python library for robotic simulation using the MuJoCo engine, developed over…
Read at source →
One step towards building safe AI systems is to remove the need for humans to write goal functions, since using a…
Read at source →
Multiagent environments where agents compete for resources are stepping stones on the path to AGI. Multiagent…
Read at source →

We’re open-sourcing OpenAI Baselines, our internal effort to reproduce reinforcement learning algorithms with…
Read at source →
We’ve created a robotics system, trained entirely in simulation and deployed on a physical robot, which can learn a new…
Read at source →
We are releasing Roboschool: open-source software for robot simulation, integrated with OpenAI Gym.
Read at source →


We’ve developed an unsupervised system which learns an excellent representation of sentiment, despite being trained…
Read at source →
We’ve created the world’s first Spam-detecting AI trained entirely in simulation and deployed on a physical robot.
Read at source →
We’ve discovered that evolution strategies (ES), an optimization technique that’s been known for decades, rivals the…
Read at source →

We’re excited to support today’s launch of Distill, a new kind of journal aimed at excellent communication of machine…
Read at source →
In this post we’ll outline new OpenAI research in which agents develop their own language.
Read at source →



Adversarial examples are inputs to machine learning models that an attacker has intentionally designed to cause the…
Read at source →

The OpenAI team is now 45 people. Together, we’re pushing the frontier of AI capabilities—whether by validating novel…
Read at source →

Reinforcement learning algorithms can break in surprising, counterintuitive ways. In this post we’ll explore one…
Read at source →
We’re releasing Universe, a software platform for measuring and training an AI’s general intelligence across the…
Read at source →

We’re working with Microsoft to start running most of our large-scale experiments on Azure.
Read at source →



