The World's Fastest
AI News Feed
Breaking AI news from 100+ sources — updated every 6 hours, 24/7.
Page 191 Stories
Hindsight Experience Replay
Hindsight Experience Replay. OpenAI News - artificial intelligence news.
Teacher–student curriculum learning
Teacher–student curriculum learning. OpenAI News - artificial intelligence news.
Faster physics in Python
We’re open-sourcing a high-performance Python library for robotic simulation using the MuJoCo engine,…
Learning from human preferences
One step towards building safe AI systems is to remove the need for humans to write goal functions, since…
Learning to cooperate, compete, and communicate
Multiagent environments where agents compete for resources are stepping stones on the path to AGI. Multiagent…
UCB exploration via Q-ensembles
UCB exploration via Q-ensembles. OpenAI News - artificial intelligence news.
OpenAI Baselines: DQN
We’re open-sourcing OpenAI Baselines, our internal effort to reproduce reinforcement learning algorithms…
Robots that learn
We’ve created a robotics system, trained entirely in simulation and deployed on a physical robot, which can…
Roboschool
We are releasing Roboschool: open-source software for robot simulation, integrated with OpenAI Gym.
Equivalence between policy gradients and soft Q-learning
Equivalence between policy gradients and soft Q-learning. OpenAI News - artificial intelligence news.
Stochastic Neural Networks for hierarchical reinforcement learning
Stochastic Neural Networks for hierarchical reinforcement learning. OpenAI ChatGPT - artificial intelligence…
Unsupervised sentiment neuron
We’ve developed an unsupervised system which learns an excellent representation of sentiment, despite being…
Spam detection in the physical world
We’ve created the world’s first Spam-detecting AI trained entirely in simulation and deployed on a…
Evolution strategies as a scalable alternative to reinforcement learning
We’ve discovered that evolution strategies (ES), an optimization technique that’s been known for decades,…
One-shot imitation learning
One-shot imitation learning. OpenAI ChatGPT - artificial intelligence news.
Learning to communicate
Learning to communicate. OpenAI ChatGPT - artificial intelligence news.
Emergence of grounded compositional language in multi-agent populations
Emergence of grounded compositional language in multi-agent populations. OpenAI ChatGPT - artificial…
Prediction and control with temporal segment models
Prediction and control with temporal segment models. OpenAI ChatGPT - artificial intelligence news.
Third-person imitation learning
Third-person imitation learning. OpenAI ChatGPT - artificial intelligence news.
Attacking machine learning with adversarial examples
Adversarial examples are inputs to machine learning models that an attacker has intentionally designed to…
Adversarial attacks on neural network policies
Adversarial attacks on neural network policies. OpenAI ChatGPT - artificial intelligence news.
Team update
The OpenAI team is now 45 people. Together, we’re pushing the frontier of AI capabilities—whether by…
PixelCNN++: Improving the PixelCNN with discretized logistic mixture likelihood and other modifications
PixelCNN++: Improving the PixelCNN with discretized logistic mixture likelihood and other modifications.…
Faulty reward functions in the wild
Reinforcement learning algorithms can break in surprising, counterintuitive ways. In this post we’ll…