Solve RL With This One Weird Trick: How to get better performance from reinforcement learning.
Reinforcement Learning

Solve RL With This One Weird Trick: How to get better performance from reinforcement learning.

The previous state-of-the-art model for playing vintage Atari games took advantage of a number of advances in reinforcement learning (RL). The new champion is a basic RL architecture plus a trick borrowed from image generation.

August 18, 20212 min read
Bye Bye Bots: OpenAI quit robotics to focus on AGI.
Reinforcement Learning

Bye Bye Bots: OpenAI quit robotics to focus on AGI.

The independent research lab OpenAI wowed technology watchers in 2019 with a robotic hand that solved Rubik’s Cube. Now it has disbanded the team that built it. OpenAI cofounder Wojciech Zaremba revealed that OpenAI shuttered its robotics program last October.

July 21, 20211 min read
Bye Bye Bots: OpenAI quit robotics to focus on AGI.
Reinforcement Learning

Bye Bye Bots: OpenAI quit robotics to focus on AGI.

The independent research lab OpenAI wowed technology watchers in 2019 with a robotic hand that solved Rubik’s Cube. Now it has disbanded the team that built it. OpenAI cofounder Wojciech Zaremba revealed that OpenAI shuttered its robotics program last October.

July 21, 20211 min read
Walking the Dog: Training a robot to walk over unsteady terrain with RL.
Reinforcement Learning

Walking the Dog: Training a robot to walk over unsteady terrain with RL.

A reinforcement learning system enabled a four-legged robot to amble over unfamiliar, rapidly changing terrain.

July 14, 20211 min read
Walking the Dog: Training a robot to walk over unsteady terrain with RL.
Reinforcement Learning

Walking the Dog: Training a robot to walk over unsteady terrain with RL.

A reinforcement learning system enabled a four-legged robot to amble over unfamiliar, rapidly changing terrain.

July 14, 20211 min read
Behavioral Cloning Shootout: AI learns to play Counter Strike  Global Offensive.
Reinforcement Learning

Behavioral Cloning Shootout: AI learns to play Counter Strike Global Offensive.

Neural networks have learned to play video games like Dota 2 via reinforcement learning by playing for the equivalent of thousands of years (compressed into far less time). In new work, an automated player learned not by playing for millennia but by watching a few days’ worth of recorded gameplay.

July 7, 20212 min read
Behavioral Cloning Shootout: AI learns to play Counter Strike  Global Offensive.
Reinforcement Learning

Behavioral Cloning Shootout: AI learns to play Counter Strike Global Offensive.

Neural networks have learned to play video games like Dota 2 via reinforcement learning by playing for the equivalent of thousands of years (compressed into far less time). In new work, an automated player learned not by playing for millennia but by watching a few days’ worth of recorded gameplay.

July 7, 20212 min read
Computers Making Computers: How Google used AI to help design its TPU v4 chip.
Reinforcement Learning

Computers Making Computers: How Google used AI to help design its TPU v4 chip.

A neural network wrote the blueprint for upcoming computer chips that will accelerate deep learning itself. Google engineers used a reinforcement learning system to arrange the billions of minuscule transistors in an upcoming version of its Tensor Processing Unit (TPU) chips.

June 16, 20212 min read
Computers Making Computers: How Google used AI to help design its TPU v4 chip.
Reinforcement Learning

Computers Making Computers: How Google used AI to help design its TPU v4 chip.

A neural network wrote the blueprint for upcoming computer chips that will accelerate deep learning itself. Google engineers used a reinforcement learning system to arrange the billions of minuscule transistors in an upcoming version of its Tensor Processing Unit (TPU) chips.

June 16, 20212 min read
Medical AI Gets a Grip: An AI System controlled DaVinci surgical robots.
Reinforcement Learning

Medical AI Gets a Grip: An AI System controlled DaVinci surgical robots.

Surgical robots perform millions of delicate operations annually under human control. Now they’re getting ready to operate on their own.

May 19, 20212 min read
Medical AI Gets a Grip: An AI System controlled DaVinci surgical robots.
Reinforcement Learning

Medical AI Gets a Grip: An AI System controlled DaVinci surgical robots.

Surgical robots perform millions of delicate operations annually under human control. Now they’re getting ready to operate on their own.

May 19, 20212 min read
Drones For Defense: How companies like Anduril are developing military drones.
Reinforcement Learning

Drones For Defense: How companies like Anduril are developing military drones.

Drone startups are taking aim at military customers. As large tech companies have backed away from defense work, startups like Anduril, Shield AI, and Teal are picking up the slack. They’re developing autonomous fliers specifically for military operations.

March 10, 20212 min read
Drones For Defense: How companies like Anduril are developing military drones.
Reinforcement Learning

Drones For Defense: How companies like Anduril are developing military drones.

Drone startups are taking aim at military customers. As large tech companies have backed away from defense work, startups like Anduril, Shield AI, and Teal are picking up the slack. They’re developing autonomous fliers specifically for military operations.

March 10, 20212 min read
Performance Guaranteed: How deep learning networks can become Bayes-optimal.
Reinforcement Learning

Performance Guaranteed: How deep learning networks can become Bayes-optimal.

Bayes-optimal algorithms always make the best decisions given their training and input, if certain assumptions hold true. New work shows that some neural networks can approach this kind of performance.

February 3, 20212 min read
Performance Guaranteed: How deep learning networks can become Bayes-optimal.
Reinforcement Learning

Performance Guaranteed: How deep learning networks can become Bayes-optimal.

Bayes-optimal algorithms always make the best decisions given their training and input, if certain assumptions hold true. New work shows that some neural networks can approach this kind of performance.

February 3, 20212 min read

Subscribe to The Batch

Stay updated with weekly AI News and Insights delivered to your inbox