It’s a Small World Model After All: More efficient world models for reinforcement learning
Reinforcement Learning

It’s a Small World Model After All: More efficient world models for reinforcement learning

World models, which learn a compressed representation of a dynamic environment like, say, a video game, have delivered top results in reinforcement learning. A new method makes them much smaller.

January 13, 20212 min read
It’s a Small World Model After All: More efficient world models for reinforcement learning
Reinforcement Learning

It’s a Small World Model After All: More efficient world models for reinforcement learning

World models, which learn a compressed representation of a dynamic environment like, say, a video game, have delivered top results in reinforcement learning. A new method makes them much smaller.

January 13, 20212 min read
Ilya Sutskever: OpenAI’s co-founder on building multimodal AI models
Reinforcement Learning

Ilya Sutskever: OpenAI’s co-founder on building multimodal AI models

The past year was the first in which general-purpose models became economically useful. GPT-3, in particular, demonstrated that large language models have surprising linguistic competence and the ability to perform a wide variety of useful tasks.

December 30, 20202 min read
Ilya Sutskever: OpenAI’s co-founder on building multimodal AI models
Reinforcement Learning

Ilya Sutskever: OpenAI’s co-founder on building multimodal AI models

The past year was the first in which general-purpose models became economically useful. GPT-3, in particular, demonstrated that large language models have surprising linguistic competence and the ability to perform a wide variety of useful tasks.

December 30, 20202 min read
How to Drive a Balloon: How high-altitude balloons navigate using AI.
Reinforcement Learning

How to Drive a Balloon: How high-altitude balloons navigate using AI.

Helium balloons that beam internet service to hard-to-serve areas are using AI to navigate amid high-altitude winds. Loon, the Alphabet division that provides wireless internet via polyethylene blimps.

December 9, 20202 min read
How to Drive a Balloon: How high-altitude balloons navigate using AI.
Reinforcement Learning

How to Drive a Balloon: How high-altitude balloons navigate using AI.

Helium balloons that beam internet service to hard-to-serve areas are using AI to navigate amid high-altitude winds. Loon, the Alphabet division that provides wireless internet via polyethylene blimps.

December 9, 20202 min read
Phantom Menace: Fighter pilot trains against augmented reality jet.
Reinforcement Learning

Phantom Menace: Fighter pilot trains against augmented reality jet.

A fighter pilot battled a true-to-life virtual enemy in midair. In the skies over southern California, an airman pitted his dogfighting skills against an AI-controlled opponent that was projected onto his augmented-reality visor.

December 2, 20202 min read
Phantom Menace: Fighter pilot trains against augmented reality jet.
Reinforcement Learning

Phantom Menace: Fighter pilot trains against augmented reality jet.

A fighter pilot battled a true-to-life virtual enemy in midair. In the skies over southern California, an airman pitted his dogfighting skills against an AI-controlled opponent that was projected onto his augmented-reality visor.

December 2, 20202 min read
RL Agents SOS!: Inside Agence, a reinforcement learning video game.
Reinforcement Learning

RL Agents SOS!: Inside Agence, a reinforcement learning video game.

A new multimedia experience lets audience members help artificially intelligent creatures work together to survive. Agence, an interactive virtual reality (VR) project blends audience participation with reinforcement learning to create an experience that’s half film, half video game.

October 21, 20202 min read
RL Agents SOS!: Inside Agence, a reinforcement learning video game.
Reinforcement Learning

RL Agents SOS!: Inside Agence, a reinforcement learning video game.

A new multimedia experience lets audience members help artificially intelligent creatures work together to survive. Agence, an interactive virtual reality (VR) project blends audience participation with reinforcement learning to create an experience that’s half film, half video game.

October 21, 20202 min read
Guess What Happens Next: Research teaches robots to predict unseen obstacles.
Reinforcement Learning

Guess What Happens Next: Research teaches robots to predict unseen obstacles.

New research teaches robots to anticipate what’s coming rather than focusing on what’s right in front of them. Researchers developed Occupancy Anticipation (OA), a navigation system that predicts unseen obstacles in addition to observing those in its field of view.

September 23, 20202 min read
Guess What Happens Next: Research teaches robots to predict unseen obstacles.
Reinforcement Learning

Guess What Happens Next: Research teaches robots to predict unseen obstacles.

New research teaches robots to anticipate what’s coming rather than focusing on what’s right in front of them. Researchers developed Occupancy Anticipation (OA), a navigation system that predicts unseen obstacles in addition to observing those in its field of view.

September 23, 20202 min read
Chess: The Next Move: Chess masters use AI to test variants of the game.
Reinforcement Learning

Chess: The Next Move: Chess masters use AI to test variants of the game.

AI has humbled human chess masters. Now it’s helping them take the game to the next level. DeepMind and retired chess champion Vladimir Kramnik trained AlphaZero, a reinforcement learning model that bested human experts in chess, Go, and Shogi, to play-test changes in the rules.

September 16, 20201 min read
Chess: The Next Move: Chess masters use AI to test variants of the game.
Reinforcement Learning

Chess: The Next Move: Chess masters use AI to test variants of the game.

AI has humbled human chess masters. Now it’s helping them take the game to the next level. DeepMind and retired chess champion Vladimir Kramnik trained AlphaZero, a reinforcement learning model that bested human experts in chess, Go, and Shogi, to play-test changes in the rules.

September 16, 20201 min read
Experience Counts: Research proposes an upgrade to experience replay.
Reinforcement Learning

Experience Counts: Research proposes an upgrade to experience replay.

If the world changes every second and you take a picture every 10 seconds, you won’t have enough pictures to observe the changes clearly, and storing a series of pictures won’t help. On the other hand, if you take a picture every tenth of a second, then storing a history will help model the world.

August 26, 20202 min read

Subscribe to The Batch

Stay updated with weekly AI News and Insights delivered to your inbox