Anima Anandkumar — The Power of Simulation: How simulation can be useful for supervised learning
Simulation

Anima Anandkumar — The Power of Simulation: How simulation can be useful for supervised learning

We’ve had great success with supervised deep learning on labeled data. Now it’s time to explore other ways to learn: training on unlabeled data, lifelong learning, and especially letting models explore a simulated environment before transferring what they learn to the real world.

January 1, 20202 min read
Anima Anandkumar — The Power of Simulation: How simulation can be useful for supervised learning
Simulation

Anima Anandkumar — The Power of Simulation: How simulation can be useful for supervised learning

We’ve had great success with supervised deep learning on labeled data. Now it’s time to explore other ways to learn: training on unlabeled data, lifelong learning, and especially letting models explore a simulated environment before transferring what they learn to the real world.

January 1, 20202 min read
Simulation Substitutes for Data: When simulation works wonders with deep learning
Simulation

Simulation Substitutes for Data: When simulation works wonders with deep learning

The future of machine learning may depend less on amassing ground-truth data than simulating the environment in which a model will operate. Deep learning works like magic with enough high-quality data. When examples are scarce, though, researchers are using simulation to fill the gap.

December 24, 20191 min read
Simulation Substitutes for Data: When simulation works wonders with deep learning
Simulation

Simulation Substitutes for Data: When simulation works wonders with deep learning

The future of machine learning may depend less on amassing ground-truth data than simulating the environment in which a model will operate. Deep learning works like magic with enough high-quality data. When examples are scarce, though, researchers are using simulation to fill the gap.

December 24, 20191 min read
Different Skills From Different Demos: Implicit reinforcement without interaction at scale, explained
Simulation

Different Skills From Different Demos: Implicit reinforcement without interaction at scale, explained

Reinforcement learning trains models by trial and error. In batch reinforcement learning (BRL), models learn by observing many demonstrations by a variety of actors. But what if one doctor is handier with a scalpel while another excels at suturing?

December 18, 20192 min read
Different Skills From Different Demos: Implicit reinforcement without interaction at scale, explained
Simulation

Different Skills From Different Demos: Implicit reinforcement without interaction at scale, explained

Reinforcement learning trains models by trial and error. In batch reinforcement learning (BRL), models learn by observing many demonstrations by a variety of actors. But what if one doctor is handier with a scalpel while another excels at suturing?

December 18, 20192 min read
Seeing the World Blindfolded: The observational dropout technique, explained
Simulation

Seeing the World Blindfolded: The observational dropout technique, explained

In reinforcement learning, if researchers want an agent to have an internal representation of its environment, they’ll build and train a world model that it can refer to. New research shows that world models can emerge from standard training, rather than needing to be built separately.

December 11, 20192 min read
Seeing the World Blindfolded: The observational dropout technique, explained
Simulation

Seeing the World Blindfolded: The observational dropout technique, explained

In reinforcement learning, if researchers want an agent to have an internal representation of its environment, they’ll build and train a world model that it can refer to. New research shows that world models can emerge from standard training, rather than needing to be built separately.

December 11, 20192 min read
Take That, Humans!
Simulation

Take That, Humans!

At the BlizzCon gaming convention last weekend, players of the strategy game StarCraft II stood in line to get walloped by DeepMind’s AI. After training for the better part of a year, the bot has become one of the world’s top players.

November 6, 20192 min read
Take That, Humans!
Simulation

Take That, Humans!

At the BlizzCon gaming convention last weekend, players of the strategy game StarCraft II stood in line to get walloped by DeepMind’s AI. After training for the better part of a year, the bot has become one of the world’s top players.

November 6, 20192 min read
New Materials Courtesy of Bayes
Simulation

New Materials Courtesy of Bayes

Would you like an umbrella that fits in your pocket? Researchers used machine learning to invent sturdy but collapsible materials that might lead to such a fantastical object.

October 23, 20191 min read
Cube Controversy
Simulation

Cube Controversy

OpenAI trained a five-fingered robotic hand to unscramble the Rubik’s Cube puzzle, bringing both acclaim and criticism. The AI research lab OpenAI trained a mechanical hand to balance, twist, and turn the cube.

October 23, 20192 min read
Cube Controversy
Simulation

Cube Controversy

OpenAI trained a five-fingered robotic hand to unscramble the Rubik’s Cube puzzle, bringing both acclaim and criticism. The AI research lab OpenAI trained a mechanical hand to balance, twist, and turn the cube.

October 23, 20192 min read
New Materials Courtesy of Bayes
Simulation

New Materials Courtesy of Bayes

Would you like an umbrella that fits in your pocket? Researchers used machine learning to invent sturdy but collapsible materials that might lead to such a fantastical object.

October 23, 20191 min read
Autonomous Drones Ready to Race
Simulation

Autonomous Drones Ready to Race

Pilots in drone races fly souped-up quadcopters around an obstacle course at 120 miles per hour. But soon they may be out of a job, as race organizers try to spice things up with drones controlled by AI.

October 16, 20191 min read

Subscribe to The Batch

Stay updated with weekly AI News and Insights delivered to your inbox