The Net Speaks in Many Tongues: NLP Model Translates 200 Different Languages
Machine Learning Research

The Net Speaks in Many Tongues: NLP Model Translates 200 Different Languages

Sentence pairs that have equivalent meanings in different languages — typically used to train machine translation systems — have been available in sufficient quantities for only around 100 languages. New work doubled that number and produced a more capable model.

October 19, 20224 min read
Long-Form Videos from Text Stories: Google's Phenaki Generates Long-Form Video from Text
Machine Learning Research

Long-Form Videos from Text Stories: Google's Phenaki Generates Long-Form Video from Text

Only a week ago, researchers unveiled a system that generates a few seconds of video based on a text prompt. New work enables a text-to-video system to produce an entire visual narrative from several sentences of text.

October 12, 20223 min read
Long-Form Videos from Text Stories: Google's Phenaki Generates Long-Form Video from Text
Machine Learning Research

Long-Form Videos from Text Stories: Google's Phenaki Generates Long-Form Video from Text

Only a week ago, researchers unveiled a system that generates a few seconds of video based on a text prompt. New work enables a text-to-video system to produce an entire visual narrative from several sentences of text.

October 12, 20223 min read
The Sound of Conversation: AI Learns to Mimic Conversational Pauses and Interruptions
Machine Learning Research

The Sound of Conversation: AI Learns to Mimic Conversational Pauses and Interruptions

In spoken conversation, people naturally take turns amid interjections and other patterns that aren’t strictly verbal. A new approach generated natural-sounding audio dialogs without training on text transcriptions that mark when one party should stop speaking and the other should chime in.

October 6, 20222 min read
The Sound of Conversation: AI Learns to Mimic Conversational Pauses and Interruptions
Machine Learning Research

The Sound of Conversation: AI Learns to Mimic Conversational Pauses and Interruptions

In spoken conversation, people naturally take turns amid interjections and other patterns that aren’t strictly verbal. A new approach generated natural-sounding audio dialogs without training on text transcriptions that mark when one party should stop speaking and the other should chime in.

October 6, 20222 min read
Text to Video Without Text-Video Training Data: Make-A-Video, an AI System from Meta, Generates Video from Text
Machine Learning Research

Text to Video Without Text-Video Training Data: Make-A-Video, an AI System from Meta, Generates Video from Text

Text-to-image generators like DALL·E 2, Midjourney, and Stable Diffusion are winning art contests and worrying artists. A new approach brings the magic of text-to-image generation to video.

October 5, 20222 min read
Text to Video Without Text-Video Training Data: Make-A-Video, an AI System from Meta, Generates Video from Text
Machine Learning Research

Text to Video Without Text-Video Training Data: Make-A-Video, an AI System from Meta, Generates Video from Text

Text-to-image generators like DALL·E 2, Midjourney, and Stable Diffusion are winning art contests and worrying artists. A new approach brings the magic of text-to-image generation to video.

October 5, 20222 min read
Cookbook for Vision Transformers: A Formula for Training Vision Transformers
Machine Learning Research

Cookbook for Vision Transformers: A Formula for Training Vision Transformers

Vision Transformers (ViTs) are overtaking convolutional neural networks (CNN) in many vision tasks, but procedures for training them are still tailored for CNNs. New research investigated how various training ingredients affect ViT performance.

September 28, 20222 min read
Cookbook for Vision Transformers: A Formula for Training Vision Transformers
Machine Learning Research

Cookbook for Vision Transformers: A Formula for Training Vision Transformers

Vision Transformers (ViTs) are overtaking convolutional neural networks (CNN) in many vision tasks, but procedures for training them are still tailored for CNNs. New research investigated how various training ingredients affect ViT performance.

September 28, 20222 min read
Automating Mattes for Visual Effects: New ML Method Produces Image Mattes Easier
Machine Learning Research

Automating Mattes for Visual Effects: New ML Method Produces Image Mattes Easier

Researchers at Baidu introduced PP-Matting, an architecture that, given an image, estimates the transparency of pixels surrounding foreground objects to create mattes without requiring additional input.

September 21, 20223 min read
Automating Mattes for Visual Effects: New ML Method Produces Image Mattes Easier
Machine Learning Research

Automating Mattes for Visual Effects: New ML Method Produces Image Mattes Easier

Researchers at Baidu introduced PP-Matting, an architecture that, given an image, estimates the transparency of pixels surrounding foreground objects to create mattes without requiring additional input.

September 21, 20223 min read
Update Any Language Model: New Method to Update Pretrained Language Models
Machine Learning Research

Update Any Language Model: New Method to Update Pretrained Language Models

The ability to update language models is essential to incorporate new information and correct undesirable behaviors. Previous methods are unwieldy and often fail as the amount of new data increases. New work offers a workaround.

September 14, 20223 min read
Update Any Language Model: New Method to Update Pretrained Language Models
Machine Learning Research

Update Any Language Model: New Method to Update Pretrained Language Models

The ability to update language models is essential to incorporate new information and correct undesirable behaviors. Previous methods are unwieldy and often fail as the amount of new data increases. New work offers a workaround.

September 14, 20223 min read
Attention to Rows and Columns: Altering Transformers' Self-Attention Mechanism for Greater Efficiency
Machine Learning Research

Attention to Rows and Columns: Altering Transformers' Self-Attention Mechanism for Greater Efficiency

A new approach alters transformers' self-attention mechanism to balance computational efficiency with performance on vision tasks.

September 7, 20222 min read
Attention to Rows and Columns: Altering Transformers' Self-Attention Mechanism for Greater Efficiency
Machine Learning Research

Attention to Rows and Columns: Altering Transformers' Self-Attention Mechanism for Greater Efficiency

A new approach alters transformers' self-attention mechanism to balance computational efficiency with performance on vision tasks.

September 7, 20222 min read

Subscribe to The Batch

Stay updated with weekly AI News and Insights delivered to your inbox