Segmented Images, No Labeled Data: Improved unsupervised learning for semantic segmentation
Transformer

Segmented Images, No Labeled Data: Improved unsupervised learning for semantic segmentation

Training a model to separate the objects in a picture typically requires labeled images for best results. Recent work upped the ante for training without labels.

January 11, 20232 min read
Precision-Guided Image Generation: Better text-to-image results with latent diffusion
Transformer

Precision-Guided Image Generation: Better text-to-image results with latent diffusion

Typical text-to-image generators can generate pictures of a cat, but not your cat. That’s because it’s hard to describe in a text prompt precisely all the things that distinguish your pet from other members of the same species.

January 4, 20232 min read
Precision-Guided Image Generation: Better text-to-image results with latent diffusion
Transformer

Precision-Guided Image Generation: Better text-to-image results with latent diffusion

Typical text-to-image generators can generate pictures of a cat, but not your cat. That’s because it’s hard to describe in a text prompt precisely all the things that distinguish your pet from other members of the same species.

January 4, 20232 min read
Language Models, Extended: Large language models grew more reliable and less biased in 2022.
Transformer

Language Models, Extended: Large language models grew more reliable and less biased in 2022.

Researchers pushed the boundaries of language models to address persistent problems of trustworthiness, bias, and updatability.

December 21, 20222 min read
Language Models, Extended: Large language models grew more reliable and less biased in 2022.
Transformer

Language Models, Extended: Large language models grew more reliable and less biased in 2022.

Researchers pushed the boundaries of language models to address persistent problems of trustworthiness, bias, and updatability.

December 21, 20222 min read
AI's Eyes Evolve: Vision transformer research exploded in 2022.
Transformer

AI's Eyes Evolve: Vision transformer research exploded in 2022.

Work on vision transformers exploded in 2022. Researchers published well over 17,000 ViT papers during the year. A major theme: combining self-attention and convolution.

December 21, 20222 min read
AI's Eyes Evolve: Vision transformer research exploded in 2022.
Transformer

AI's Eyes Evolve: Vision transformer research exploded in 2022.

Work on vision transformers exploded in 2022. Researchers published well over 17,000 ViT papers during the year. A major theme: combining self-attention and convolution.

December 21, 20222 min read
Memorize Less; Retrieve More: How small language models can perform specialized tasks.
Transformer

Memorize Less; Retrieve More: How small language models can perform specialized tasks.

Large language models are trained only to predict the next word based on previous ones. Yet, given a modest fine-tuning set, they acquire enough information to learn how to perform tasks such as answering questions.

December 14, 20223 min read
Memorize Less; Retrieve More: How small language models can perform specialized tasks.
Transformer

Memorize Less; Retrieve More: How small language models can perform specialized tasks.

Large language models are trained only to predict the next word based on previous ones. Yet, given a modest fine-tuning set, they acquire enough information to learn how to perform tasks such as answering questions.

December 14, 20223 min read
Seeing What Comes Next: Transformers predict future video frames.
Transformer

Seeing What Comes Next: Transformers predict future video frames.

If a robot can predict what it’s likely to see next, it may have a better basis for choosing an appropriate action — but it has to predict quickly. Transformers, for all their utility in computer vision, aren’t well suited to this because of their steep computational and memory requirements...

December 7, 20223 min read
Seeing What Comes Next: Transformers predict future video frames.
Transformer

Seeing What Comes Next: Transformers predict future video frames.

If a robot can predict what it’s likely to see next, it may have a better basis for choosing an appropriate action — but it has to predict quickly. Transformers, for all their utility in computer vision, aren’t well suited to this because of their steep computational and memory requirements...

December 7, 20223 min read
What the Missing Frames Showed: Machine Learning Describes Masked Video Events
Transformer

What the Missing Frames Showed: Machine Learning Describes Masked Video Events

Neural networks can describe in words what’s happening in pictures and videos — but can they make sensible guesses about things that happened before or will happen afterward? Researchers probed this ability.

November 16, 20223 min read
What the Missing Frames Showed: Machine Learning Describes Masked Video Events
Transformer

What the Missing Frames Showed: Machine Learning Describes Masked Video Events

Neural networks can describe in words what’s happening in pictures and videos — but can they make sensible guesses about things that happened before or will happen afterward? Researchers probed this ability.

November 16, 20223 min read
Right-Sizing Models for the Dataset: Finding the Best Data-To-Parameter Ratio for NLP Models
Transformer

Right-Sizing Models for the Dataset: Finding the Best Data-To-Parameter Ratio for NLP Models

The route to improving transformer-based language models like GPT-3 and Gopher, which are trained on immense quantities of text scraped from the web, has been to increase their size. But research shows that, given a processing budget, bigger doesn’t necessarily mean better.

November 9, 20222 min read
Right-Sizing Models for the Dataset: Finding the Best Data-To-Parameter Ratio for NLP Models
Transformer

Right-Sizing Models for the Dataset: Finding the Best Data-To-Parameter Ratio for NLP Models

The route to improving transformer-based language models like GPT-3 and Gopher, which are trained on immense quantities of text scraped from the web, has been to increase their size. But research shows that, given a processing budget, bigger doesn’t necessarily mean better.

November 9, 20222 min read

Subscribe to The Batch

Stay updated with weekly AI News and Insights delivered to your inbox