Vision Transformer

24 Posts

Synthetic Data Helps Image Classification: StableRep, a method that trains vision transformers on images generated by Stable Diffusion
Vision Transformer

Synthetic Data Helps Image Classification: StableRep, a method that trains vision transformers on images generated by Stable Diffusion

Generated images can be more effective than real ones in training a vision model to classify images. Yonglong Tian, Lijie Fan, and colleagues at Google and MIT introduced StableRep, a self-supervised method that trains vision transformers on images generated by...

November 8, 20232 min read
Synthetic Data Helps Image Classification: StableRep, a method that trains vision transformers on images generated by Stable Diffusion
Vision Transformer

Synthetic Data Helps Image Classification: StableRep, a method that trains vision transformers on images generated by Stable Diffusion

Generated images can be more effective than real ones in training a vision model to classify images. Yonglong Tian, Lijie Fan, and colleagues at Google and MIT introduced StableRep, a self-supervised method that trains vision transformers on images generated by...

November 8, 20232 min read
Masked Pretraining for CNNs: ConvNeXt V2, the new model family that boosts ConvNet performance
Vision Transformer

Masked Pretraining for CNNs: ConvNeXt V2, the new model family that boosts ConvNet performance

Vision transformers have bested convolutional neural networks (CNNs) in a number of key vision tasks. Have CNNs hit their limit? New research suggests otherwise.

September 13, 20232 min read
Masked Pretraining for CNNs: ConvNeXt V2, the new model family that boosts ConvNet performance
Vision Transformer

Masked Pretraining for CNNs: ConvNeXt V2, the new model family that boosts ConvNet performance

Vision transformers have bested convolutional neural networks (CNNs) in a number of key vision tasks. Have CNNs hit their limit? New research suggests otherwise.

September 13, 20232 min read
Vision Transformers Made Manageable: FlexiViT, the vision transformer that allows users to specify the patch size
Vision Transformer

Vision Transformers Made Manageable: FlexiViT, the vision transformer that allows users to specify the patch size

Vision transformers typically process images in patches of fixed size. Smaller patches yield higher accuracy but require more computation. A new training method lets AI engineers adjust the tradeoff.

August 23, 20232 min read
Vision Transformers Made Manageable: FlexiViT, the vision transformer that allows users to specify the patch size
Vision Transformer

Vision Transformers Made Manageable: FlexiViT, the vision transformer that allows users to specify the patch size

Vision transformers typically process images in patches of fixed size. Smaller patches yield higher accuracy but require more computation. A new training method lets AI engineers adjust the tradeoff.

August 23, 20232 min read
AI's Eyes Evolve: Vision transformer research exploded in 2022.
Vision Transformer

AI's Eyes Evolve: Vision transformer research exploded in 2022.

Work on vision transformers exploded in 2022. Researchers published well over 17,000 ViT papers during the year. A major theme: combining self-attention and convolution.

December 21, 20222 min read
AI's Eyes Evolve: Vision transformer research exploded in 2022.
Vision Transformer

AI's Eyes Evolve: Vision transformer research exploded in 2022.

Work on vision transformers exploded in 2022. Researchers published well over 17,000 ViT papers during the year. A major theme: combining self-attention and convolution.

December 21, 20222 min read
Cookbook for Vision Transformers: A Formula for Training Vision Transformers
Vision Transformer

Cookbook for Vision Transformers: A Formula for Training Vision Transformers

Vision Transformers (ViTs) are overtaking convolutional neural networks (CNN) in many vision tasks, but procedures for training them are still tailored for CNNs. New research investigated how various training ingredients affect ViT performance.

September 28, 20222 min read
Cookbook for Vision Transformers: A Formula for Training Vision Transformers
Vision Transformer

Cookbook for Vision Transformers: A Formula for Training Vision Transformers

Vision Transformers (ViTs) are overtaking convolutional neural networks (CNN) in many vision tasks, but procedures for training them are still tailored for CNNs. New research investigated how various training ingredients affect ViT performance.

September 28, 20222 min read
Attention to Rows and Columns: Altering Transformers' Self-Attention Mechanism for Greater Efficiency
Vision Transformer

Attention to Rows and Columns: Altering Transformers' Self-Attention Mechanism for Greater Efficiency

A new approach alters transformers' self-attention mechanism to balance computational efficiency with performance on vision tasks.

September 7, 20222 min read
Attention to Rows and Columns: Altering Transformers' Self-Attention Mechanism for Greater Efficiency
Vision Transformer

Attention to Rows and Columns: Altering Transformers' Self-Attention Mechanism for Greater Efficiency

A new approach alters transformers' self-attention mechanism to balance computational efficiency with performance on vision tasks.

September 7, 20222 min read
Object-Detection Transformers Simplified: New Research Improves Object Detection With Vision Transformers
Vision Transformer

Object-Detection Transformers Simplified: New Research Improves Object Detection With Vision Transformers

ViTDet, a new system from Facebook, adds an object detector to a plain pretrained transformer.

August 31, 20222 min read
Object-Detection Transformers Simplified: New Research Improves Object Detection With Vision Transformers
Vision Transformer

Object-Detection Transformers Simplified: New Research Improves Object Detection With Vision Transformers

ViTDet, a new system from Facebook, adds an object detector to a plain pretrained transformer.

August 31, 20222 min read
Cutting the Carbon Cost of Training: A New Tool Helps NLP Models Lower Their Gas Emissions
Vision Transformer

Cutting the Carbon Cost of Training: A New Tool Helps NLP Models Lower Their Gas Emissions

You can reduce your model’s carbon emissions by being choosy about when and where you train it.

July 20, 20222 min read

Subscribe to The Batch

Stay updated with weekly AI News and Insights delivered to your inbox