Gopher

6 Posts

Right-Sizing Models for the Dataset: Finding the Best Data-To-Parameter Ratio for NLP Models
Gopher

Right-Sizing Models for the Dataset: Finding the Best Data-To-Parameter Ratio for NLP Models

The route to improving transformer-based language models like GPT-3 and Gopher, which are trained on immense quantities of text scraped from the web, has been to increase their size. But research shows that, given a processing budget, bigger doesn’t necessarily mean better.

November 9, 20222 min read
Right-Sizing Models for the Dataset: Finding the Best Data-To-Parameter Ratio for NLP Models
Gopher

Right-Sizing Models for the Dataset: Finding the Best Data-To-Parameter Ratio for NLP Models

The route to improving transformer-based language models like GPT-3 and Gopher, which are trained on immense quantities of text scraped from the web, has been to increase their size. But research shows that, given a processing budget, bigger doesn’t necessarily mean better.

November 9, 20222 min read
Yale Song: Foundation models for vision
Gopher

Yale Song: Foundation models for vision

Large models pretrained on immense quantities of text have been proven to provide strong foundations for solving specialized language tasks. My biggest hope for AI in 2022 is...

December 29, 20212 min read
Yale Song: Foundation models for vision
Gopher

Yale Song: Foundation models for vision

Large models pretrained on immense quantities of text have been proven to provide strong foundations for solving specialized language tasks. My biggest hope for AI in 2022 is...

December 29, 20212 min read
Large Language Models Shrink: Gopher and RETRO prove lean language models can push boundaries.
Gopher

Large Language Models Shrink: Gopher and RETRO prove lean language models can push boundaries.

DeepMind released three papers that push the boundaries — and examine the issues — of large language models.

December 15, 20212 min read
Large Language Models Shrink: Gopher and RETRO prove lean language models can push boundaries.
Gopher

Large Language Models Shrink: Gopher and RETRO prove lean language models can push boundaries.

DeepMind released three papers that push the boundaries — and examine the issues — of large language models.

December 15, 20212 min read

Subscribe to The Batch

Stay updated with weekly AI News and Insights delivered to your inbox