Language Models Want to Be Free: How EleutherAI is developing a GPT-3 clone.
Datasets

Language Models Want to Be Free: How EleutherAI is developing a GPT-3 clone.

A grassroots research collective aims to make a GPT-3 clone that’s available to everyone. EleutherAI, a loose-knit group of independent researchers, is developing GPT-Neo, an open source, free-to-use version of OpenAI’s gargantuan language model.

February 3, 20211 min read
Language Models Want to Be Free: How EleutherAI is developing a GPT-3 clone.
Datasets

Language Models Want to Be Free: How EleutherAI is developing a GPT-3 clone.

A grassroots research collective aims to make a GPT-3 clone that’s available to everyone. EleutherAI, a loose-knit group of independent researchers, is developing GPT-Neo, an open source, free-to-use version of OpenAI’s gargantuan language model.

February 3, 20211 min read
Quake Watch: AI model detects earthquakes and estimates epicenters.
Datasets

Quake Watch: AI model detects earthquakes and estimates epicenters.

Detecting earthquakes is an important step toward warning surrounding communities that damaging seismic waves may be headed their way. A new model detects tremors and provides clues to their epicenter.

January 27, 20212 min read
Quake Watch: AI model detects earthquakes and estimates epicenters.
Datasets

Quake Watch: AI model detects earthquakes and estimates epicenters.

Detecting earthquakes is an important step toward warning surrounding communities that damaging seismic waves may be headed their way. A new model detects tremors and provides clues to their epicenter.

January 27, 20212 min read
U.S. New Year’s Resolutions for AI: All the AI programs authorized in the 2021 NDAA.
Datasets

U.S. New Year’s Resolutions for AI: All the AI programs authorized in the 2021 NDAA.

U.S. lawmakers authorized a slew of national programs that promote artificial intelligence research, development, and deployment, and support efforts to make sure the results are ethical and trustworthy.

January 6, 20212 min read
U.S. New Year’s Resolutions for AI: All the AI programs authorized in the 2021 NDAA.
Datasets

U.S. New Year’s Resolutions for AI: All the AI programs authorized in the 2021 NDAA.

U.S. lawmakers authorized a slew of national programs that promote artificial intelligence research, development, and deployment, and support efforts to make sure the results are ethical and trustworthy.

January 6, 20212 min read
Representing the Underrepresented: Many important AI datasets contain bias.
Datasets

Representing the Underrepresented: Many important AI datasets contain bias.

Some of deep learning’s bedrock datasets came under scrutiny as researchers combed them for built-in biases. Researchers found that popular datasets impart biases against socially marginalized groups to trained models due to the ways the datasets were compiled, labeled, and used.

December 23, 20202 min read
Representing the Underrepresented: Many important AI datasets contain bias.
Datasets

Representing the Underrepresented: Many important AI datasets contain bias.

Some of deep learning’s bedrock datasets came under scrutiny as researchers combed them for built-in biases. Researchers found that popular datasets impart biases against socially marginalized groups to trained models due to the ways the datasets were compiled, labeled, and used.

December 23, 20202 min read
Prosperity of the Commons: Tools from MLCommons for improved model development
Datasets

Prosperity of the Commons: Tools from MLCommons for improved model development

A new consortium of companies, schools, and research labs is building open tools for next-generation machine learning. MLCommons aims to foster innovation in machine learning by developing new benchmarks, datasets, and best practices.

December 9, 20201 min read
Prosperity of the Commons: Tools from MLCommons for improved model development
Datasets

Prosperity of the Commons: Tools from MLCommons for improved model development

A new consortium of companies, schools, and research labs is building open tools for next-generation machine learning. MLCommons aims to foster innovation in machine learning by developing new benchmarks, datasets, and best practices.

December 9, 20201 min read
Unsupervised Prejudice: Image classification models learned bias from ImageNet.
Datasets

Unsupervised Prejudice: Image classification models learned bias from ImageNet.

Social biases are well documented in decisions made by supervised models trained on ImageNet’s labels. But they also crept into the output of unsupervised models pretrained on the same dataset.

November 18, 20202 min read
Unsupervised Prejudice: Image classification models learned bias from ImageNet.
Datasets

Unsupervised Prejudice: Image classification models learned bias from ImageNet.

Social biases are well documented in decisions made by supervised models trained on ImageNet’s labels. But they also crept into the output of unsupervised models pretrained on the same dataset.

November 18, 20202 min read
Battling Bias in Synthetic Data: How synthetic data startups are working to avoid bias
Datasets

Battling Bias in Synthetic Data: How synthetic data startups are working to avoid bias

Synthetic datasets can inherit flaws in the real-world data they’re based on. Startups are working on solutions. Generating synthetic datasets for training machine learning systems is a booming business.

October 21, 20202 min read
Battling Bias in Synthetic Data: How synthetic data startups are working to avoid bias
Datasets

Battling Bias in Synthetic Data: How synthetic data startups are working to avoid bias

Synthetic datasets can inherit flaws in the real-world data they’re based on. Startups are working on solutions. Generating synthetic datasets for training machine learning systems is a booming business.

October 21, 20202 min read
GANs for Smaller Data: Training GANs on small data without overfitting
Datasets

GANs for Smaller Data: Training GANs on small data without overfitting

Trained on a small dataset, generative adversarial networks (GANs) tend to generate either replicas of the training data or noisy output. A new method spurs them to produce satisfying variations.

October 14, 20202 min read

Subscribe to The Batch

Stay updated with weekly AI News and Insights delivered to your inbox