Transformer Variants Head to Head: A benchmark for comparing different AI transformers.
Benchmarks

Transformer Variants Head to Head: A benchmark for comparing different AI transformers.

The transformer architecture has inspired a plethora of variations. Yet researchers have used a patchwork of metrics to evaluate their performance, making them hard to compare. New work aims to level the playing field.

March 3, 20212 min read
Computation as a National Resource: An effort to estimate computing capacity for 37 nations.
Benchmarks

Computation as a National Resource: An effort to estimate computing capacity for 37 nations.

How much processing power do various nations have on hand to drive their AI strategy? An international trade group aims to find out. The Organisation for Economic Co-operation and Development (OECD) is launching an effort to measure the computing capacity available in countries around the world.

February 17, 20211 min read
Computation as a National Resource: An effort to estimate computing capacity for 37 nations.
Benchmarks

Computation as a National Resource: An effort to estimate computing capacity for 37 nations.

How much processing power do various nations have on hand to drive their AI strategy? An international trade group aims to find out. The Organisation for Economic Co-operation and Development (OECD) is launching an effort to measure the computing capacity available in countries around the world.

February 17, 20211 min read
Prosperity of the Commons: Tools from MLCommons for improved model development
Benchmarks

Prosperity of the Commons: Tools from MLCommons for improved model development

A new consortium of companies, schools, and research labs is building open tools for next-generation machine learning. MLCommons aims to foster innovation in machine learning by developing new benchmarks, datasets, and best practices.

December 9, 20201 min read
Prosperity of the Commons: Tools from MLCommons for improved model development
Benchmarks

Prosperity of the Commons: Tools from MLCommons for improved model development

A new consortium of companies, schools, and research labs is building open tools for next-generation machine learning. MLCommons aims to foster innovation in machine learning by developing new benchmarks, datasets, and best practices.

December 9, 20201 min read
Dynamic Benchmarks: A platform for fooling language models
Benchmarks

Dynamic Benchmarks: A platform for fooling language models

Benchmarks provide a scientific basis for evaluating model performance, but they don’t necessarily map well to human cognitive abilities. Facebook aims to close the gap through a dynamic benchmarking method that keeps humans in the loop.

October 14, 20202 min read
Dynamic Benchmarks: A platform for fooling language models
Benchmarks

Dynamic Benchmarks: A platform for fooling language models

Benchmarks provide a scientific basis for evaluating model performance, but they don’t necessarily map well to human cognitive abilities. Facebook aims to close the gap through a dynamic benchmarking method that keeps humans in the loop.

October 14, 20202 min read
Do Muppets Have Common Sense?: The Bert NLP model scores high on common sense test.
Benchmarks

Do Muppets Have Common Sense?: The Bert NLP model scores high on common sense test.

Two years after it pointed a new direction for language models, Bert still hovers near the top of several natural language processing leaderboards. A new study considers whether Bert simply excels at tracking word order or or learns something closer to common sense.

September 16, 20202 min read
Do Muppets Have Common Sense?: The Bert NLP model scores high on common sense test.
Benchmarks

Do Muppets Have Common Sense?: The Bert NLP model scores high on common sense test.

Two years after it pointed a new direction for language models, Bert still hovers near the top of several natural language processing leaderboards. A new study considers whether Bert simply excels at tracking word order or or learns something closer to common sense.

September 16, 20202 min read
Optimizer Shootout: An evaluation of 14 deep learning optimizers
Benchmarks

Optimizer Shootout: An evaluation of 14 deep learning optimizers

Everyone has a favorite optimization method, but it’s not always clear which one works best in a given situation. New research aims to establish a set of benchmarks. Researchers evaluated 14 popular optimizers using the Deep Optimization Benchmark Suite some of them introduced last year.

September 9, 20202 min read
Optimizer Shootout: An evaluation of 14 deep learning optimizers
Benchmarks

Optimizer Shootout: An evaluation of 14 deep learning optimizers

Everyone has a favorite optimization method, but it’s not always clear which one works best in a given situation. New research aims to establish a set of benchmarks. Researchers evaluated 14 popular optimizers using the Deep Optimization Benchmark Suite some of them introduced last year.

September 9, 20202 min read
Built for Speed: Nvidia topped MLPerf's training benchmarks in 2020.
Benchmarks

Built for Speed: Nvidia topped MLPerf's training benchmarks in 2020.

Chips specially designed for AI are becoming much faster at training neural networks, judging from recent trials. MLPerf, an organization that’s developing standards for hardware performance in machine learning tasks, released results from its third benchmark competition.

August 5, 20201 min read
Built for Speed: Nvidia topped MLPerf's training benchmarks in 2020.
Benchmarks

Built for Speed: Nvidia topped MLPerf's training benchmarks in 2020.

Chips specially designed for AI are becoming much faster at training neural networks, judging from recent trials. MLPerf, an organization that’s developing standards for hardware performance in machine learning tasks, released results from its third benchmark competition.

August 5, 20201 min read
Running Fast, Standing Still: Some state of the art machine learning progress is illusory.
Benchmarks

Running Fast, Standing Still: Some state of the art machine learning progress is illusory.

Machine learning researchers report better and better results, but some of that progress may be illusory. Some models that appear to set a new state of the art haven’t been compared properly to their predecessors, Science News reports based on several published surveys.

June 10, 20201 min read
Running Fast, Standing Still: Some state of the art machine learning progress is illusory.
Benchmarks

Running Fast, Standing Still: Some state of the art machine learning progress is illusory.

Machine learning researchers report better and better results, but some of that progress may be illusory. Some models that appear to set a new state of the art haven’t been compared properly to their predecessors, Science News reports based on several published surveys.

June 10, 20201 min read

Subscribe to The Batch

Stay updated with weekly AI News and Insights delivered to your inbox