Neural Networks Study Math: A sequence to sequence model for solving math problems.
Language

Neural Networks Study Math: A sequence to sequence model for solving math problems.

In tasks that involve generating natural language, neural networks often map an input sequence of words to an output sequence of words. Facebook researchers used a similar technique on sequences of mathematical symbols, training a model to map math problems to math solutions.

January 15, 20202 min read
Language Modeling on One GPU: Single-headed attention competes with transformers.
Language

Language Modeling on One GPU: Single-headed attention competes with transformers.

The latest large, pretrained language models rely on trendy layers based on transformer networks. New research shows that these newfangled layers may not be necessary.

January 8, 20202 min read
Language Modeling on One GPU: Single-headed attention competes with transformers.
Language

Language Modeling on One GPU: Single-headed attention competes with transformers.

The latest large, pretrained language models rely on trendy layers based on transformer networks. New research shows that these newfangled layers may not be necessary.

January 8, 20202 min read
Richard Socher — Boiling the Information Ocean: Using AI summarization to help with information overload
Language

Richard Socher — Boiling the Information Ocean: Using AI summarization to help with information overload

Ignorance is a choice in the Internet age. Virtually all of human knowledge is available for the cost of typing a few words into a search box.

January 1, 20202 min read
Richard Socher — Boiling the Information Ocean: Using AI summarization to help with information overload
Language

Richard Socher — Boiling the Information Ocean: Using AI summarization to help with information overload

Ignorance is a choice in the Internet age. Virtually all of human knowledge is available for the cost of typing a few words into a search box.

January 1, 20202 min read
Yann LeCun — Learning From Observation: The power of self-supervised learning
Language

Yann LeCun — Learning From Observation: The power of self-supervised learning

How is it that many people learn to drive a car fairly safely in 20 hours of practice, while current imitation learning algorithms take hundreds of thousands of hours, and reinforcement learning algorithms take millions of hours? Clearly we’re missing something big.

January 1, 20202 min read
Yann LeCun — Learning From Observation: The power of self-supervised learning
Language

Yann LeCun — Learning From Observation: The power of self-supervised learning

How is it that many people learn to drive a car fairly safely in 20 hours of practice, while current imitation learning algorithms take hundreds of thousands of hours, and reinforcement learning algorithms take millions of hours? Clearly we’re missing something big.

January 1, 20202 min read
Natural Language Processing Models Get Literate: Why 2019 was a breakthrough year for NLP
Language

Natural Language Processing Models Get Literate: Why 2019 was a breakthrough year for NLP

Earlier language models powered by Word2Vec and GloVe embeddings yielded confused chatbots, grammar tools with middle-school reading comprehension, and not-half-bad translations. The latest generation is so good, some people consider it dangerous.

December 24, 20192 min read
Natural Language Processing Models Get Literate: Why 2019 was a breakthrough year for NLP
Language

Natural Language Processing Models Get Literate: Why 2019 was a breakthrough year for NLP

Earlier language models powered by Word2Vec and GloVe embeddings yielded confused chatbots, grammar tools with middle-school reading comprehension, and not-half-bad translations. The latest generation is so good, some people consider it dangerous.

December 24, 20192 min read
Inside AI’s Muppet Empire: Why Are So Many NLP Models Named After Muppets?
Language

Inside AI’s Muppet Empire: Why Are So Many NLP Models Named After Muppets?

As language models show increasing power, a parallel trend has received less notice: The vogue for naming models after characters in the children’s TV show Sesame Street.

December 18, 20191 min read
Inside AI’s Muppet Empire: Why Are So Many NLP Models Named After Muppets?
Language

Inside AI’s Muppet Empire: Why Are So Many NLP Models Named After Muppets?

As language models show increasing power, a parallel trend has received less notice: The vogue for naming models after characters in the children’s TV show Sesame Street.

December 18, 20191 min read
Keeping the Facts Straight: NLP system FactCC fact checks texts.
Language

Keeping the Facts Straight: NLP system FactCC fact checks texts.

Automatically generated text summaries are becoming common in search engines and news websites. But existing summarizers often mix up facts. For instance, a victim’s name might get switched for the perpetrator’s.

December 11, 20192 min read
Keeping the Facts Straight: NLP system FactCC fact checks texts.
Language

Keeping the Facts Straight: NLP system FactCC fact checks texts.

Automatically generated text summaries are becoming common in search engines and news websites. But existing summarizers often mix up facts. For instance, a victim’s name might get switched for the perpetrator’s.

December 11, 20192 min read
Bigger Corpora, Better Answers: Using knowledge graphs to improve question answering NLP
Language

Bigger Corpora, Better Answers: Using knowledge graphs to improve question answering NLP

Models that summarize documents and answer questions work pretty well with limited source material, but they can slip into incoherence when they draw from a sizeable corpus. Recent work addresses this problem.

December 4, 20192 min read
Bigger Corpora, Better Answers: Using knowledge graphs to improve question answering NLP
Language

Bigger Corpora, Better Answers: Using knowledge graphs to improve question answering NLP

Models that summarize documents and answer questions work pretty well with limited source material, but they can slip into incoherence when they draw from a sizeable corpus. Recent work addresses this problem.

December 4, 20192 min read

Subscribe to The Batch

Stay updated with weekly AI News and Insights delivered to your inbox