Transformer-XL

4 Posts

Selective Attention: More efficient NLP training without sacrificing performance
Transformer-XL

Selective Attention: More efficient NLP training without sacrificing performance

Large transformer networks work wonders with natural language, but they require enormous amounts of computation. New research slashes processor cycles without compromising performance.

November 18, 20201 min read
Selective Attention: More efficient NLP training without sacrificing performance
Transformer-XL

Selective Attention: More efficient NLP training without sacrificing performance

Large transformer networks work wonders with natural language, but they require enormous amounts of computation. New research slashes processor cycles without compromising performance.

November 18, 20201 min read
What Language Models Know
Transformer-XL

What Language Models Know

Watson set a high bar for language understanding in 2011, when it famously whipped human competitors in the televised trivia game show Jeopardy! IBM’s special-purpose AI required around $1 billion. Research suggests that today’s best language models can accomplish similar tasks right off the shelf.

September 11, 20192 min read
What Language Models Know
Transformer-XL

What Language Models Know

Watson set a high bar for language understanding in 2011, when it famously whipped human competitors in the televised trivia game show Jeopardy! IBM’s special-purpose AI required around $1 billion. Research suggests that today’s best language models can accomplish similar tasks right off the shelf.

September 11, 20192 min read

Subscribe to The Batch

Stay updated with weekly AI News and Insights delivered to your inbox