
A Transformer Alternative Emerges: Mamba, a new approach that may outperform transformers
An architectural innovation improves upon transformers — up to 2 billion parameters, at least...

An architectural innovation improves upon transformers — up to 2 billion parameters, at least...

Large language models sometimes generate false statements. New work makes them more likely to produce factual output.

Large language models sometimes generate false statements. New work makes them more likely to produce factual output.

Humanoid robots can play football (known as soccer in the United States) in the real world, thanks to reinforcement learning.

Humanoid robots can play football (known as soccer in the United States) in the real world, thanks to reinforcement learning.

Research aims to help users select large language models that minimize expenses while maintaining quality.

Research aims to help users select large language models that minimize expenses while maintaining quality.

Robots equipped with large language models are asking their human overseers for help.

Robots equipped with large language models are asking their human overseers for help.

The technique known as reinforcement learning from human feedback fine-tunes large language models to be helpful and avoid generating harmful responses such as suggesting illegal or dangerous activities. An alternative method streamlines this approach and achieves better results.

The technique known as reinforcement learning from human feedback fine-tunes large language models to be helpful and avoid generating harmful responses such as suggesting illegal or dangerous activities. An alternative method streamlines this approach and achieves better results.

Machine learning models typically learn language by training on tasks like predicting the next word in a given text. Researchers trained a language model in a less focused, more human-like way.

Machine learning models typically learn language by training on tasks like predicting the next word in a given text. Researchers trained a language model in a less focused, more human-like way.

Large language models are not good at math. Researchers devised a way to make them better. Tiedong Liu and Bryan Kian Hsiang Low at the National University of Singapore proposed a method to fine-tune large language models for arithmetic tasks.

Large language models are not good at math. Researchers devised a way to make them better. Tiedong Liu and Bryan Kian Hsiang Low at the National University of Singapore proposed a method to fine-tune large language models for arithmetic tasks.
Stay updated with weekly AI News and Insights delivered to your inbox