Google Releases Open Source LLMs: All we know about Google's Gemma-7B and Gemma-2B models
Machine Learning Research

Google Releases Open Source LLMs: All we know about Google's Gemma-7B and Gemma-2B models

Google asserted its open source bona fides with new models. Google released weights for Gemma-7B, an 8.5 billion-parameter large language model intended to run GPUs, and Gemma-2B, a 2.5 billion-parameter version intended for deployment on CPUs and edge devices.

March 6, 20242 min read
Human Feedback Without Reinforcement Learning: Direct Preference Optimization (DPO) fine-tunes pretrained large language models on human preferences without the cumbersome step of reinforcement learning.
Machine Learning Research

Human Feedback Without Reinforcement Learning: Direct Preference Optimization (DPO) fine-tunes pretrained large language models on human preferences without the cumbersome step of reinforcement learning.

Reinforcement learning from human feedback (RLHF) is widely used to fine-tune pretrained models to deliver outputs that align with human preferences. New work aligns pretrained models without the cumbersome step of reinforcement learning.

February 29, 20242 min read
Human Feedback Without Reinforcement Learning: Direct Preference Optimization (DPO) fine-tunes pretrained large language models on human preferences without the cumbersome step of reinforcement learning.
Machine Learning Research

Human Feedback Without Reinforcement Learning: Direct Preference Optimization (DPO) fine-tunes pretrained large language models on human preferences without the cumbersome step of reinforcement learning.

Reinforcement learning from human feedback (RLHF) is widely used to fine-tune pretrained models to deliver outputs that align with human preferences. New work aligns pretrained models without the cumbersome step of reinforcement learning.

February 29, 20242 min read
Swiss Army LLM
Machine Learning Research

Swiss Army LLM

The combination of  language models that are equipped for retrieval augmented generation can retrieve text from a database to improve their output. Further work extends this capability to retrieve information from any application that comes with an API. 

February 28, 20243 min read
Swiss Army LLM
Machine Learning Research

Swiss Army LLM

The combination of  language models that are equipped for retrieval augmented generation can retrieve text from a database to improve their output. Further work extends this capability to retrieve information from any application that comes with an API. 

February 28, 20243 min read
Better, Faster Network Pruning: Researchers devise pruning method that boosts AI speed
Machine Learning Research

Better, Faster Network Pruning: Researchers devise pruning method that boosts AI speed

Pruning weights from a neural network makes it smaller and faster, but it can take a lot of computation to choose weights that can be removed without degrading the network’s performance.

February 28, 20242 min read
Better, Faster Network Pruning: Researchers devise pruning method that boosts AI speed
Machine Learning Research

Better, Faster Network Pruning: Researchers devise pruning method that boosts AI speed

Pruning weights from a neural network makes it smaller and faster, but it can take a lot of computation to choose weights that can be removed without degrading the network’s performance.

February 28, 20242 min read
Memory-Efficient Optimizer: A method to reduce memory needs when fine-tuning AI models
Machine Learning Research

Memory-Efficient Optimizer: A method to reduce memory needs when fine-tuning AI models

Researchers devised a way to reduce memory requirements when fine-tuning large language models. Kai Lv and colleagues at Fudan University proposed low memory optimization (LOMO), a modification of stochastic gradient descent that stores less data than other optimizers during fine-tuning.

February 22, 20242 min read
Memory-Efficient Optimizer: A method to reduce memory needs when fine-tuning AI models
Machine Learning Research

Memory-Efficient Optimizer: A method to reduce memory needs when fine-tuning AI models

Researchers devised a way to reduce memory requirements when fine-tuning large language models. Kai Lv and colleagues at Fudan University proposed low memory optimization (LOMO), a modification of stochastic gradient descent that stores less data than other optimizers during fine-tuning.

February 22, 20242 min read
Better Images, Less Training: Würstchen, a speedy, high-quality image generator
Machine Learning Research

Better Images, Less Training: Würstchen, a speedy, high-quality image generator

The longer text-to-image models train, the better their output — but the training is costly. Researchers built a system that produced superior images after far less training.

February 14, 20243 min read
Better Images, Less Training: Würstchen, a speedy, high-quality image generator
Machine Learning Research

Better Images, Less Training: Würstchen, a speedy, high-quality image generator

The longer text-to-image models train, the better their output — but the training is costly. Researchers built a system that produced superior images after far less training.

February 14, 20243 min read
LLMs Can Get Inside Your Head: AI models show promise in understanding human beliefs, research reveals
Machine Learning Research

LLMs Can Get Inside Your Head: AI models show promise in understanding human beliefs, research reveals

Most people understand that others’ mental states can differ from their own. For instance, if your friend leaves a smartphone on a table and you privately put it in your pocket, you understand that your friend continues to believe it was on the table.

February 7, 20242 min read
LLMs Can Get Inside Your Head: AI models show promise in understanding human beliefs, research reveals
Machine Learning Research

LLMs Can Get Inside Your Head: AI models show promise in understanding human beliefs, research reveals

Most people understand that others’ mental states can differ from their own. For instance, if your friend leaves a smartphone on a table and you privately put it in your pocket, you understand that your friend continues to believe it was on the table.

February 7, 20242 min read
More Consistent Generated Videos: Lumiere, a system that achieves unprecedented motion realism in video
Machine Learning Research

More Consistent Generated Videos: Lumiere, a system that achieves unprecedented motion realism in video

Text-to-video has struggled to produce consistent motions like walking and rotation. A new approach achieves more realistic motion.

January 31, 20242 min read
More Consistent Generated Videos: Lumiere, a system that achieves unprecedented motion realism in video
Machine Learning Research

More Consistent Generated Videos: Lumiere, a system that achieves unprecedented motion realism in video

Text-to-video has struggled to produce consistent motions like walking and rotation. A new approach achieves more realistic motion.

January 31, 20242 min read

Subscribe to The Batch

Stay updated with weekly AI News and Insights delivered to your inbox