Synthetic Data Factory: AgentInstruct, a framework for generating diverse synthetic data for LLM fine-tuning
Machine Learning Research

Synthetic Data Factory: AgentInstruct, a framework for generating diverse synthetic data for LLM fine-tuning

Researchers increasingly fine-tune models on synthetic data, but generated datasets may not be sufficiently diverse. New work used agentic workflows to produce diverse synthetic datasets.

July 31, 20242 min read
Synthetic Data Factory: AgentInstruct, a framework for generating diverse synthetic data for LLM fine-tuning
Machine Learning Research

Synthetic Data Factory: AgentInstruct, a framework for generating diverse synthetic data for LLM fine-tuning

Researchers increasingly fine-tune models on synthetic data, but generated datasets may not be sufficiently diverse. New work used agentic workflows to produce diverse synthetic datasets.

July 31, 20242 min read
Expressive Synthetic Talking Heads: Microsoft's VASA-1 delivers more lifelike talking-head videos
Machine Learning Research

Expressive Synthetic Talking Heads: Microsoft's VASA-1 delivers more lifelike talking-head videos

Previous systems that produce a talking-head video from a photo and a spoken-word audio clip animate the lips and other parts of the face separately.

July 24, 20243 min read
Expressive Synthetic Talking Heads: Microsoft's VASA-1 delivers more lifelike talking-head videos
Machine Learning Research

Expressive Synthetic Talking Heads: Microsoft's VASA-1 delivers more lifelike talking-head videos

Previous systems that produce a talking-head video from a photo and a spoken-word audio clip animate the lips and other parts of the face separately.

July 24, 20243 min read
Hallucination Detector: Oxford scientists propose effective method to detect AI hallucinations
Machine Learning Research

Hallucination Detector: Oxford scientists propose effective method to detect AI hallucinations

Large language models can produce output that’s convincing but false. Researchers proposed a way to identify such hallucinations. 

July 17, 20243 min read
Hallucination Detector: Oxford scientists propose effective method to detect AI hallucinations
Machine Learning Research

Hallucination Detector: Oxford scientists propose effective method to detect AI hallucinations

Large language models can produce output that’s convincing but false. Researchers proposed a way to identify such hallucinations. 

July 17, 20243 min read
How Open Are Open Models?: Radboud University study ranks AI models on openness
Machine Learning Research

How Open Are Open Models?: Radboud University study ranks AI models on openness

The word “open” can mean many things with respect to AI. A new paper outlines the variations and ranks popular models for openness.

July 17, 20242 min read
How Open Are Open Models?: Radboud University study ranks AI models on openness
Machine Learning Research

How Open Are Open Models?: Radboud University study ranks AI models on openness

The word “open” can mean many things with respect to AI. A new paper outlines the variations and ranks popular models for openness.

July 17, 20242 min read
Efficient Subject Consistency For Stable Diffusion
Machine Learning Research

Efficient Subject Consistency For Stable Diffusion

Published in mid-2022, DreamBooth enables Stable Diffusion to depict variations on a given subject; say, a particular dog and the same dog with angel wings or wearing a chef’s hat.

July 16, 20243 min read
Efficient Subject Consistency For Stable Diffusion
Machine Learning Research

Efficient Subject Consistency For Stable Diffusion

Published in mid-2022, DreamBooth enables Stable Diffusion to depict variations on a given subject; say, a particular dog and the same dog with angel wings or wearing a chef’s hat.

July 16, 20243 min read
Like LoRA, But for Pretraining: GaLore, a memory-saving method for pretraining and fine-tuning LLMs
Machine Learning Research

Like LoRA, But for Pretraining: GaLore, a memory-saving method for pretraining and fine-tuning LLMs

Low-rank adaptation (LoRA) reduces memory requirements when fine-tuning large language models, but it isn’t as conducive to pretraining.

July 10, 20243 min read
Like LoRA, But for Pretraining: GaLore, a memory-saving method for pretraining and fine-tuning LLMs
Machine Learning Research

Like LoRA, But for Pretraining: GaLore, a memory-saving method for pretraining and fine-tuning LLMs

Low-rank adaptation (LoRA) reduces memory requirements when fine-tuning large language models, but it isn’t as conducive to pretraining.

July 10, 20243 min read
Model Merging Evolves: Researchers developed automated system for efficient model merging
Machine Learning Research

Model Merging Evolves: Researchers developed automated system for efficient model merging

The technique of model merging combines separate models into a single, more capable model without further training, but it requires expertise and manual effort. Researchers automated the process.

July 3, 20242 min read
Model Merging Evolves: Researchers developed automated system for efficient model merging
Machine Learning Research

Model Merging Evolves: Researchers developed automated system for efficient model merging

The technique of model merging combines separate models into a single, more capable model without further training, but it requires expertise and manual effort. Researchers automated the process.

July 3, 20242 min read
24 Hours on an Old Consumer GPU: Optimizing LLMs for low-resource hardware
Machine Learning Research

24 Hours on an Old Consumer GPU: Optimizing LLMs for low-resource hardware

BERT, a large language model released in 2018 and built upon the then-new transformer architecture, marked a paradigm shift in AI.

July 2, 20242 min read

Subscribe to The Batch

Stay updated with weekly AI News and Insights delivered to your inbox