Latent Diffusion

8 Posts

Generated Video Gets Real(er): OpenAI’s Sora, a new player in text-to-video generation
Latent Diffusion

Generated Video Gets Real(er): OpenAI’s Sora, a new player in text-to-video generation

OpenAI’s new video generator raises the bar for detail and realism in generated videos — but the company released few details about how it built the system.

February 22, 20243 min read
Generated Video Gets Real(er): OpenAI’s Sora, a new player in text-to-video generation
Latent Diffusion

Generated Video Gets Real(er): OpenAI’s Sora, a new player in text-to-video generation

OpenAI’s new video generator raises the bar for detail and realism in generated videos — but the company released few details about how it built the system.

February 22, 20243 min read
Better Images, Less Training: Würstchen, a speedy, high-quality image generator
Latent Diffusion

Better Images, Less Training: Würstchen, a speedy, high-quality image generator

The longer text-to-image models train, the better their output — but the training is costly. Researchers built a system that produced superior images after far less training.

February 14, 20243 min read
Better Images, Less Training: Würstchen, a speedy, high-quality image generator
Latent Diffusion

Better Images, Less Training: Würstchen, a speedy, high-quality image generator

The longer text-to-image models train, the better their output — but the training is costly. Researchers built a system that produced superior images after far less training.

February 14, 20243 min read
Music Generation For the Masses: Stability.ai launches Stable Audio, a text-to-music generator.
Latent Diffusion

Music Generation For the Masses: Stability.ai launches Stable Audio, a text-to-music generator.

Stability.ai, maker of the Stable Diffusion image generator and StableLM text generator, launched Stable Audio, a system that generates music and sound effects from text. You can play with it and listen to examples here. The service is free for 20 generations per month up to 45 seconds long.

September 21, 20232 min read
Music Generation For the Masses: Stability.ai launches Stable Audio, a text-to-music generator.
Latent Diffusion

Music Generation For the Masses: Stability.ai launches Stable Audio, a text-to-music generator.

Stability.ai, maker of the Stable Diffusion image generator and StableLM text generator, launched Stable Audio, a system that generates music and sound effects from text. You can play with it and listen to examples here. The service is free for 20 generations per month up to 45 seconds long.

September 21, 20232 min read
What the Brain Sees: How a text-to-image model generates images from brain scans
Latent Diffusion

What the Brain Sees: How a text-to-image model generates images from brain scans

A pretrained text-to-image generator enabled researchers to see — roughly — what other people looked at based on brain scans. Yu Takagi and Shinji Nishimoto developed a method that uses Stable Diffusion to reconstruct images viewed by test subjects...

June 22, 20232 min read
What the Brain Sees: How a text-to-image model generates images from brain scans
Latent Diffusion

What the Brain Sees: How a text-to-image model generates images from brain scans

A pretrained text-to-image generator enabled researchers to see — roughly — what other people looked at based on brain scans. Yu Takagi and Shinji Nishimoto developed a method that uses Stable Diffusion to reconstruct images viewed by test subjects...

June 22, 20232 min read

Subscribe to The Batch

Stay updated with weekly AI News and Insights delivered to your inbox