Vision

530 Posts

Painting With Text, Voice, and Images: ChatGPT now accepts voice and image inputs and outputs.
Vision

Painting With Text, Voice, and Images: ChatGPT now accepts voice and image inputs and outputs.

The updates expand ChatGPT into a voice-controlled, interactive system for text and image interpretation and production.

September 27, 20232 min read
Painting With Text, Voice, and Images: ChatGPT now accepts voice and image inputs and outputs.
Vision

Painting With Text, Voice, and Images: ChatGPT now accepts voice and image inputs and outputs.

The updates expand ChatGPT into a voice-controlled, interactive system for text and image interpretation and production.

September 27, 20232 min read
Where Is Meta’s Generative Play?: Why Meta still lacks a flagship generative AI service
Vision

Where Is Meta’s Generative Play?: Why Meta still lacks a flagship generative AI service

While Microsoft and Google scramble to supercharge their businesses with text generation, Meta has yet to launch a flagship generative AI service. Reporters went looking for reasons why.

June 28, 20232 min read
Where Is Meta’s Generative Play?: Why Meta still lacks a flagship generative AI service
Vision

Where Is Meta’s Generative Play?: Why Meta still lacks a flagship generative AI service

While Microsoft and Google scramble to supercharge their businesses with text generation, Meta has yet to launch a flagship generative AI service. Reporters went looking for reasons why.

June 28, 20232 min read
Deep Learning at (Small) Scale: How to run PilotNet on a Raspberry Pi Pico microcontroller
Vision

Deep Learning at (Small) Scale: How to run PilotNet on a Raspberry Pi Pico microcontroller

TinyML shows promise for bringing deep learning to applications where electrical power is scarce, processing in the cloud is impractical, and/or data privacy is paramount.

May 24, 20232 min read
Deep Learning at (Small) Scale: How to run PilotNet on a Raspberry Pi Pico microcontroller
Vision

Deep Learning at (Small) Scale: How to run PilotNet on a Raspberry Pi Pico microcontroller

TinyML shows promise for bringing deep learning to applications where electrical power is scarce, processing in the cloud is impractical, and/or data privacy is paramount.

May 24, 20232 min read
Don’t Steal My Style: Glaze tool prevents AI from learning an artist's style.
Vision

Don’t Steal My Style: Glaze tool prevents AI from learning an artist's style.

Asked to produce “a landscape by Thomas Kinkade,” a text-to-image generator fine-tuned on the pastoral painter’s work can mimic his style in seconds, often for pennies. A new technique aims to make it harder for algorithms to mimic an artist’s style.

May 10, 20233 min read
Don’t Steal My Style: Glaze tool prevents AI from learning an artist's style.
Vision

Don’t Steal My Style: Glaze tool prevents AI from learning an artist's style.

Asked to produce “a landscape by Thomas Kinkade,” a text-to-image generator fine-tuned on the pastoral painter’s work can mimic his style in seconds, often for pennies. A new technique aims to make it harder for algorithms to mimic an artist’s style.

May 10, 20233 min read
Battlefield Chat: A military chatbot can create battle plans.
Vision

Battlefield Chat: A military chatbot can create battle plans.

Large language models may soon help military analysts and commanders make decisions on the battlefield.

May 10, 20232 min read
Battlefield Chat: A military chatbot can create battle plans.
Vision

Battlefield Chat: A military chatbot can create battle plans.

Large language models may soon help military analysts and commanders make decisions on the battlefield.

May 10, 20232 min read
Image Generators Copy Training Data: Spotting similarities between generated images and data
Vision

Image Generators Copy Training Data: Spotting similarities between generated images and data

We know that image generators create wonderful original works, but do they sometimes replicate their training data? Recent work found that replication does occur.

April 26, 20232 min read
Image Generators Copy Training Data: Spotting similarities between generated images and data
Vision

Image Generators Copy Training Data: Spotting similarities between generated images and data

We know that image generators create wonderful original works, but do they sometimes replicate their training data? Recent work found that replication does occur.

April 26, 20232 min read
Eyes on the Olympics: The 2024 Paris Olympics may have AI surveillance.
Vision

Eyes on the Olympics: The 2024 Paris Olympics may have AI surveillance.

French lawmakers said “oui” to broad uses of AI-powered surveillance. France’s National Assembly authorized authorities to test systems that detect unlawful, dangerous, or unusual behavior at next year’s Summer Olympics in Paris. The bill will become law unless the country’s top court blocks it.

April 19, 20231 min read
Eyes on the Olympics: The 2024 Paris Olympics may have AI surveillance.
Vision

Eyes on the Olympics: The 2024 Paris Olympics may have AI surveillance.

French lawmakers said “oui” to broad uses of AI-powered surveillance. France’s National Assembly authorized authorities to test systems that detect unlawful, dangerous, or unusual behavior at next year’s Summer Olympics in Paris. The bill will become law unless the country’s top court blocks it.

April 19, 20231 min read
Vision and Language Tightly Bound: Training on a single loss function improves multimiodal AI.
Vision

Vision and Language Tightly Bound: Training on a single loss function improves multimiodal AI.

Recent multimodal models process both text and images as sequences of tokens, but they learn to represent these distinct data types using separate loss functions. Recent work unifies the loss function as well.

March 15, 20232 min read

Subscribe to The Batch

Stay updated with weekly AI News and Insights delivered to your inbox