Context Is Everything: Gemini 1.5 Pro, a leap in multimodal AI amid controversy over v1.0
Google

Context Is Everything: Gemini 1.5 Pro, a leap in multimodal AI amid controversy over v1.0

An update of Google’s flagship multimodal model keeps track of colossal inputs, while an earlier version generated some questionable outputs.

February 28, 20243 min read
GPT-4 Biothreat Risk is Low: Study finds GPT-4 no more risky than online search in aiding bioweapon development.
Google

GPT-4 Biothreat Risk is Low: Study finds GPT-4 no more risky than online search in aiding bioweapon development.

GPT-4 poses negligible additional risk that a malefactor could build a biological weapon, according to a new study. OpenAI compared the ability of GPT-4 and web search to contribute to the creation of a dangerous virus or bacterium. The large language model was barely more helpful than the web.

February 7, 20242 min read
GPT-4 Biothreat Risk is Low: Study finds GPT-4 no more risky than online search in aiding bioweapon development.
Google

GPT-4 Biothreat Risk is Low: Study finds GPT-4 no more risky than online search in aiding bioweapon development.

GPT-4 poses negligible additional risk that a malefactor could build a biological weapon, according to a new study. OpenAI compared the ability of GPT-4 and web search to contribute to the creation of a dangerous virus or bacterium. The large language model was barely more helpful than the web.

February 7, 20242 min read
More Consistent Generated Videos: Lumiere, a system that achieves unprecedented motion realism in video
Google

More Consistent Generated Videos: Lumiere, a system that achieves unprecedented motion realism in video

Text-to-video has struggled to produce consistent motions like walking and rotation. A new approach achieves more realistic motion.

January 31, 20242 min read
More Consistent Generated Videos: Lumiere, a system that achieves unprecedented motion realism in video
Google

More Consistent Generated Videos: Lumiere, a system that achieves unprecedented motion realism in video

Text-to-video has struggled to produce consistent motions like walking and rotation. A new approach achieves more realistic motion.

January 31, 20242 min read
Learning the Language of Geometry: AlphaGeometry, a system that nears expert proficiency in proving complex geometry theorems
Google

Learning the Language of Geometry: AlphaGeometry, a system that nears expert proficiency in proving complex geometry theorems

Machine learning algorithms often struggle with geometry. A language model learned to prove relatively difficult theorems. 

January 24, 20242 min read
Learning the Language of Geometry: AlphaGeometry, a system that nears expert proficiency in proving complex geometry theorems
Google

Learning the Language of Geometry: AlphaGeometry, a system that nears expert proficiency in proving complex geometry theorems

Machine learning algorithms often struggle with geometry. A language model learned to prove relatively difficult theorems. 

January 24, 20242 min read
Sing a Tune, Generate an Accompaniment: SingSong, a tool that generates instrumental music for unaccompanied input vocals
Google

Sing a Tune, Generate an Accompaniment: SingSong, a tool that generates instrumental music for unaccompanied input vocals

A neural network makes music for unaccompanied vocal tracks. Chris Donahue, Antoine Caillon, Adam Roberts, and colleagues at Google proposed SingSong, a system that generates musical accompaniments for sung melodies. You can listen to its output here.

January 17, 20242 min read
Sing a Tune, Generate an Accompaniment: SingSong, a tool that generates instrumental music for unaccompanied input vocals
Google

Sing a Tune, Generate an Accompaniment: SingSong, a tool that generates instrumental music for unaccompanied input vocals

A neural network makes music for unaccompanied vocal tracks. Chris Donahue, Antoine Caillon, Adam Roberts, and colleagues at Google proposed SingSong, a system that generates musical accompaniments for sung melodies. You can listen to its output here.

January 17, 20242 min read
AGI Defined: Researchers propose a taxonomy for artificial general intelligence (AGI).
Google

AGI Defined: Researchers propose a taxonomy for artificial general intelligence (AGI).

How will we know if someone succeeds in building artificial general intelligence (AGI)? A recent paper defines milestones on the road from calculator to superintelligence.

January 10, 20243 min read
AGI Defined: Researchers propose a taxonomy for artificial general intelligence (AGI).
Google

AGI Defined: Researchers propose a taxonomy for artificial general intelligence (AGI).

How will we know if someone succeeds in building artificial general intelligence (AGI)? A recent paper defines milestones on the road from calculator to superintelligence.

January 10, 20243 min read
Sharper Vision for Cancer: An AI-powered microscope that helps pathologists detect cancer
Google

Sharper Vision for Cancer: An AI-powered microscope that helps pathologists detect cancer

A microscope enhanced with augmented reality is helping pathologists recognize cancerous tissue. The United States Department of Defense is using microscopes that use machine learning models based on research from Google to detect cancers.

January 3, 20242 min read
Sharper Vision for Cancer: An AI-powered microscope that helps pathologists detect cancer
Google

Sharper Vision for Cancer: An AI-powered microscope that helps pathologists detect cancer

A microscope enhanced with augmented reality is helping pathologists recognize cancerous tissue. The United States Department of Defense is using microscopes that use machine learning models based on research from Google to detect cancers.

January 3, 20242 min read
Google’s Multimodal Challenger: All you need to know about Gemini, Google's new multimodal model
Google

Google’s Multimodal Challenger: All you need to know about Gemini, Google's new multimodal model

Google unveiled Gemini, its bid to catch up to, and perhaps surpass, OpenAI’s GPT-4. Google demonstrated the Gemini family of models that accept any combination of text (including code), images, video, and audio and output text and images. The demonstrations and metrics were impressive...

December 13, 20233 min read
Google’s Multimodal Challenger: All you need to know about Gemini, Google's new multimodal model
Google

Google’s Multimodal Challenger: All you need to know about Gemini, Google's new multimodal model

Google unveiled Gemini, its bid to catch up to, and perhaps surpass, OpenAI’s GPT-4. Google demonstrated the Gemini family of models that accept any combination of text (including code), images, video, and audio and output text and images. The demonstrations and metrics were impressive...

December 13, 20233 min read

Subscribe to The Batch

Stay updated with weekly AI News and Insights delivered to your inbox