Open Video Gen Closes the Gap: Tencent releases HunyuanVideo, an open source model rivaling commercial video generators
Machine Learning Research

Open Video Gen Closes the Gap: Tencent releases HunyuanVideo, an open source model rivaling commercial video generators

The gap is narrowing between closed and open models for video generation.

December 18, 20242 min read
Open Video Gen Closes the Gap: Tencent releases HunyuanVideo, an open source model rivaling commercial video generators
Machine Learning Research

Open Video Gen Closes the Gap: Tencent releases HunyuanVideo, an open source model rivaling commercial video generators

The gap is narrowing between closed and open models for video generation.

December 18, 20242 min read
Phi-4 Beats Models Five Times Its Size: Microsoft’s Phi-4 blends synthetic and organic data to surpass larger models in math and reasoning benchmarks
Machine Learning Research

Phi-4 Beats Models Five Times Its Size: Microsoft’s Phi-4 blends synthetic and organic data to surpass larger models in math and reasoning benchmarks

Microsoft updated its smallest model family with a single, surprisingly high-performance model.

December 18, 20243 min read
Phi-4 Beats Models Five Times Its Size: Microsoft’s Phi-4 learned from a blend of synthetic and organic data to surpass larger models in math and reasoning benchmarks
Machine Learning Research

Phi-4 Beats Models Five Times Its Size: Microsoft’s Phi-4 learned from a blend of synthetic and organic data to surpass larger models in math and reasoning benchmarks

Microsoft updated its smallest model family with a single, surprisingly high-performance model.

December 18, 20243 min read
Getting the Facts Right: A memory method that reduces hallucinations in LLMs
Machine Learning Research

Getting the Facts Right: A memory method that reduces hallucinations in LLMs

Large language models that remember more hallucinate less.

December 11, 20242 min read
Getting the Facts Right: A memory method that reduces hallucinations in LLMs
Machine Learning Research

Getting the Facts Right: A memory method that reduces hallucinations in LLMs

Large language models that remember more hallucinate less.

December 11, 20242 min read
Game Worlds on Tap: Genie 2 brings interactive 3D worlds to life
Machine Learning Research

Game Worlds on Tap: Genie 2 brings interactive 3D worlds to life

A new model improves on recent progress in generating interactive virtual worlds from still images.

December 11, 20242 min read
Game Worlds on Tap: Genie 2 brings interactive 3D worlds to life
Machine Learning Research

Game Worlds on Tap: Genie 2 brings interactive 3D worlds to life

A new model improves on recent progress in generating interactive virtual worlds from still images.

December 11, 20242 min read
Higher Reasoning: OpenAI debuts o1 and pro mode for $200/month
Machine Learning Research

Higher Reasoning: OpenAI debuts o1 and pro mode for $200/month

OpenAI launched not only its highly anticipated o1 model but also an operating mode that enables the model to deliver higher performance — at a hefty price.

December 11, 20243 min read
Higher Reasoning: OpenAI debuts o1 and pro mode for $200/month
Machine Learning Research

Higher Reasoning: OpenAI debuts o1 and pro mode for $200/month

OpenAI launched not only its highly anticipated o1 model but also an operating mode that enables the model to deliver higher performance — at a hefty price.

December 11, 20243 min read
Breaking Jailbreaks: New E-DPO method strengthens defenses against jailbreak prompts
Machine Learning Research

Breaking Jailbreaks: New E-DPO method strengthens defenses against jailbreak prompts

Jailbreak prompts can prod a large language model (LLM) to overstep built-in boundaries, leading it to do things like respond to queries it was trained to refuse to answer. Researchers devised a way to further boost the probability that LLMs will respond in ways that respect such limits.

December 4, 20242 min read
Breaking Jailbreaks: New E-DPO method strengthens defenses against jailbreak prompts
Machine Learning Research

Breaking Jailbreaks: New E-DPO method strengthens defenses against jailbreak prompts

Jailbreak prompts can prod a large language model (LLM) to overstep built-in boundaries, leading it to do things like respond to queries it was trained to refuse to answer. Researchers devised a way to further boost the probability that LLMs will respond in ways that respect such limits.

December 4, 20242 min read
Mistral’s Vision-Language Contender: Mistral unveils Pixtral Large, a rival to top vision-language models
Machine Learning Research

Mistral’s Vision-Language Contender: Mistral unveils Pixtral Large, a rival to top vision-language models

Mistral AI unveiled Pixtral Large, which rivals top models at processing combinations of text and images.

December 4, 20242 min read
Mistral’s Vision-Language Contender: Mistral unveils Pixtral Large, a rival to top vision-language models
Machine Learning Research

Mistral’s Vision-Language Contender: Mistral unveils Pixtral Large, a rival to top vision-language models

Mistral AI unveiled Pixtral Large, which rivals top models at processing combinations of text and images.

December 4, 20242 min read
Object Detection for Small Devices: Grounding DINO 1.5, an edge device model built for faster, smarter object detection
Machine Learning Research

Object Detection for Small Devices: Grounding DINO 1.5, an edge device model built for faster, smarter object detection

An open source model is designed to perform sophisticated object detection on edge devices like phones, cars, medical equipment, and smart doorbells.

November 27, 20243 min read

Subscribe to The Batch

Stay updated with weekly AI News and Insights delivered to your inbox