Gemini Thinks Faster: Google’s Gemini 2.0 Flash Thinking advances in reasoning, outperforms DeepSeek-R1
Machine Learning Research

Gemini Thinks Faster: Google’s Gemini 2.0 Flash Thinking advances in reasoning, outperforms DeepSeek-R1

Google updated the December-vintage reasoning model Gemini 2.0 Flash Thinking and other Flash models, gaining ground on OpenAI o1 and DeepSeek-R1.

February 5, 20252 min read
Training for Computer Use: UI-TARS shows strong computer use capabilities in benchmarks
Machine Learning Research

Training for Computer Use: UI-TARS shows strong computer use capabilities in benchmarks

As Anthropic, Google, OpenAI, and others roll out agents that are capable of computer use, new work shows how underlying models can be trained to do this.

February 5, 20253 min read
Training for Computer Use: UI-TARS shows strong computer use capabilities in benchmarks
Machine Learning Research

Training for Computer Use: UI-TARS shows strong computer use capabilities in benchmarks

As Anthropic, Google, OpenAI, and others roll out agents that are capable of computer use, new work shows how underlying models can be trained to do this.

February 5, 20253 min read
Reasoning in High Gear: o3-mini, a faster, more affordable reasoning model for coding, math, and science
Machine Learning Research

Reasoning in High Gear: o3-mini, a faster, more affordable reasoning model for coding, math, and science

OpenAI introduced a successor to its o1 models that’s faster, less expensive, and especially strong in coding, math, and science.

February 5, 20253 min read
Reasoning in High Gear: o3-mini, a faster, more affordable reasoning model for coding, math, and science
Machine Learning Research

Reasoning in High Gear: o3-mini, a faster, more affordable reasoning model for coding, math, and science

OpenAI introduced a successor to its o1 models that’s faster, less expensive, and especially strong in coding, math, and science.

February 5, 20253 min read
Fine-Tuning Fine Points: Active inheritance, a smarter way to fine-tune models on synthetic data
Machine Learning Research

Fine-Tuning Fine Points: Active inheritance, a smarter way to fine-tune models on synthetic data

The practice of fine-tuning models on synthetic data is becoming well established. But synthetic training data, even if it represents the training task well, may include characteristics like toxicity that impart unwelcome properties in the trained model’s output...

January 29, 20253 min read
Fine-Tuning Fine Points: Active inheritance, a smarter way to fine-tune models on synthetic data
Machine Learning Research

Fine-Tuning Fine Points: Active inheritance, a smarter way to fine-tune models on synthetic data

The practice of fine-tuning models on synthetic data is becoming well established. But synthetic training data, even if it represents the training task well, may include characteristics like toxicity that impart unwelcome properties in the trained model’s output...

January 29, 20253 min read
Computer Use Gains Momentum: OpenAI’s Operator automates online tasks with a new AI agent
Machine Learning Research

Computer Use Gains Momentum: OpenAI’s Operator automates online tasks with a new AI agent

OpenAI introduced an AI agent that performs simple web tasks on a user’s behalf.

January 29, 20252 min read
Computer Use Gains Momentum: OpenAI’s Operator automates online tasks with a new AI agent
Machine Learning Research

Computer Use Gains Momentum: OpenAI’s Operator automates online tasks with a new AI agent

OpenAI introduced an AI agent that performs simple web tasks on a user’s behalf.

January 29, 20252 min read
Reinforcement Learning Heats Up: How DeepSeek-R1 and Kimi k1.5 use reinforcement learning to improve reasoning
Machine Learning Research

Reinforcement Learning Heats Up: How DeepSeek-R1 and Kimi k1.5 use reinforcement learning to improve reasoning

Reinforcement learning is emerging as an avenue for building large language models with advanced reasoning capabilities.

January 29, 20252 min read
Reinforcement Learning Heats Up: How DeepSeek-R1 and Kimi k1.5 use reinforcement learning to improve reasoning
Machine Learning Research

Reinforcement Learning Heats Up: How DeepSeek-R1 and Kimi k1.5 use reinforcement learning to improve reasoning

Reinforcement learning is emerging as an avenue for building large language models with advanced reasoning capabilities.

January 29, 20252 min read
Generated Chip Designs Work in Mysterious Ways: Researchers used deep learning and an evolutionary algorithm to design chips in minutes
Machine Learning Research

Generated Chip Designs Work in Mysterious Ways: Researchers used deep learning and an evolutionary algorithm to design chips in minutes

Designing integrated circuits typically requires years of human expertise. Recent work set AI to the task with surprising results.

January 22, 20252 min read
Generated Chip Designs Work in Mysterious Ways: Researchers used deep learning and an evolutionary algorithm to design chips in minutes
Machine Learning Research

Generated Chip Designs Work in Mysterious Ways: Researchers used deep learning and an evolutionary algorithm to design chips in minutes

Designing integrated circuits typically requires years of human expertise. Recent work set AI to the task with surprising results.

January 22, 20252 min read
DeepSeek Sharpens Its Reasoning: DeepSeek-R1, an affordable rival to OpenAI’s o1
Machine Learning Research

DeepSeek Sharpens Its Reasoning: DeepSeek-R1, an affordable rival to OpenAI’s o1

A new open model rivals OpenAI’s o1, and it’s free to use or modify.

January 22, 20254 min read
DeepSeek Sharpens Its Reasoning: DeepSeek-R1, an affordable rival to OpenAI’s o1
Machine Learning Research

DeepSeek Sharpens Its Reasoning: DeepSeek-R1, an affordable rival to OpenAI’s o1

A new open model rivals OpenAI’s o1, and it’s free to use or modify.

January 22, 20254 min read

Subscribe to The Batch

Stay updated with weekly AI News and Insights delivered to your inbox