Better Spatial Perception for Robots: MolmoAct creates spatial maps for robots to plot their actions before executing text directions
Machine Learning Research

Better Spatial Perception for Robots: MolmoAct creates spatial maps for robots to plot their actions before executing text directions

Robot control systems that accept only text input struggle to translate words into motions in space. Researchers developed a system that enables robots to plan spatial paths before they execute text instructions.

October 15, 20253 min read
Fine-Tuning Simplified: Thinking Machines’ new Tinker API makes it easier to fine-tune models on many GPUs
Machine Learning Research

Fine-Tuning Simplified: Thinking Machines’ new Tinker API makes it easier to fine-tune models on many GPUs

The first offering from Thinking Machines Lab, the startup founded by former OpenAI CTO Mira Murati, aims to simplify — and democratize — the process of fine-tuning AI models.

October 15, 20252 min read
DeepSeek Cuts Inference Costs: DeepSeek-V3.2-Exp streamlines processing using a "lightning indexer," boosting efficiency
Machine Learning Research

DeepSeek Cuts Inference Costs: DeepSeek-V3.2-Exp streamlines processing using a "lightning indexer," boosting efficiency

DeepSeek’s latest large language model can cut inference costs by more than half and processes long contexts dramatically faster relative to its predecessor.

October 15, 20253 min read
LoRA Adapters On Tap: Text-to-LoRA generates task-specific LoRA adapters directly from natural language descriptions
Machine Learning Research

LoRA Adapters On Tap: Text-to-LoRA generates task-specific LoRA adapters directly from natural language descriptions

The approach known as LoRA streamlines fine-tuning by training a small adapter that modifies a pretrained model’s weights at inference. Researchers built a model that generates such adapters directly.

October 8, 20252 min read
Qwen3 Goes Big (and Smaller): Alibaba expands Qwen3 family with a 1 trillion-parameter Max model, open-weights Qwen3-VL, and the Qwen3-Omni voice model
Machine Learning Research

Qwen3 Goes Big (and Smaller): Alibaba expands Qwen3 family with a 1 trillion-parameter Max model, open-weights Qwen3-VL, and the Qwen3-Omni voice model

Alibaba rounded out the Qwen3 family with its biggest large language model to date as well as smaller models that process text, images, video, and/or audio.

October 8, 20253 min read
Claude Levels Up: Anthropic launches Claude Sonnet 4.5 and the Claude Agent SDK, and overhauls Claude Code for developers
Machine Learning Research

Claude Levels Up: Anthropic launches Claude Sonnet 4.5 and the Claude Agent SDK, and overhauls Claude Code for developers

Anthropic updated its mid-size Claude Sonnet model, making it the first member of the Claude family to reach version 4.5. It also enhanced the Claude Code agentic coding tool with long-desired features.

October 8, 20253 min read
Earth Modeled in 10-Meter Squares: Google’s AlphaEarth Foundations tracks the whole planet’s climate, land use, potential for disasters, in detail and at scale
Machine Learning Research

Earth Modeled in 10-Meter Squares: Google’s AlphaEarth Foundations tracks the whole planet’s climate, land use, potential for disasters, in detail and at scale

Researchers built a model that integrates satellite imagery and other sensor readings across the entire surface of the Earth to reveal patterns of climate, land use, and other features.

October 1, 20253 min read
AI Generates Viral Genomes: Researchers use genomic language models to create custom viruses
Machine Learning Research

AI Generates Viral Genomes: Researchers use genomic language models to create custom viruses

Researchers used AI models to create novel viruses from scratch.

October 1, 20253 min read
Faster Reinforcement Learning: New technique auto-selects training examples to speed up fine-tuning
Machine Learning Research

Faster Reinforcement Learning: New technique auto-selects training examples to speed up fine-tuning

Fine-tuning large language models via reinforcement learning is computationally expensive, but researchers found a way to streamline the process.

September 24, 20252 min read
What ChatGPT Users Want: ChatGPT users now more likely to be young, female, and seeking info, study shows
Machine Learning Research

What ChatGPT Users Want: ChatGPT users now more likely to be young, female, and seeking info, study shows

What do ChatGPT’s 700 million weekly active users do with it? OpenAI teamed up with a Harvard economist to find out.

September 24, 20253 min read
Agents of Commerce: Google’s AP2 gives developers new tools to build agentic payments
Machine Learning Research

Agents of Commerce: Google’s AP2 gives developers new tools to build agentic payments

Google launched an open protocol for agentic payments that enables agents based on any large language model to purchase items over the internet.

September 24, 20252 min read
Transformers Energized: Energy-Based Transformers (EBTs) use gradient descent to gradually predict the next token
Machine Learning Research

Transformers Energized: Energy-Based Transformers (EBTs) use gradient descent to gradually predict the next token

A new type of transformer can check its work. Instead of guessing the next output token in one shot like a typical transformer, it starts with a rough version of the token and improves it step by step.

September 17, 20253 min read
Qwen3-Next Accelerates: Alibaba’s new model uses hybrid attention layers and a sparse MoE architecture for speed and performance
Machine Learning Research

Qwen3-Next Accelerates: Alibaba’s new model uses hybrid attention layers and a sparse MoE architecture for speed and performance

Alibaba updated its popular Qwen3 open-weights models with a number of fresh, speed-boosting tweaks.

September 17, 20253 min read
10 Million Tokens of Input Context: ATLAS, a transformer-like architecture, can process a context window as large as ten million tokens
Machine Learning Research

10 Million Tokens of Input Context: ATLAS, a transformer-like architecture, can process a context window as large as ten million tokens

An alternative to attention enables large language models to track relationships among words across extraordinarily wide spans of text.

September 10, 20253 min read
Cybersecurity for Agents: Meta releases LlamaFirewall, an open-source defense against AI hijacking
Machine Learning Research

Cybersecurity for Agents: Meta releases LlamaFirewall, an open-source defense against AI hijacking

Autonomous agents built on large language models introduce distinct security concerns. Researchers designed a system to protect agents from common vulnerabilities.

September 3, 20252 min read

Subscribe to The Batch

Stay updated with weekly AI News and Insights delivered to your inbox