GLM-5.1 Aims for Long-Running Tasks: Z.ai’s GLM 5.1 evaluates interim results and may change its approach hundreds of times before it delivers final output
Machine Learning Research

GLM-5.1 Aims for Long-Running Tasks: Z.ai’s GLM 5.1 evaluates interim results and may change its approach hundreds of times before it delivers final output

Z.ai updated its flagship open-weights large language model to work autonomously on single tasks for up to eight hours.

April 24, 20263 min read
Simulating Diverse Human Cohorts: Persona generation simulates human characters across a controllable range of points of view
Machine Learning Research

Simulating Diverse Human Cohorts: Persona generation simulates human characters across a controllable range of points of view

If you want to understand how the public will respond to your offerings, large language models can simulate users who answer questions about capabilities, features, promotions, or prices.

April 17, 20263 min read
US States Move Forward With AI Laws: Most states are regulating AI despite President Trump’s opposition to state-level laws
Machine Learning Research

US States Move Forward With AI Laws: Most states are regulating AI despite President Trump’s opposition to state-level laws

U.S. states are continuing to enact laws that regulate AI, despite President Trump’s efforts to discourage state-by-state legislation in favor of national laws.

April 17, 20264 min read
Big Pharma Bets Big on AI: Pharmaceutical kingpin Eli Lilly gave Insilico $2.75 billion for AI-driven drug development
Machine Learning Research

Big Pharma Bets Big on AI: Pharmaceutical kingpin Eli Lilly gave Insilico $2.75 billion for AI-driven drug development

Generative AI has proven that it can produce text, images, audio, video, and code. The world’s most valuable pharmaceutical company is betting billions that it can produce drugs as well.

April 17, 20263 min read
Life After Llama: With Muse Spark, Meta pivots away from its open-weights Llama strategy
Machine Learning Research

Life After Llama: With Muse Spark, Meta pivots away from its open-weights Llama strategy

Meta pivoted from its open-weights strategy to deliver a closed alternative.

April 17, 20263 min read
How Liquids and Gases Behave: A dynamic fluids model appears to solve transformers’ pixellation problem
Machine Learning Research

How Liquids and Gases Behave: A dynamic fluids model appears to solve transformers’ pixellation problem

Simulating complex physical systems through traditional numerical methods is slow and expensive, and simulations based on machine learning are usually specialized for a specific type of system, such as water in a pipe or atmosphere surrounding a planet.

April 10, 20263 min read
Dark DNA Unveiled: Google’s AlphaGenome interprets DNA that regulates genetic expression
Machine Learning Research

Dark DNA Unveiled: Google’s AlphaGenome interprets DNA that regulates genetic expression

An open-weights model could help scientists compare the impact of genetic variations, identify mutations that cause diseases, and develop treatments.

April 10, 20263 min read
Claude Mythos Preview Raises Security Worries: Why Claude’s advanced Mythos Preview model will be limited-release-only
Machine Learning Research

Claude Mythos Preview Raises Security Worries: Why Claude’s advanced Mythos Preview model will be limited-release-only

Anthropic took unusual steps to prepare the world for a forthcoming large language model that it said poses extraordinary risks to cybersecurity.

April 10, 20264 min read
Learning Long Context at Inference: Test-Time Training End-to-End (TTT-E2E) retrains model weights to handle long inputs
Machine Learning Research

Learning Long Context at Inference: Test-Time Training End-to-End (TTT-E2E) retrains model weights to handle long inputs

Large language models typically become less accurate and slower when they process longer contexts, but researchers enabled an LLM to keep accuracy stable and inference time constant as its context grew.

April 3, 20264 min read
Gemini’s Music Generator: Google debuted Lyria 3, an app that turns text or images into 30-second songs
Machine Learning Research

Gemini’s Music Generator: Google debuted Lyria 3, an app that turns text or images into 30-second songs

Google added a music generator to Gemini and YouTube, putting a model that produces synthetic songs in front of hundreds of millions of users.

April 3, 20263 min read
Inside Claude Code: Claude Code’s source code leaked, exposing potential future features Kairos and autoDream
Machine Learning Research

Inside Claude Code: Claude Code’s source code leaked, exposing potential future features Kairos and autoDream

The inner workings of the popular coding agent Claude Code are available for all to see.

April 3, 20263 min read
Context As An External Variable: Recursive Language Models offer path to aramatically expand beyond the context window
Machine Learning Research

Context As An External Variable: Recursive Language Models offer path to aramatically expand beyond the context window

When processing long contexts, large language models often lose track of details or devolve into nonsense. Researchers reduced these effects by managing context externally.

March 27, 20263 min read
xAI’s Cost-Effective Video Generator: Grok Imagine 1.0 sharply cuts costs for high-quality video generation
Machine Learning Research

xAI’s Cost-Effective Video Generator: Grok Imagine 1.0 sharply cuts costs for high-quality video generation

xAI launched a video generator that topped an independent quality ranking at a fraction of competitors’ prices.

March 27, 20263 min read
Open-Source Speed Demon: Nvidia’s open Nemotron 3 Super 120B-A12B model sets new paces in its class
Machine Learning Research

Open-Source Speed Demon: Nvidia’s open Nemotron 3 Super 120B-A12B model sets new paces in its class

Nvidia, the dominant supplier of AI chips, released a competitive open-source large language model whose speed tops its size class — the first open-weights leader to come from the United States since last year, when Meta delivered Llama 4.

March 27, 20264 min read
A Single Tokenizer for Visual Media: Apple’s AToken, a multimodal model with a single encoder and tokenizer for images, videos, and 3D objects
Machine Learning Research

A Single Tokenizer for Visual Media: Apple’s AToken, a multimodal model with a single encoder and tokenizer for images, videos, and 3D objects

Multimodal models typically use different tokenizers to embed different media types, and different encoders when training to generate media rather than classify it.

March 20, 20264 min read

Subscribe to The Batch

Stay updated with weekly AI News and Insights delivered to your inbox