Teaching Models to Tell the Truth: OpenAI fine-tuned a version of GPT-5 to confess when it was breaking the rules
Large Language Models (LLMs)

Teaching Models to Tell the Truth: OpenAI fine-tuned a version of GPT-5 to confess when it was breaking the rules

Large language models occasionally conceal their failures to comply with constraints they’ve been trained or prompted to observe. Researchers trained an LLM to admit when it disobeyed.

January 9, 20262 min read
Chatbots That Build Community by Sharon Zhou: Sharon Zhou of AMD on expanding chat to serve groups and connect us with other people
Large Language Models (LLMs)

Chatbots That Build Community by Sharon Zhou: Sharon Zhou of AMD on expanding chat to serve groups and connect us with other people

Next year, I’m excited to see AI break out of 1:1 relationships with each of us. In 2026, AI has the potential to bring people together and unite us with human connection, rather than polarize and isolate us. It’s about time for ChatGPT to enter your group chats.

January 2, 20262 min read
Agents Write Code Faster, Cheaper: Software developers used more versatile AI-powered tools to write code
Large Language Models (LLMs)

Agents Write Code Faster, Cheaper: Software developers used more versatile AI-powered tools to write code

Coding apps moved beyond autofill-style code completion to agentic systems that manage a wide range of software development tasks.

December 26, 20253 min read
Thinking Models Solve Bigger Problems: Reasoning models, beginning with OpenAI’s o1 and DeepSeek’s R1, transformed the industry
Large Language Models (LLMs)

Thinking Models Solve Bigger Problems: Reasoning models, beginning with OpenAI’s o1 and DeepSeek’s R1, transformed the industry

Think step by step. Explain your reasoning. Work backwards from the answer. As 2025 began, models executed these reasoning strategies only when prompted. Now most new large language models do it as a matter of course, improving performance across a wide range of tasks.

December 26, 20253 min read
Adapting LLMs to Any Sort of Data: SEMI (Sample-Efficient Modality Integration) tackles new domains with few-shot examples
Large Language Models (LLMs)

Adapting LLMs to Any Sort of Data: SEMI (Sample-Efficient Modality Integration) tackles new domains with few-shot examples

Enabling a pretrained large language model to process a data type other than text (say, images), possibly in a specialized domain (say, radiology), typically requires thousands to millions of examples that pair the other data (perhaps x-rays) with text.

December 17, 20253 min read
OpenAI’s Answer to Gemini 3: GPT-5.2 arrives, touting variable reasoning and coding performance
Large Language Models (LLMs)

OpenAI’s Answer to Gemini 3: GPT-5.2 arrives, touting variable reasoning and coding performance

OpenAI launched GPT-5.2 only weeks after its CEO Sam Altman reportedly issued a “code red” alarm in response to Google's Gemini 3.

December 17, 20253 min read
Claude Does More With Fewer Tokens: Claude Opus 4.5 retakes the coding crown at one-third the price of its predecessor
Large Language Models (LLMs)

Claude Does More With Fewer Tokens: Claude Opus 4.5 retakes the coding crown at one-third the price of its predecessor

Claude Opus 4.5, the latest version of Anthropic’s flagship model, extends the earlier version’s strengths in coding, computer use, and agentic workflows while generating fewer tokens.

December 10, 20253 min read
Toward Steering LLM Personality: Persona Vectors allow model builders to identify and edit out sycophancy, hallucinations, and more
Large Language Models (LLMs)

Toward Steering LLM Personality: Persona Vectors allow model builders to identify and edit out sycophancy, hallucinations, and more

Large language models can develop character traits like cheerfulness or sycophancy during fine-tuning. Researchers developed a method to identify, monitor, and control such traits.

November 26, 20253 min read
Microsoft and Anthropic Form Alliance: Claude becomes the first leading language model available from all three cloud giants
Large Language Models (LLMs)

Microsoft and Anthropic Form Alliance: Claude becomes the first leading language model available from all three cloud giants

Having recently revised its agreement with longtime partner OpenAI, Microsoft pledged to invest billions of dollars in Anthropic, one of OpenAI’s top competitors.

November 26, 20253 min read
More-Efficient Agentic Search: Researchers fine-tune models to search their own parameters to boost recall
Large Language Models (LLMs)

More-Efficient Agentic Search: Researchers fine-tune models to search their own parameters to boost recall

Large language models may have learned knowledge that’s relevant to a given prompt, but they don’t always recall it consistently. Fine-tuning a model to search its parameters as though it were searching the web can help it find knowledge in its own weights.

November 19, 20253 min read
Top Agentic Results, Open Weights: Kimi K2 Thinking outperforms proprietary models with new techniques for agentic tool use
Large Language Models (LLMs)

Top Agentic Results, Open Weights: Kimi K2 Thinking outperforms proprietary models with new techniques for agentic tool use

The latest open-weights large language model from Moonshot AI challenges top proprietary LLMs at agentic tasks by executing hundreds of tool calls sequentially and pausing to think between each.

November 19, 20254 min read
Masking Private Data in Training Sets: Google researchers released VaultGemma, an open-weights model redacting personal information
Large Language Models (LLMs)

Masking Private Data in Training Sets: Google researchers released VaultGemma, an open-weights model redacting personal information

Large language models often memorize details in their training data, including private information that may appear only once, like a person’s name, address, or phone number. Researchers built the first open-weights language model that’s guaranteed not to remember such facts.

November 5, 20253 min read
Open-Weights Coding Leader: MiniMax-M2’s lightweight footprint and low costs belie that its top performance
Large Language Models (LLMs)

Open-Weights Coding Leader: MiniMax-M2’s lightweight footprint and low costs belie that its top performance

An open-weights model from Shanghai-based MiniMax challenges top proprietary models on key benchmarks for coding and agentic tasks.

November 5, 20253 min read
Web Data Diminishes: What if online publishers make it harder and more expensive to train models?
Large Language Models (LLMs)

Web Data Diminishes: What if online publishers make it harder and more expensive to train models?

For decades, AI developers have treated the web as an open faucet of training data. Now publishers are shutting the tap. Will web data dry up?

October 29, 20253 min read
Chatbots Lead Users Into Rabbit Holes: When paranoia, delusions, and other signs of mental illness meet AI
Large Language Models (LLMs)

Chatbots Lead Users Into Rabbit Holes: When paranoia, delusions, and other signs of mental illness meet AI

Conversations with chatbots are loosening users’ grips on reality, fueling the sorts of delusions that can trigger episodes of severe mental illness. Are AI models driving us insane?

October 29, 20253 min read

Subscribe to The Batch

Stay updated with weekly AI News and Insights delivered to your inbox