Google Gets Character.AI Co-Founders: Google acquires Character.AI talent and tech in strategic move
Large Language Models (LLMs)

Google Gets Character.AI Co-Founders: Google acquires Character.AI talent and tech in strategic move

Character.AI followed an emerging pattern for ambitious AI startups, trading its leadership to a tech giant in exchange for funds and a strategic makeover. 

August 7, 20242 min read
Art Attack: ArtPrompt, a technique that exploits ASCII art to bypass LLM safety measures
Large Language Models (LLMs)

Art Attack: ArtPrompt, a technique that exploits ASCII art to bypass LLM safety measures

Seemingly an innocuous form of expression, ASCII art opens a new vector for jailbreak attacks on large language models (LLMs), enabling them to generate outputs that their developers tuned them to avoid producing.

August 7, 20242 min read
Art Attack: ArtPrompt, a technique that exploits ASCII art to bypass LLM safety measures
Large Language Models (LLMs)

Art Attack: ArtPrompt, a technique that exploits ASCII art to bypass LLM safety measures

Seemingly an innocuous form of expression, ASCII art opens a new vector for jailbreak attacks on large language models (LLMs), enabling them to generate outputs that their developers tuned them to avoid producing.

August 7, 20242 min read
Synthetic Data Factory: AgentInstruct, a framework for generating diverse synthetic data for LLM fine-tuning
Large Language Models (LLMs)

Synthetic Data Factory: AgentInstruct, a framework for generating diverse synthetic data for LLM fine-tuning

Researchers increasingly fine-tune models on synthetic data, but generated datasets may not be sufficiently diverse. New work used agentic workflows to produce diverse synthetic datasets.

July 31, 20242 min read
Synthetic Data Factory: AgentInstruct, a framework for generating diverse synthetic data for LLM fine-tuning
Large Language Models (LLMs)

Synthetic Data Factory: AgentInstruct, a framework for generating diverse synthetic data for LLM fine-tuning

Researchers increasingly fine-tune models on synthetic data, but generated datasets may not be sufficiently diverse. New work used agentic workflows to produce diverse synthetic datasets.

July 31, 20242 min read
Web Data Increasingly Off Limits: Online publishers crack down on AI training data access
Large Language Models (LLMs)

Web Data Increasingly Off Limits: Online publishers crack down on AI training data access

Online publishers are moving to stop AI developers from training models on their content.

July 31, 20243 min read
Web Data Increasingly Off Limits: Online publishers crack down on AI training data access
Large Language Models (LLMs)

Web Data Increasingly Off Limits: Online publishers crack down on AI training data access

Online publishers are moving to stop AI developers from training models on their content.

July 31, 20243 min read
Search Gets Conversational: OpenAI launches SearchGPT to rival Google and Microsoft
Large Language Models (LLMs)

Search Gets Conversational: OpenAI launches SearchGPT to rival Google and Microsoft

OpenAI is testing an AI-powered search engine in a bid to compete head-to-head with both Google and its close partner Microsoft Bing. 

July 31, 20242 min read
Search Gets Conversational: OpenAI launches SearchGPT to rival Google and Microsoft
Large Language Models (LLMs)

Search Gets Conversational: OpenAI launches SearchGPT to rival Google and Microsoft

OpenAI is testing an AI-powered search engine in a bid to compete head-to-head with both Google and its close partner Microsoft Bing. 

July 31, 20242 min read
The State of the Art Is Open: Meta’s Llama 3.1 outperforms GPT-4 in key areas
Large Language Models (LLMs)

The State of the Art Is Open: Meta’s Llama 3.1 outperforms GPT-4 in key areas

Meta raised the bar for large language models with open weights and published details about how it built one that outperforms GPT-4o and Claude 3.5 Sonnet by some measures.

July 31, 20244 min read
The State of the Art Is Open: Meta’s Llama 3.1 outperforms GPT-4 in key areas
Large Language Models (LLMs)

The State of the Art Is Open: Meta’s Llama 3.1 outperforms GPT-4 in key areas

Meta raised the bar for large language models with open weights and published details about how it built one that outperforms GPT-4o and Claude 3.5 Sonnet by some measures.

July 31, 20244 min read
Mini but Mighty: OpenAI’s GPT-4o Mini offers big performance at a small price
Large Language Models (LLMs)

Mini but Mighty: OpenAI’s GPT-4o Mini offers big performance at a small price

A slimmed-down version of Open AI’s multimodal flagship packs a low-price punch.

July 24, 20242 min read
Mini but Mighty: OpenAI’s GPT-4o Mini offers big performance at a small price
Large Language Models (LLMs)

Mini but Mighty: OpenAI’s GPT-4o Mini offers big performance at a small price

A slimmed-down version of Open AI’s multimodal flagship packs a low-price punch.

July 24, 20242 min read
Hallucination Detector: Oxford scientists propose effective method to detect AI hallucinations
Large Language Models (LLMs)

Hallucination Detector: Oxford scientists propose effective method to detect AI hallucinations

Large language models can produce output that’s convincing but false. Researchers proposed a way to identify such hallucinations. 

July 17, 20243 min read
Hallucination Detector: Oxford scientists propose effective method to detect AI hallucinations
Large Language Models (LLMs)

Hallucination Detector: Oxford scientists propose effective method to detect AI hallucinations

Large language models can produce output that’s convincing but false. Researchers proposed a way to identify such hallucinations. 

July 17, 20243 min read

Subscribe to The Batch

Stay updated with weekly AI News and Insights delivered to your inbox