More Factual LLMs: FactTune, a method to fine-tune LLMs for factual accuracy without human feedback
Large Language Models (LLMs)

More Factual LLMs: FactTune, a method to fine-tune LLMs for factual accuracy without human feedback

Large language models sometimes generate false statements. New work makes them more likely to produce factual output.

April 3, 20242 min read
Microsoft Absorbs Inflection: Microsoft pays Inflection AI $650 Million, hires most of its staff
Large Language Models (LLMs)

Microsoft Absorbs Inflection: Microsoft pays Inflection AI $650 Million, hires most of its staff

Microsoft took over most of the once high-flying chatbot startup Inflection AI in an unusual deal.

April 3, 20242 min read
Microsoft Absorbs Inflection: Microsoft pays Inflection AI $650 Million, hires most of its staff
Large Language Models (LLMs)

Microsoft Absorbs Inflection: Microsoft pays Inflection AI $650 Million, hires most of its staff

Microsoft took over most of the once high-flying chatbot startup Inflection AI in an unusual deal.

April 3, 20242 min read
Cutting the Cost of Pretrained Models: FrugalGPT, a method to cut AI costs and maintain quality
Large Language Models (LLMs)

Cutting the Cost of Pretrained Models: FrugalGPT, a method to cut AI costs and maintain quality

Research aims to help users select large language models that minimize expenses while maintaining quality.

March 21, 20242 min read
Cutting the Cost of Pretrained Models: FrugalGPT, a method to cut AI costs and maintain quality
Large Language Models (LLMs)

Cutting the Cost of Pretrained Models: FrugalGPT, a method to cut AI costs and maintain quality

Research aims to help users select large language models that minimize expenses while maintaining quality.

March 21, 20242 min read
Some Models Pose Security Risk: Security flaws exposed in Hugging Face’s repository and security features
Large Language Models (LLMs)

Some Models Pose Security Risk: Security flaws exposed in Hugging Face’s repository and security features

Security researchers sounded the alarm about holes in Hugging Face’s platform.

March 21, 20242 min read
Some Models Pose Security Risk: Security flaws exposed in Hugging Face’s repository and security features
Large Language Models (LLMs)

Some Models Pose Security Risk: Security flaws exposed in Hugging Face’s repository and security features

Security researchers sounded the alarm about holes in Hugging Face’s platform.

March 21, 20242 min read
Schooling Language Models in Math: GOAT (Good at Arithmetic Tasks), a method to boost large language models' arithmetic abilities
Large Language Models (LLMs)

Schooling Language Models in Math: GOAT (Good at Arithmetic Tasks), a method to boost large language models' arithmetic abilities

Large language models are not good at math. Researchers devised a way to make them better. Tiedong Liu and Bryan Kian Hsiang Low at the National University of Singapore proposed a method to fine-tune large language models for arithmetic tasks.

March 6, 20242 min read
Schooling Language Models in Math: GOAT (Good at Arithmetic Tasks), a method to boost large language models' arithmetic abilities
Large Language Models (LLMs)

Schooling Language Models in Math: GOAT (Good at Arithmetic Tasks), a method to boost large language models' arithmetic abilities

Large language models are not good at math. Researchers devised a way to make them better. Tiedong Liu and Bryan Kian Hsiang Low at the National University of Singapore proposed a method to fine-tune large language models for arithmetic tasks.

March 6, 20242 min read
Google Releases Open Source LLMs: All we know about Google's Gemma-7B and Gemma-2B models
Large Language Models (LLMs)

Google Releases Open Source LLMs: All we know about Google's Gemma-7B and Gemma-2B models

Google asserted its open source bona fides with new models. Google released weights for Gemma-7B, an 8.5 billion-parameter large language model intended to run GPUs, and Gemma-2B, a 2.5 billion-parameter version intended for deployment on CPUs and edge devices.

March 6, 20242 min read
Google Releases Open Source LLMs: All we know about Google's Gemma-7B and Gemma-2B models
Large Language Models (LLMs)

Google Releases Open Source LLMs: All we know about Google's Gemma-7B and Gemma-2B models

Google asserted its open source bona fides with new models. Google released weights for Gemma-7B, an 8.5 billion-parameter large language model intended to run GPUs, and Gemma-2B, a 2.5 billion-parameter version intended for deployment on CPUs and edge devices.

March 6, 20242 min read
Mistral AI Extends Its Portfolio: Mistral enhances AI landscape in Europe with Microsoft partnership and new language models.
Large Language Models (LLMs)

Mistral AI Extends Its Portfolio: Mistral enhances AI landscape in Europe with Microsoft partnership and new language models.

European AI champion Mistral AI unveiled new large language models and formed an alliance with Microsoft. 

March 6, 20242 min read
Mistral AI Extends Its Portfolio: Mistral enhances AI landscape in Europe with Microsoft partnership and new language models.
Large Language Models (LLMs)

Mistral AI Extends Its Portfolio: Mistral enhances AI landscape in Europe with Microsoft partnership and new language models.

European AI champion Mistral AI unveiled new large language models and formed an alliance with Microsoft. 

March 6, 20242 min read
Swiss Army LLM
Large Language Models (LLMs)

Swiss Army LLM

The combination of  language models that are equipped for retrieval augmented generation can retrieve text from a database to improve their output. Further work extends this capability to retrieve information from any application that comes with an API. 

February 28, 20243 min read
Swiss Army LLM
Large Language Models (LLMs)

Swiss Army LLM

The combination of  language models that are equipped for retrieval augmented generation can retrieve text from a database to improve their output. Further work extends this capability to retrieve information from any application that comes with an API. 

February 28, 20243 min read

Subscribe to The Batch

Stay updated with weekly AI News and Insights delivered to your inbox