
More Factual LLMs: FactTune, a method to fine-tune LLMs for factual accuracy without human feedback
Large language models sometimes generate false statements. New work makes them more likely to produce factual output.
116 Posts

Large language models sometimes generate false statements. New work makes them more likely to produce factual output.

Large language models sometimes generate false statements. New work makes them more likely to produce factual output.

Researchers used an AI system to identify animal cell types from gene sequences, including a cell type that conventional approaches had discovered only in the past year.

Researchers used an AI system to identify animal cell types from gene sequences, including a cell type that conventional approaches had discovered only in the past year.

Research aims to help users select large language models that minimize expenses while maintaining quality.

Research aims to help users select large language models that minimize expenses while maintaining quality.

Machine learning models typically learn language by training on tasks like predicting the next word in a given text. Researchers trained a language model in a less focused, more human-like way.

Machine learning models typically learn language by training on tasks like predicting the next word in a given text. Researchers trained a language model in a less focused, more human-like way.

Researchers proposed a way for robots to find objects in households where things get moved around. Andrey Kurenkov and colleagues at Stanford University introduced Node Edge Predictor, a model that learned to predict where objects were located in houses.

Researchers proposed a way for robots to find objects in households where things get moved around. Andrey Kurenkov and colleagues at Stanford University introduced Node Edge Predictor, a model that learned to predict where objects were located in houses.

A new index ranks popular AI models in terms of information their developers provide about their training, architecture, and usage. Few score well.

A new index ranks popular AI models in terms of information their developers provide about their training, architecture, and usage. Few score well.

Large language models increasingly reply to prompts with a believably human response. Can they also mimic human behavior?

Large language models increasingly reply to prompts with a believably human response. Can they also mimic human behavior?

It wasn’t your imagination: OpenAI’s large language models have changed. Researchers at Stanford and UC Berkeley found that the performance of GPT-4 and GPT-3.5 has drifted in recent months. In a limited selection of tasks, some prompts yielded better results than before, some worse.
Stay updated with weekly AI News and Insights delivered to your inbox