Bias

196 Posts

Context Is Everything: Gemini 1.5 Pro, a leap in multimodal AI amid controversy over v1.0
Bias

Context Is Everything: Gemini 1.5 Pro, a leap in multimodal AI amid controversy over v1.0

An update of Google’s flagship multimodal model keeps track of colossal inputs, while an earlier version generated some questionable outputs.

February 28, 20243 min read
Context Is Everything: Gemini 1.5 Pro, a leap in multimodal AI amid controversy over v1.0
Bias

Context Is Everything: Gemini 1.5 Pro, a leap in multimodal AI amid controversy over v1.0

An update of Google’s flagship multimodal model keeps track of colossal inputs, while an earlier version generated some questionable outputs.

February 28, 20243 min read
New Leaderboards Rank Safety, More: Hugging Face introduces leaderboards to evaluate model performance and trustworthiness.
Bias

New Leaderboards Rank Safety, More: Hugging Face introduces leaderboards to evaluate model performance and trustworthiness.

Hugging Face introduced four leaderboards to rank the performance and trustworthiness of large language models (LLMs). The open source AI repository now ranks performance on tests of workplace utility, trust and safety, tendency to generate falsehoods, and reasoning.

February 7, 20242 min read
New Leaderboards Rank Safety, More: Hugging Face introduces leaderboards to evaluate model performance and trustworthiness.
Bias

New Leaderboards Rank Safety, More: Hugging Face introduces leaderboards to evaluate model performance and trustworthiness.

Hugging Face introduced four leaderboards to rank the performance and trustworthiness of large language models (LLMs). The open source AI repository now ranks performance on tests of workplace utility, trust and safety, tendency to generate falsehoods, and reasoning.

February 7, 20242 min read
Seeing Darker-Skinned Pedestrians: Children and people with darker skin face higher street risks with object detectors, research finds.
Bias

Seeing Darker-Skinned Pedestrians: Children and people with darker skin face higher street risks with object detectors, research finds.

In a study, models used to detect people walking on streets and sidewalks performed less well on adults with darker skin and children of all skin tones.

December 6, 20232 min read
Seeing Darker-Skinned Pedestrians: Children and people with darker skin face higher street risks with object detectors, research finds.
Bias

Seeing Darker-Skinned Pedestrians: Children and people with darker skin face higher street risks with object detectors, research finds.

In a study, models used to detect people walking on streets and sidewalks performed less well on adults with darker skin and children of all skin tones.

December 6, 20232 min read
Amazon Joins Chatbot Fray: The pros and cons of Q, Amazon’s new enterprise chatbot
Bias

Amazon Joins Chatbot Fray: The pros and cons of Q, Amazon’s new enterprise chatbot

Amazon launched a chatbot for large companies even as internal tests indicated potential problems. Amazon introduced Q, an AI-powered assistant that enables employees to query documents and corporate systems.

December 6, 20232 min read
Amazon Joins Chatbot Fray: The pros and cons of Q, Amazon’s new enterprise chatbot
Bias

Amazon Joins Chatbot Fray: The pros and cons of Q, Amazon’s new enterprise chatbot

Amazon launched a chatbot for large companies even as internal tests indicated potential problems. Amazon introduced Q, an AI-powered assistant that enables employees to query documents and corporate systems.

December 6, 20232 min read
Testing for Large Language Models: Meet Giskard, an automated quality manager for LLMs.
Bias

Testing for Large Language Models: Meet Giskard, an automated quality manager for LLMs.

An open source tool automatically tests language and tabular-data models for social biases and other common issues. Giskard is a software framework that evaluates models using a suite of heuristics and tests based on GPT-4.

November 29, 20232 min read
Testing for Large Language Models: Meet Giskard, an automated quality manager for LLMs.
Bias

Testing for Large Language Models: Meet Giskard, an automated quality manager for LLMs.

An open source tool automatically tests language and tabular-data models for social biases and other common issues. Giskard is a software framework that evaluates models using a suite of heuristics and tests based on GPT-4.

November 29, 20232 min read
More Scraped Data, Greater Bias: Research shows that training on larger datasets can increase social bias.
Bias

More Scraped Data, Greater Bias: Research shows that training on larger datasets can increase social bias.

How can we build large-scale language and vision models that don’t inherit social biases? Conventional wisdom suggests training on larger datasets, but research challenges this assumption.

October 4, 20232 min read
More Scraped Data, Greater Bias: Research shows that training on larger datasets can increase social bias.
Bias

More Scraped Data, Greater Bias: Research shows that training on larger datasets can increase social bias.

How can we build large-scale language and vision models that don’t inherit social biases? Conventional wisdom suggests training on larger datasets, but research challenges this assumption.

October 4, 20232 min read
Stable Biases: Stable Diffusion may amplify biases in its training data.
Bias

Stable Biases: Stable Diffusion may amplify biases in its training data.

Stable Diffusion may amplify biases in its training data in ways that promote deeply ingrained social stereotypes.

July 12, 20233 min read
Stable Biases: Stable Diffusion may amplify biases in its training data.
Bias

Stable Biases: Stable Diffusion may amplify biases in its training data.

Stable Diffusion may amplify biases in its training data in ways that promote deeply ingrained social stereotypes.

July 12, 20233 min read
Where Is Meta’s Generative Play?: Why Meta still lacks a flagship generative AI service
Bias

Where Is Meta’s Generative Play?: Why Meta still lacks a flagship generative AI service

While Microsoft and Google scramble to supercharge their businesses with text generation, Meta has yet to launch a flagship generative AI service. Reporters went looking for reasons why.

June 28, 20232 min read

Subscribe to The Batch

Stay updated with weekly AI News and Insights delivered to your inbox