
Context Is Everything: Gemini 1.5 Pro, a leap in multimodal AI amid controversy over v1.0
An update of Google’s flagship multimodal model keeps track of colossal inputs, while an earlier version generated some questionable outputs.
196 Posts

An update of Google’s flagship multimodal model keeps track of colossal inputs, while an earlier version generated some questionable outputs.

An update of Google’s flagship multimodal model keeps track of colossal inputs, while an earlier version generated some questionable outputs.

Hugging Face introduced four leaderboards to rank the performance and trustworthiness of large language models (LLMs). The open source AI repository now ranks performance on tests of workplace utility, trust and safety, tendency to generate falsehoods, and reasoning.

Hugging Face introduced four leaderboards to rank the performance and trustworthiness of large language models (LLMs). The open source AI repository now ranks performance on tests of workplace utility, trust and safety, tendency to generate falsehoods, and reasoning.

In a study, models used to detect people walking on streets and sidewalks performed less well on adults with darker skin and children of all skin tones.

In a study, models used to detect people walking on streets and sidewalks performed less well on adults with darker skin and children of all skin tones.

Amazon launched a chatbot for large companies even as internal tests indicated potential problems. Amazon introduced Q, an AI-powered assistant that enables employees to query documents and corporate systems.

Amazon launched a chatbot for large companies even as internal tests indicated potential problems. Amazon introduced Q, an AI-powered assistant that enables employees to query documents and corporate systems.

An open source tool automatically tests language and tabular-data models for social biases and other common issues. Giskard is a software framework that evaluates models using a suite of heuristics and tests based on GPT-4.

An open source tool automatically tests language and tabular-data models for social biases and other common issues. Giskard is a software framework that evaluates models using a suite of heuristics and tests based on GPT-4.

How can we build large-scale language and vision models that don’t inherit social biases? Conventional wisdom suggests training on larger datasets, but research challenges this assumption.

How can we build large-scale language and vision models that don’t inherit social biases? Conventional wisdom suggests training on larger datasets, but research challenges this assumption.

Stable Diffusion may amplify biases in its training data in ways that promote deeply ingrained social stereotypes.

Stable Diffusion may amplify biases in its training data in ways that promote deeply ingrained social stereotypes.

While Microsoft and Google scramble to supercharge their businesses with text generation, Meta has yet to launch a flagship generative AI service. Reporters went looking for reasons why.
Stay updated with weekly AI News and Insights delivered to your inbox