
Benchmarks for Industry: Vals AI evaluates large language models on industry-specific tasks.
How well do large language models respond to professional-level queries in various industry domains? A new company aims to find out.
268 Posts

How well do large language models respond to professional-level queries in various industry domains? A new company aims to find out.

How well do large language models respond to professional-level queries in various industry domains? A new company aims to find out.

A new breed of audio generator produces synthetic performances of songs in a variety of popular styles.

A new breed of audio generator produces synthetic performances of songs in a variety of popular styles.

Google is empowering developers to build autonomous agents using little or no custom code.

Google is empowering developers to build autonomous agents using little or no custom code.

Humanoid robots can play football (known as soccer in the United States) in the real world, thanks to reinforcement learning.

Humanoid robots can play football (known as soccer in the United States) in the real world, thanks to reinforcement learning.

AI agents are typically designed to operate a particular software environment. Recent work enabled a single agent to take actions in a variety of three-dimensional virtual worlds.

AI agents are typically designed to operate a particular software environment. Recent work enabled a single agent to take actions in a variety of three-dimensional virtual worlds.

Google is paying newsrooms to use a system that helps transform press releases into articles.

Google is paying newsrooms to use a system that helps transform press releases into articles.

Google asserted its open source bona fides with new models. Google released weights for Gemma-7B, an 8.5 billion-parameter large language model intended to run GPUs, and Gemma-2B, a 2.5 billion-parameter version intended for deployment on CPUs and edge devices.

Google asserted its open source bona fides with new models. Google released weights for Gemma-7B, an 8.5 billion-parameter large language model intended to run GPUs, and Gemma-2B, a 2.5 billion-parameter version intended for deployment on CPUs and edge devices.

An update of Google’s flagship multimodal model keeps track of colossal inputs, while an earlier version generated some questionable outputs.
Stay updated with weekly AI News and Insights delivered to your inbox