Harm

234 Posts

GPT Store Shows Lax Moderation: A report exposes policy violations in OpenAI’s GPT Store.
Harm

GPT Store Shows Lax Moderation: A report exposes policy violations in OpenAI’s GPT Store.

OpenAI has been moderating its GPT Store with a very light touch. In a survey of the GPT Store’s offerings, TechCrunch found numerous examples of custom ChatGPT instances that appear to violate the store’s own policies.

April 18, 20242 min read
GPT Store Shows Lax Moderation: A report exposes policy violations in OpenAI’s GPT Store.
Harm

GPT Store Shows Lax Moderation: A report exposes policy violations in OpenAI’s GPT Store.

OpenAI has been moderating its GPT Store with a very light touch. In a survey of the GPT Store’s offerings, TechCrunch found numerous examples of custom ChatGPT instances that appear to violate the store’s own policies.

April 18, 20242 min read
Toward Managing AI Bio Risk: Over 150 scientists commit to ensure AI safety in synthetic biology research.
Harm

Toward Managing AI Bio Risk: Over 150 scientists commit to ensure AI safety in synthetic biology research.

Scientists pledged to control their use of AI to produce potentially hazardous biological materials.

April 3, 20242 min read
Toward Managing AI Bio Risk: Over 150 scientists commit to ensure AI safety in synthetic biology research.
Harm

Toward Managing AI Bio Risk: Over 150 scientists commit to ensure AI safety in synthetic biology research.

Scientists pledged to control their use of AI to produce potentially hazardous biological materials.

April 3, 20242 min read
Deepfakes Become Politics as Usual: Deepfakes dominate as India’s election season unfolds.
Harm

Deepfakes Become Politics as Usual: Deepfakes dominate as India’s election season unfolds.

Synthetic depictions of politicians are taking center stage as the world’s biggest democratic election kicks off.

March 21, 20242 min read
Deepfakes Become Politics as Usual: Deepfakes dominate as India’s election season unfolds.
Harm

Deepfakes Become Politics as Usual: Deepfakes dominate as India’s election season unfolds.

Synthetic depictions of politicians are taking center stage as the world’s biggest democratic election kicks off.

March 21, 20242 min read
U.S. Restricts AI Robocalls: U.S. cracks down on AI-generated voice robocalls to combat election interference.
Harm

U.S. Restricts AI Robocalls: U.S. cracks down on AI-generated voice robocalls to combat election interference.

The United States outlawed unsolicited phone calls that use AI-generated voices. 

February 14, 20242 min read
U.S. Restricts AI Robocalls: U.S. cracks down on AI-generated voice robocalls to combat election interference.
Harm

U.S. Restricts AI Robocalls: U.S. cracks down on AI-generated voice robocalls to combat election interference.

The United States outlawed unsolicited phone calls that use AI-generated voices. 

February 14, 20242 min read
GPT-4 Biothreat Risk is Low: Study finds GPT-4 no more risky than online search in aiding bioweapon development.
Harm

GPT-4 Biothreat Risk is Low: Study finds GPT-4 no more risky than online search in aiding bioweapon development.

GPT-4 poses negligible additional risk that a malefactor could build a biological weapon, according to a new study. OpenAI compared the ability of GPT-4 and web search to contribute to the creation of a dangerous virus or bacterium. The large language model was barely more helpful than the web.

February 7, 20242 min read
GPT-4 Biothreat Risk is Low: Study finds GPT-4 no more risky than online search in aiding bioweapon development.
Harm

GPT-4 Biothreat Risk is Low: Study finds GPT-4 no more risky than online search in aiding bioweapon development.

GPT-4 poses negligible additional risk that a malefactor could build a biological weapon, according to a new study. OpenAI compared the ability of GPT-4 and web search to contribute to the creation of a dangerous virus or bacterium. The large language model was barely more helpful than the web.

February 7, 20242 min read
New Leaderboards Rank Safety, More: Hugging Face introduces leaderboards to evaluate model performance and trustworthiness.
Harm

New Leaderboards Rank Safety, More: Hugging Face introduces leaderboards to evaluate model performance and trustworthiness.

Hugging Face introduced four leaderboards to rank the performance and trustworthiness of large language models (LLMs). The open source AI repository now ranks performance on tests of workplace utility, trust and safety, tendency to generate falsehoods, and reasoning.

February 7, 20242 min read
New Leaderboards Rank Safety, More: Hugging Face introduces leaderboards to evaluate model performance and trustworthiness.
Harm

New Leaderboards Rank Safety, More: Hugging Face introduces leaderboards to evaluate model performance and trustworthiness.

Hugging Face introduced four leaderboards to rank the performance and trustworthiness of large language models (LLMs). The open source AI repository now ranks performance on tests of workplace utility, trust and safety, tendency to generate falsehoods, and reasoning.

February 7, 20242 min read
Nude Deepfakes Spur Legislators: Taylor Swift deepfake outrage prompts U.S. lawmakers to propose anti-AI pornography laws.
Harm

Nude Deepfakes Spur Legislators: Taylor Swift deepfake outrage prompts U.S. lawmakers to propose anti-AI pornography laws.

Sexually explicit deepfakes of Taylor Swift galvanized public demand for laws against nonconsensual, AI-enabled pornography.

February 7, 20243 min read
Nude Deepfakes Spur Legislators: Taylor Swift deepfake outrage prompts U.S. lawmakers to propose anti-AI pornography laws.
Harm

Nude Deepfakes Spur Legislators: Taylor Swift deepfake outrage prompts U.S. lawmakers to propose anti-AI pornography laws.

Sexually explicit deepfakes of Taylor Swift galvanized public demand for laws against nonconsensual, AI-enabled pornography.

February 7, 20243 min read
Standard for Media Watermarks: C2PA introduces watermark tech to combat media misinformation.
Harm

Standard for Media Watermarks: C2PA introduces watermark tech to combat media misinformation.

An alliance of major tech and media companies introduced a watermark designed to distinguish real from fake media starting with images. The Coalition for Content Provenance and Authenticity (C2PA) offers an open standard that marks media files with information about their creation and editing.

January 17, 20243 min read

Subscribe to The Batch

Stay updated with weekly AI News and Insights delivered to your inbox