Nemotron models boost Llama’s speed but maintain accuracy: NotebookLM “reads” audio and video
Data Points

Nemotron models boost Llama’s speed but maintain accuracy: NotebookLM “reads” audio and video

Hacking ChatGPT’s long-term memory function. U.S. trade commission targets companies who lie about AI. A new OpenAI model for screening text and images. Apple, Meta, and others hold off on AI Pact

September 30, 20243 min read
Nemotron models boost Llama’s speed but maintain accuracy: NotebookLM “reads” audio and video
Data Points

Nemotron models boost Llama’s speed but maintain accuracy: NotebookLM “reads” audio and video

Hacking ChatGPT’s long-term memory function. U.S. trade commission targets companies who lie about AI. A new OpenAI model for screening text and images. Apple, Meta, and others hold off on AI Pact

September 30, 20243 min read
Molmo’s impressive open multimodal models: Llama 3.2 adds vision models and small LLMs
Data Points

Molmo’s impressive open multimodal models: Llama 3.2 adds vision models and small LLMs

Google’s top Gemini models cut prices, boost performance. Microsoft’s new approach to correct hallucinations. OpenAI releases a multilingual training dataset. Chain-of-thought reasoning has limits.

September 27, 20244 min read
Molmo’s impressive open multimodal models: Llama 3.2 adds vision models and small LLMs
Data Points

Molmo’s impressive open multimodal models: Llama 3.2 adds vision models and small LLMs

Google’s top Gemini models cut prices, boost performance. Microsoft’s new approach to correct hallucinations. OpenAI releases a multilingual training dataset. Chain-of-thought reasoning has limits.

September 27, 20244 min read
Using GPT to debunk conspiracy theories: Nvidia’s new open vision-language models
Data Points

Using GPT to debunk conspiracy theories: Nvidia’s new open vision-language models

Kling 1.5 text-to-video model adds 1080p, editing tools. Codeforces restricts AI use in competition. Identifying whalesong using bioacoustic markers. A prize challenge for improving robotics world models.

September 23, 20243 min read
Using GPT to debunk conspiracy theories: Nvidia’s new open vision-language models
Data Points

Using GPT to debunk conspiracy theories: Nvidia’s new open vision-language models

Kling 1.5 text-to-video model adds 1080p, editing tools. Codeforces restricts AI use in competition. Identifying whalesong using bioacoustic markers. A prize challenge for improving robotics world models.

September 23, 20243 min read
Alibaba’s impressive suite of open models: Mistral cuts prices across its lineup
Data Points

Alibaba’s impressive suite of open models: Mistral cuts prices across its lineup

Runway adds video-to-video and API. Moshi, a new open speech model. LlamaCoder’s open webapp builder alternative. California restricts synthetic actors and election deepfakes.

September 20, 20243 min read
Alibaba’s impressive suite of open models: Mistral cuts prices across its lineup
Data Points

Alibaba’s impressive suite of open models: Mistral cuts prices across its lineup

Runway adds video-to-video and API. Moshi, a new open speech model. LlamaCoder’s open webapp builder alternative. California restricts synthetic actors and election deepfakes.

September 20, 20243 min read
A new language model tool for web scraping and conversion: Plus, a plan to combat AI sexual abuse imagery
Data Points

A new language model tool for web scraping and conversion: Plus, a plan to combat AI sexual abuse imagery

Hugging Face open-sources an LLM evaluation suite, Adobe announces its Firefly Video model, Meta researchers blend image diffusion with text transformers, NotebookLM can now generate synthetic podcasts.

September 16, 20243 min read
A new language model tool for web scraping and conversion: Plus, a plan to combat AI sexual abuse imagery
Data Points

A new language model tool for web scraping and conversion: Plus, a plan to combat AI sexual abuse imagery

Hugging Face open-sources an LLM evaluation suite, Adobe announces its Firefly Video model, Meta researchers blend image diffusion with text transformers, NotebookLM can now generate synthetic podcasts.

September 16, 20243 min read
OpenAI’s o1 models recognize and fix mistakes: Plus, explaining Reflection 70B’s replication controversy
Data Points

OpenAI’s o1 models recognize and fix mistakes: Plus, explaining Reflection 70B’s replication controversy

Copilot adds fine-tuning for faster code completion, DataGemma uses RAG and RIG for fact-retrieval, Mistral introduces its open multimodal model, results of the latest summit on military AI.

September 13, 20243 min read
OpenAI’s o1 models recognize and fix mistakes: Plus, explaining Reflection 70B’s replication controversy
Data Points

OpenAI’s o1 models recognize and fix mistakes: Plus, explaining Reflection 70B’s replication controversy

Copilot adds fine-tuning for faster code completion, DataGemma uses RAG and RIG for fact-retrieval, Mistral introduces its open multimodal model, results of the latest summit on military AI.

September 13, 20243 min read
Replit Agent builds and deploys applications using natural language prompts: DeepSeek-V2.5’s open model blends coding and chat
Data Points

Replit Agent builds and deploys applications using natural language prompts: DeepSeek-V2.5’s open model blends coding and chat

NVIDIA’s Blackwell chips impress on hardware tests, fine-tuned versions of Llama 3.1 add reflection, most AI jailbreaks may not amount to much, new architecture extends context windows to 100M tokens.

September 9, 20243 min read
Replit Agent builds and deploys applications using natural language prompts: DeepSeek-V2.5’s open model blends coding and chat
Data Points

Replit Agent builds and deploys applications using natural language prompts: DeepSeek-V2.5’s open model blends coding and chat

NVIDIA’s Blackwell chips impress on hardware tests, fine-tuned versions of Llama 3.1 add reflection, most AI jailbreaks may not amount to much, new architecture extends context windows to 100M tokens.

September 9, 20243 min read
LAION cleans up its image dataset: Plus, OLMoE outcompetes smaller open models
Data Points

LAION cleans up its image dataset: Plus, OLMoE outcompetes smaller open models

AlphaProteo, a DeepMind system that designs novel proteins, updates and price drops for Command-R and Command-R+, Anthropic shows off easy software projects in Claude, YouTube builds system to detect synthetic music and faces.

September 6, 20243 min read

Subscribe to Data Points

Your accelerated guide to AI news and research