Google Research introduces Science One autonomous research framework
Google Research published the Science One Framework, a verifiable autonomous research framework built on Chain-of-Evidence. It is presented as work in general science.
Linked and quoted from the labs, the press and the people using them. Times in UTC.
Every development we picked up, including smaller stories beyond the labs. Newest first, times in UTC.
Google Research published the Science One Framework, a verifiable autonomous research framework built on Chain-of-Evidence. It is presented as work in general science.
Google released Gemini Robotics ER 2, a model for robotic applications. It adds video understanding, tool orchestration and multi-robot collaboration.
Anthropic reported degraded performance on Claude Opus 4.8. The incident, lasting about 41 minutes, has been resolved.
Anthropic reported elevated errors across many models. The incident, rated major impact, has been resolved after about 291 minutes.
Anthropic reported elevated errors across all models, rated critical. The issue has been resolved.
Google launched Lyria 3.5 in Google Flow Music. The update brings advances in musicality, lyrics, vocals, and creative control.
Pangram has raised $9 million to scale its AI detection software. The startup also released a new AI text detection model, Pangram 4, and an AI image detection model in research preview.
An OpenAI staff member said usage limits were reset for all ChatGPT Work and Codex users. They also gave an update on GPT-5.6 Sol usage limits after reports that Sol consumed Codex limits faster.
Fish Audio raised $52 million in seed funding led by Coreline Ventures and Capital Today. The startup aims to make voice the default interface for AI models.
Google DeepMind announced Gemini Robotics 2, a model that brings whole body intelligence to robots. The announcement was published on deepmind.google.
MiniMax released MiniMax-H3 on Hugging Face, a model supporting text-to-video, image-to-video, image-text-to-video and video-to-video generation. The repository lists 3,664,216 downloads and 5,623 likes.
OpenAI staff said usage limits have been reset for all paid users of Codex and ChatGPT Work. The reset follows fast adoption of ChatGPT Work.
Microsoft introduced MAI-Cyber-1-Flash, a compact security model that scores 96 percent on the CyberGym benchmark within its MDASH multi-agent system. Microsoft says costs drop about 50 percent versus frontier models since only tough cases go to GPT-5.4.
Microsoft released a Fantasy Premier League Companion tool for managers. The post announcing it appeared on Microsoft's Source blog.
Anthropic reported elevated errors affecting Claude Opus 5 and Haiku 4.5. The incident was resolved after 57 minutes, with errors returning to baseline.
Enigma raised a $71 million seed round led by Index Ventures and Ribbit Capital, with participation from Conviction Partners. The company aims to make controlling a robot as easy as adjusting the volume.
Anthropic reported elevated errors on Claude Opus 5. The incident lasted about 63 minutes and errors returned to baseline as of 4:47 PST.
Anthropic reported elevated errors on Claude Opus 5 starting at 2:03 PST. The incident was resolved after 48 minutes when errors returned to baseline.
Cohere introduced North Automations, a feature for intelligent workflow orchestration. The announcement was published on Cohere's official site.
Anthropic reported elevated errors for Opus 5, marked as a major incident. The incident has been resolved after about 87 minutes.
Anthropic reported elevated errors affecting Claude Fable 5, Claude Sonnet 5, Claude Haiku 4.5 and other models. The incident, rated major, lasted 34 minutes and has been resolved.
Anthropic reported elevated errors affecting Mythos 5, Fable 5, Opus 5 and Claude Haiku 4.5. The incident, rated major, lasted 64 minutes and has been resolved.
Anthropic's Claude Opus 5 leads the Artificial Analysis Intelligence Index with 61 points, ahead of Claude Fable 5 and GPT-5.6 Sol. It scores highest in analytical quality and coding and costs up to half as much as Fable 5 at lower reasoning tiers.