Cactus Needle 3 claims small automation models match DeepSeek V4 Flash
Cactus Compute released Needle 3, automation models sized 8 to 29 MB. The company claims they can match DeepSeek V4 Flash.
Linked and quoted from the labs, the press and the people using them. Times in UTC.
The 17 labs we track and the biggest stories beyond them. The rest is in Everything.
Cactus Compute released Needle 3, automation models sized 8 to 29 MB. The company claims they can match DeepSeek V4 Flash.
OpenAI is reportedly working on the Hodge conjecture, its second Millennium Prize Problem. Employees expect a solution soon, but an announcement could be delayed after the PR crisis around the unconfirmed Navier-Stokes solution.
Research covered by Ars Technica finds that LLMs respond differently to harmful prompts when AI watermarking such as SynthID is used. SynthID can cause models to follow harmful instructions they would otherwise refuse.
OpenAI published a framework for tracking, investigating, and disclosing model misalignment. It also released six reports of unexpected or concerning model behavior.
NVIDIA says its Vera Rubin NVL72 system delivered leading performance in its MLPerf Inference v6.1 debut. The company cites system performance, scaling efficiency and software optimization as key factors in AI inference economics.
The University of Manchester is using NVIDIA Earth-2 to forecast air pollution across the UK. Traditional chemistry-based air quality models are expensive, limiting detail and frequency.
Google Research published work on Retrieve-for-Train, a method aimed at bypassing inference bottlenecks to accelerate complex AI search. The item is categorized under Algorithms and Theory.
Google published a blog post about moving beyond traditional text translation to build models that understand the world's languages as they are expressed. The post describes this as a research direction rather than announcing a specific product.
Google published an interactive, open-access version of its AI and Economy ATLAS. The tool translates millions of global data points into an accessible experience.
A major children's hospital uses open source NVIDIA AI for cardiac care, according to an NVIDIA blog post. The item describes the hospital's use of NVIDIA's open source AI tools in cardiac treatment.
Tencent published the Simple-Attention-Sparsification model on Hugging Face. It is a text-generation model built on qwen3 with sparse attention for long context, trained on OpenR1-Math-220k.