OpenAI introduces MentalHealthBench for AI mental health responses
OpenAI released MentalHealthBench, an expert-informed benchmark for evaluating helpful and safe AI responses in realistic mental health conversations.
Linked and quoted from the labs, the press and the people using them. Times in UTC.
Every development we picked up, including smaller stories beyond the labs. Newest first, times in UTC.
OpenAI released MentalHealthBench, an expert-informed benchmark for evaluating helpful and safe AI responses in realistic mental health conversations.
OpenAI says a new internal model solved more than 100 open math problems after a month of training. The company is backing an independent advisory group at the Institute for Advanced Study but excluded its research pace from the group's advisory role.
OpenAI used its own large language models to help design its Jalapeño chip, according to an IEEE Spectrum report. The chip design work was assisted by the company's LLMs.
OpenAI is reportedly working on the Hodge conjecture, its second Millennium Prize Problem. Employees expect a solution soon, but an announcement could be delayed after the PR crisis around the unconfirmed Navier-Stokes solution.