Anthropic says Opus 5 is nearly immune to prompt injections in its own software. Prompt injection, where an attacker slips past an AI model's instructions through manipulated inputs like hidden text ...
According to Artificial Analysis, Anthropic's Claude Opus 5 is currently the most capable model, achieving an Intelligence Index score of 61, with particular strengths in analytical quality and ...
Read full article about: Hundreds asked ChatGPT for poison and bioweapon recipes and some got step-by-step high school level guides In summer 2025, OpenAI internally flagged GPT-5 as high-risk because ...
Sakana AI has released Fugu Ultra v1.1, an update to the AI router that distributes each query across a pool of publicly available top-tier models. The company claims performance gains of up to 7.9 ...
Read full article about: Claude's voice mode now runs on Anthropic's most capable models across all platforms Anthropic has expanded Claude's voice mode to run on the more powerful Opus and Sonnet ...
OpenAI is planning a portable, screenless smart speaker as an AI companion for the home. The device is meant to feel alive, but Apple's trade secrets lawsuit could delay its launch. OpenAI's ...
The German consortium behind the AI model Soofi S has acknowledged in version 3.0 of its tech report that test questions from the science benchmark GPQA accidentally ...
The second Anthropic Economic Index analyzes how Claude usage is shifting across the economy. One key finding: the longer people use the AI model, the better their results get. That could widen ...
A new study from MIT and the University of Southern California shows that lawsuits filed without a lawyer at US federal courts have nearly doubled since ChatGPT went mainstream. One in five complaints ...
The neuroscientist Jean-Rémi King leads the Brain & AI team in Meta’s AI division. In an interview with The Decoder, he discusses the connection between AI and neuroscience, the challenges of ...
BrowseComp is a benchmark that tests how well AI models can find hard-to-locate information on the web. When Anthropic turned its Claude Opus 4.6 model loose on the benchmark in a multi-agent setup, ...
Anthropic will brief leading finance ministries and central banks on vulnerabilities in the global financial system's cyber defenses that its new AI model Claude Mythos Preview has uncovered.
Some results have been hidden because they may be inaccessible to you
Show inaccessible results