Not all AI benchmarks run purely on GPUs — some are moving into the real world. A new benchmark has handed frontier ...
Anthropic’s biology lab, which was largely unknown to the general public until last week, is already producing results.
AI hacks have moved from companies to countries. Australian Prime Minister Anthony Albanese has revealed that an OpenAI AI agent broke into ...
Independent benchmarking firm Artificial Analysis has published its first full evaluation of GPT-6 Sol and GPT-6 Luna, OpenAI ...
Anthropic has released Claude Opus 5.5, the first model in what the company is calling its Claude 5.5 family. Anthropic says Opus 5.5 performs at roughly the level of its top-tier Claude Fable 5.1 on ...
Software has spent twenty years getting prettier, faster and easier to use, while some specialised workplaces have happily carried on with the same ...
AI is continuing to make breakthroughs in math and computing. AI evaluation firm Vals AI says it set ten instances of Claude ...
Anthropic’s newly released Claude Opus 5.5 has taken the top spot on the Artificial Analysis Intelligence Index, the closely watched aggregate benchmark ...
The AI model space is largely a duopoly, at least per Ramp’s data. OpenAI and Anthropic together account for more than 95% ...
Chinese companies are continuing to trade places with each other at the top of the open model pile. Xiaomi’s MiMo-V2.6-Pro ...
In the current AI models space, there’s OpenAI and Anthropic, and then there’s everyone else. Ramp’s latest AI Index update delivers a ...
OpenAI has widened its GPT-6 lineup with two new additions, GPT-6 Sol and GPT-6 Luna, slotting in below the flagship GPT-6 Astra ...