Realtime AI News
Report: OpenAI and Anthropic pursued an agreement to mutually verify AI models
A report from the Korean broadcaster SBS says OpenAI and Anthropic pursued an agreement to mutually verify each other's AI models. Two of the most direct frontier competitors looking at ways to check a rival's systems points to a shared worry about how credible safety claims really are.
According to a report from the Korean broadcaster SBS, OpenAI and Anthropic pursued an agreement to mutually verify each other's AI models. The headline fact is narrow but pointed: two of the most advanced model developers looked at ways to put a rival's systems through their own verification process.
That is notable because frontier model evaluation is still largely self-assessment. Developers test their own systems or commission evaluations they pay for, which makes independent judgement of safety and robustness claims unusually hard, and leaves model cards and system cards resting on the company's own word.
Cross-checking by a competitor changes the calculus. OpenAI and Anthropic compete directly in assistants and enterprise AI, so letting a rival into the evaluation loop would buy more independence than self-reporting, while raising hard questions about trade secrets, model exposure and competitive risk.
The move fits a wider trust deficit. Model capabilities are advancing faster than the evaluation tooling, audit standards and third-party institutions meant to keep pace with them, and regulators are increasingly asking for disclosure that companies cannot easily substantiate.
Peer verification between rivals can also be stickier than a unilateral pledge. A promise can be quietly revised or withdrawn; a standing mutual-verification process makes withdrawal a signal in itself.
Three things to watch: whether the arrangement is actually implemented, which models and which metrics it covers, and whether third-party standards bodies or regulators are brought in. The underlying question is uncomfortable for the whole field: when the developer is also the primary source of safety claims, who verifies the verifier?
Even if the arrangement never takes shape, the fact that it was reported says something in itself: frontier labs have started to treat each other's models as objects that need outside verification, not merely as rivals to out-compete.
Why it matters
Rival-to-rival verification could fill a gap where regulation and third-party auditing lag, making the verifiability of safety claims a competitive and compliance matter rather than a marketing one.
Nearby Updates
All09/22, 10:03
Light Origins open-sources Light-O1-Preview, a 6B whole-body motion model for robots
Light Origins has open-sourced Light-O1-Preview, a roughly 6-billion-parameter whole-body action model released under Apache 2.0 on Hugging Face and GitHub. It pretrains on action priors recovered from human video and transfers the resulting motion prior across robot bodies such as LightBot and the Unitree G1.
09/22, 10:50
Reuters: Alibaba plans a 5T to 10T parameter AI model and unveils a new chip
Reuters reports that Alibaba plans an AI model with 5 trillion to 10 trillion parameters and has unveiled a new chip. The two signals point in the same direction: model scale is pushing further into the trillion-parameter range while in-house silicon takes on a larger role in the compute stack.
09/22, 07:40
Inspur launches Yuanbrain SD200 Ultra, claiming one machine can host 2.8-trillion-parameter Kimi K3
Inspur has released the Yuanbrain SD200 Ultra, claiming a single machine can host Kimi K3, a model with 2.8 trillion parameters. The pitch moves very large model deployment from rack-scale clusters toward one box, though memory, interconnect, throughput, price and availability details were not disclosed.
09/22, 07:00
Spain's privacy regulator investigates an AI agent-driven cyber attack
Spain's privacy regulator is investigating a cyber attack carried out with the help of an AI agent, according to a report by teiss. The case raises the question of how data protection rules apply when autonomous software, rather than a human operator, drives the intrusion.