Realtime AI News
Report finds China AI developers publish safety tests for just 3.6% of model releases
A new report found that Chinese AI developers publish safety tests for only about 3.6% of their model releases. The figure points to how limited safety transparency remains when models go public.
A new report has found that Chinese AI developers publish safety tests for only about 3.6% of their model releases. In other words, for the vast majority of models that go live, outsiders have little public material showing how the model was assessed for safety.
Safety tests are meant to examine how a model behaves when it is misused, produces harmful content, or poses other risks. Whether those results are published matters to everyone outside the developer — researchers, regulators, and ordinary users — who want an independent view of a model's boundaries.
The 3.6% figure draws attention because risks of misuse and loss of control grow alongside model capability. When releases come with little transparent safety information, outside oversight struggles to take root.
One caveat: publishing test results is not the same as a model being unsafe, and it does not mean no internal testing happened. What the report speaks to is the degree of disclosure, not absolute safety.
For the industry, the finding could fuel more debate over norms for model releases. As expectations around AI safety rise, attaching safety information at launch may shift from a bonus to a baseline requirement.
What to watch next is whether the share changes over time, and whether developers and regulators converge on clearer standards for disclosing safety tests.
Why it matters
The finding puts safety transparency at the point of release in the spotlight, and could push developers and regulators toward clearer disclosure norms.
Nearby Updates
All10/09, 21:04
AIFA launches creator program and AI agent upgrades for aifa.one
AIFA has launched a creator program for its aifa.one platform alongside AI agent upgrades. The announcement combines a push to attract creators with a refresh of the platform's agent capabilities.
10/09, 21:15
Qwen Lists Qwen-Image-2.1-Turbo, a Text-to-Image Model Built on Qwen-Image-2.1
Qwen has published Qwen-Image-2.1-Turbo to its official Hugging Face repository, presented as a fine-tuned variant of Qwen/Qwen-Image-2.1 for text-to-image work. The listing uses the diffusers library and safetensors weights, and had already collected 55 likes when it was captured.
10/09, 22:04
Amperity launches Pér, an AI agent with a full picture of each customer
Customer data platform Amperity has introduced Pér, an AI agent it says can hold a complete picture of every customer and act on it. The launch pairs customer-data understanding with agentic execution aimed at marketing and customer operations.
10/09, 22:52
Meta and Sierra develop Personal Agent Protocol for AI agents
Meta and Sierra are developing Personal Agent Protocol, an open standard governing how personal AI agents interact with businesses on consumers' behalf, with partners including Shopify, Stripe, and Walmart. The protocol leaves access decisions to consumers and lets companies set what agents may do, and the partners plan to publish a v0.1 specification plus payments extensions.