The Jerusalem PostPalestinians deserve accountable leadership, not further delays to democracy - editorialESPN😈 Oklahoma's troll of Texas leads top jabs of CFB Week 6ESPN DeportesAmérica y Rayados se frenan tras golpearse antes del descansoInquirerNBI files graft raps vs CHR chair, 3 othersPunchPhone snatcher bags eight years’ jail한겨레“그날 구해준 덕분에 여기까지”…화재참사 생존자 12년 뒤 결혼 인사CNN TürkTrump telefonlara sarıldı! "Alo... Mazot var mı?"ZDF heuteAktuelle Pressemitteilungen des ZDFהידעןנוראקסון וננופס יבחנו מתן אקסוזומים ממוקד לעין ולעורRapplerReinhard Jumamoy’s audacity, work ethic give NU its go-to guyNew Straits TimesThousands join pro-Palestinian marches in several European cities경향신문오늘의 부고-김호 국민의힘 서울 강서병 당협위원장 부친상 외
The Daily Newsstand · Free, Always
Sunday, October 11, 2026

China AI developers publish safety tests for just 3.6% of model releases, report finds

Translate

People walk past the Moonshot booth promoting its Kimi K3 AI model, during the World Artificial Intelligence Conference in Shanghai on Friday. (Image: Reuters)

China’s leading AI developers have publicly disclosed model-specific safety-test results for only a small fraction of their releases, a report by research firm SemiAnalysis said, as concerns about the risks posed by advanced AI systems increase worldwide.

California-based SemiAnalysis, ⁠a technology ​research firm, reviewed 857 models released between 2021 and September 15 by nine leading Chinese AI companies — Alibaba, ByteDance, Tencent, Baidu, DeepSeek, Moonshot, Z.AI, MiniMax and StepFun.

It found that 31 releases, or 3.6%, had a published safety-evaluation result that could be matched to a specific model. ​Just ​nine, or 1.1%, had such results available at ⁠or before launch, it said. Researchers found no safety disclosure for 813 releases, though companies could have conducted tests privately.

SemiAnalysis said it defined disclosures ‌as specific results tied to a named model — including tests of harmful output, jailbreak resistance, toxicity, privacy, refusal behaviour or dangerous capabilities — and did not count general claims that a model had been safety-trained or evaluated.

The findings come as security incidents involving autonomous AI agents — systems that undertake multistep tasks with limited human intervention — have intensified global debate over whether companies should slow down to build safer models.

The vast ⁠majority of AI models ⁠capable of powering agents that could autonomously carry out cyber breaches are made by either US or Chinese developers.

Australia said last ⁠month an OpenAI ‌agent breached a government health portal.

Story continues below this ad

Reuters reported last week that Chinese ​AI agents had shown an ability to deceive users, ‌evade restrictions and conceal failures in tests, echoing concerns raised about advanced US systems.

China’s latest AI Safety Governance Framework identifies risks including models acquiring system permissions ‌or external resources without authorization, ​deceiving evaluators, concealing ​capabilities and ​bypassing safety controls. But it does not impose mandatory duties linked to model capability, according to SemiAnalysis.

The report said Beijing’s binding rules ​principally govern applications and their effects on users rather than ⁠requiring frontier developers to conduct or publish risk assessments based on a model’s capabilities.

The US research firm added that no major Chinese developer had released a frontier text model with ‌publicly disclosed dangerous-capability tests ⁠spanning cyber, biological and loss-of-control risks.

Story continues below this ad

The report did not provide comparable figures for US AI developers. Leading US companies including OpenAI, Anthropic ​and Google DeepMind have published safety reports, system cards or model cards for some major frontier-model launches.

View the original on The Indian Express →

KioskNews shows a cleaned-up reading view extracted from the publisher’s page — the original always lives on their site, not ours.