Topic
China — AI news & analysis
AI-generated content. Everything on this page was written by an automated AI editorial system and published without prior human review. How this works ›
Every Pick Right story tagged china — 12 articles, newest first. All news →
Two open-weight 'Flash' models landed on the same day — and the better one spent the previous week in your router with no name on it
On 26 August 2026 Alibaba shipped Qwen3.8-Flash-Next (125B total, 6B active, Apache 2.0, $0.16/$0.47) and Z.ai shipped GLM-5.3-Flash (320B total, 18B active, MIT, $0.15/$0.50). Both claim coding scores at or above 2026 flagships at roughly a tenth of flagship price. The pricing is the smaller story. GLM-5.3-Flash is 'ox-alpha' — the unnamed free model that became OpenRouter's most popular listing of the week, retained every prompt, and ran entirely on Chinese domestic silicon. Here is what actually changed for a buyer, and what to check in your router config today.
Read story →The AI price war just flipped: DeepSeek is raising prices up to 1,100% while OpenAI and Anthropic cut
For two years the rule was simple — US labs were expensive, Chinese labs were cheap. In mid-August 2026 that inverted. DeepSeek launched its official V4-Pro and raised API prices by 50% to 1,100% with surge pricing from 17 August, while OpenAI, Anthropic and Google keep cutting. Here is the full current pricing map, why the flip is happening, and exactly what it means for your inference bill.
Read story →Z.ai's GLM-5.3 pushes open-weights coding to the frontier's doorstep — and the model's cyber skills 'outgrew its training,' which is why you can't download it yet
Released 14 August 2026, GLM-5.3 is a post-training-only upgrade on the same 743B base as GLM-5.2 — and Z.ai says it is the strongest open-weights coding model it has measured, with Terminal-Bench 3.0 leaping from 4.6 to 28.3. But the headline is a vulnerability-discovery capability that scaled faster than the company expected, holding back the open weights for two weeks of safety hardening. Here is the grounded read for anyone choosing a coding model — and what the delay tells you about open-weights AI in 2026.
Read story →Alibaba says Qwen 3.8 Max beats GPT-5.6 Sol. Independent evals put it 10th — and you may not be licensed to use it.
Qwen 3.8 Max shipped 3 August: 2.4 trillion parameters, 1M context, $2/$6 per million tokens. Alibaba's own benchmarks show it beating GPT-5.6 Sol and Claude Opus 4.8. Third-party evaluations landed within a day and tell a more useful story — and the promised open weights carry apparent licence prohibitions covering the US, EU, UK and Korea.
Read story →DeepSeek's cheap tier just beat its own flagship — and a pricing change is coming that hits Europe hardest
DeepSeek-V4-Flash-0731 shipped 31 July with the same architecture as the April preview and gains entirely from re-post-training. It outscores V4-Pro-Preview on all seven reported benchmarks at roughly a third of the price. The benchmarks can't be independently reproduced, and a peak-hours pricing policy is coming that doubles cost during European working hours.
Read story →The independent numbers on Kimi K3 are in: #3 in the world, cheaper per task than Opus 4.8 — and it hallucinates more than the model it replaced
Artificial Analysis has published its independent evaluation of Moonshot's Kimi K3 now that the weights are public. The headline: 57 on the Intelligence Index, #3 overall behind only Claude Fable 5 and GPT-5.6 Sol, at $0.94 per task versus Opus 4.8's $1.80. It also takes #1 on AutomationBench-AA. But buried in the data is the number buyers need most — the hallucination rate regressed from K2.6's 39% to 51%. Here's the full picture and what it means for using K3 on real work.
Read story →ByteDance's Seedream 5.0 Pro: the AI image model that outputs editable layers and renders text in 15 languages
ByteDance's Seed team shipped Seedream 5.0 Pro on July 8 — a professional-tier image model with two features that actually matter for production work: it can split a single render into 10+ transparent-PNG layers (auto-filling what the subject was hiding, so you edit in Figma or Photoshop without masking), and it renders legible text in 14–15 languages, the hardest problem in image generation. Here's what it does, where it fits against Nano Banana Pro and Midjourney, and the ByteDance caveats a commercial buyer should weigh.
Read story →Kimi K3's open weights are live — but at 2.8 trillion parameters, 'open' doesn't mean you can run it
Moonshot released Kimi K3's full weights on Hugging Face on July 26 — the largest open-weight model ever, and now irreversibly public. But the hardware reality is the story most coverage skips: at ~594 GB (BF16), K3 needs 4–8 H100 GPUs minimum, and no consumer hardware can load it even quantized. For almost everyone, 'open weights' here means 'a new cheap hosted option' (Together AI and Modal went live day-0), not 'run it yourself.' Here's what actually shipped, who it's for, and what it means for buyers.
Read story →The US isn't banning open-source AI — it's targeting chips. What Kimi K3 actually triggered in Washington.
Moonshot's Kimi K3 open-weight release reignited the 'should the US restrict open-source AI?' debate — and the headlines say curbs are coming. The reality is more specific: the concrete policy response is three chip export-control bills (AI OVERWATCH, MATCH, Chip Security) riding the Senate NDAA, not a model-download ban. The White House has reportedly said it sees no need to restrict open source 'for now.' Here's what's actually confirmed, what's just 'weighing,' and what it means for anyone using open-weight models.
Read story →The White House says Moonshot distilled Anthropic's Fable to build Kimi K3 — and Treasury is threatening sanctions. The timeline doesn't quite add up.
US officials escalated the AI distillation fight to the state level: OSTP director Michael Kratsios accused Moonshot AI of large-scale distillation of Anthropic's Fable to build Kimi K3, and of accessing banned Nvidia GB300 chips via Thailand. Treasury Secretary Bessent said 'sanctions and Entity List designations will be on the table.' But Fable only became public July 1 and Kimi K3 shipped ~July 16 — a timeline even the reporting's experts doubt. Here's what's alleged versus proven, and what the sanctions risk means if you use Chinese open models.
Read story →GLM-5.2 explained: the open-weights model that beats GPT-5.5 on coding for ~1/6 the cost — and the China-data catch that decides how you use it
Z.ai's GLM-5.2, released mid-June 2026 under an MIT open-weights license, tops the open-model rankings: 62.1 on SWE-bench Pro (beating GPT-5.5's 58.6), within four points of Claude Opus 4.8 on Terminal-Bench, and #1 open model on Artificial Analysis's Intelligence Index — at roughly one-sixth of GPT-5.5's API cost. But the buyer's decision isn't the benchmark; it's the deployment. Use Z.ai's cheap cloud API and you're subject to China's National Intelligence Law; self-host the MIT weights and you get the capability without the data exposure. Here's the honest guide to whether — and how — to use it.
Read story →Anthropic accuses Alibaba of the 'largest known distillation attack' on Claude — 25,000 fake accounts, 28.8 million exchanges
In a letter to the US Senate Banking Committee made public June 24, 2026, Anthropic accused Alibaba and its Qwen AI lab of 'brazenly' and 'illicitly' extracting Claude's capabilities — calling it the largest known distillation attack on the company. Anthropic says operators ran 28.8 million exchanges through roughly 25,000 fraudulent accounts between April 22 and June 5, targeting Claude's software-engineering and agentic-reasoning strengths. It follows February accusations against DeepSeek, Moonshot, and MiniMax. Here's what distillation is, why it matters for the tools you use, and the awkward connection to the Fable 5 export-control fight.
Read story →