china
Z.ai changed only the post-training, and the model learned to find exploits News
Z.ai released GLM-5.3 on August 14 using the same base model as GLM-5.2, with every gain coming from post-training, and the largest jump was in finding and exploiting software vulnerabilities.
Qwen3.8-27B shares its predecessor's bones, but not its contract News
Alibaba's Qwen3.8-27B shipped with the same coarse architecture as Qwen3.6-27B, prompting accusations it was a relabel with knowledge stripped out, but the published comparison shows knowledge scores flat or slightly up.
DeepSeek starts charging rush-hour prices on August 17 News
DeepSeek is replacing flat API pricing with peak and off-peak rates on August 17, and the steepest change hits cached input on its Pro model, which goes up twelvefold during Beijing business hours.
Chinese models passed American ones in OpenRouter traffic in June News
OpenRouter's own analysis dates the crossover where Chinese models overtook American ones in token share to early June 2026, driven by DeepSeek V4 Flash taking 70 percent of DeepSeek's agentic traffic.
China's biggest memory maker is booked through 2027 News
ChangXin Memory Technologies has reportedly sold out its DRAM output through the end of 2027 as PC brands rushed to secure supply, and consumer memory prices have stayed near their highs since.
DeepSeek warns of a significant API price rise, five days after being called 100 times cheaper News
DeepSeek added a footnote to its official pricing page warning that it plans to raise API prices significantly in the near future with no figure and no date attached, five days after an independent benchmark study priced its model at roughly 100 times less per task than Western frontier models.
Ant published Ling-3.0-flash's weights under plain MIT, with no rider attached News
Ant Group's InclusionAI lab published the full weights for its 124-billion-parameter Ling-3.0-flash model on Hugging Face this week under an unmodified MIT license, with no acceptable-use policy, revenue threshold, or branding requirement anywhere in the release.
The "2x GB200 bandwidth" Chinese chip claim is a 2027 projection, and the arithmetic gives 1.67x News
A widely shared claim that a Chinese accelerator delivers twice the memory bandwidth of NVIDIA's GB200 traces to a roadmap part expected in early 2027, compared 64-at-a-time against a full NVIDIA rack, and the published numbers work out to 1.67 times at rack level while the single chip lands below a shipping GB200.
China did not give away free models. It built a governance body. News
Reports that China offered free AI models to the Global South at a Geneva summit describe a discussion session; the concrete instrument came nine days later in Shanghai with the founding of an intergovernmental AI cooperation organisation.
Moonshot releases Kimi K3: a 2.8-trillion-parameter open-weight model, 1.56 terabytes on disk News
Moonshot AI published the full downloadable weights for Kimi K3, a 2.8-trillion-parameter model that uses only 104 billion parameters per token, handles text and images, and reads just over a million tokens of context.
Lobbying Filings Show Anthropic Named Distillation and Export Controls. OpenAI's Did Not. News
After the New York Times reported that both labs privately pressed Washington over Chinese open-weight models, their own second-quarter lobbying disclosures tell sharply different stories about what each one admits to working on.
Kimi K3's Open Weights Are Still a Countdown, Not a Download News
On the eve of its promised release, Moonshot AI's Kimi K3 page on Hugging Face is a timer with no weights, no license file, and no technical report behind it.
DeepSeek paused its funding round after a leaked meeting transcript went viral News
Bloomberg reports DeepSeek told prospective backers it would not sign expected agreements for now, a suspension its sources attribute in part to viral posts about a leaked investor-meeting transcript attributed to founder Liang Wenfeng.
A Huawei-chip training report shows what leaving CUDA actually costs News
SLAI's technical report documents full-parameter post-training of a DeepSeek-V4 model on Huawei Ascend hardware, and the work list -- rebuilt collectives, converted checkpoints, hand-written kernels -- is the real measure of chip independence.
White House Says Moonshot Distilled Anthropic's Fable to Build Kimi K3 News
OSTP Director Michael Kratsios said the US government has information that Moonshot AI distilled Anthropic's Fable model to build Kimi K3, but no supporting evidence has been made public.
Axios: U.S. Officials Revive an Effort to Discourage Chinese Open-Weight AI News
Axios reports that internal U.S. efforts to restrict Chinese open-weight AI models have revived after Kimi's rise, but no ban, rule, or executive order has been announced.
Alibaba Ships Qwen3.6 as Open Weights, Betting on Efficiency Over Size News
Alibaba released its Qwen3.6 line under Apache 2.0, led by a 35-billion-parameter mixture-of-experts model that activates only about 3 billion parameters per token and targets agentic coding.
Xi Jinping Pitches Open-Source AI and Launches a Global AI Body in Shanghai News
At the 2026 World AI Conference, Xi Jinping urged the world to 'encourage open source, openness, collaboration and sharing' and announced a new China-led World AI Cooperation Organization headquartered in Shanghai.
Kimi K3, a Frontier Chinese Model, Triggers a Global Chip Selloff News
Moonshot AI's Kimi K3 took the top spot on a frontend-coding leaderboard and helped push AI and semiconductor stocks down for a third straight day, with traders calling it a 'Kimi moment.'
China Bans AI Romantic Companions for Minors in a World-First Rule News
Five Chinese agencies enacted the world's first dedicated regulation of emotionally interactive AI, banning virtual romantic partners for minors and pushing platforms like ByteDance and Alibaba to pull companion features.
What Ring-2.6-1T's model card actually says News
Ant Group's openly downloadable trillion-parameter model is real and MIT-licensed, but its benchmark claims are vendor-supplied and measured against a previous generation of rivals -- not the current frontier.
DeepSeek is designing its own AI chip -- and raising outside money for the first time News
Chinese AI startup DeepSeek is developing its own chip aimed at running trained models rather than training them, and is simultaneously raising its first-ever outside capital -- about $7 billion at a $52-59 billion valuation.
A 2025 Nobel chemist is leaving the US to run an AI materials lab in China News
Omar Yaghi, who won the 2025 Nobel Prize in Chemistry, has taken a full-time position at Tsinghua University in Beijing to lead a new AI-assisted materials-discovery institute, citing US grant cuts and a lack of American engagement with AI.
Meituan open-sources LongCat-2.0, a trillion-parameter model it says was trained end-to-end on Chinese chips News
Meituan released LongCat-2.0, a 1.6-trillion-parameter open-weight (MIT) model that ran anonymously as 'Owl Alpha' for two months and was, the company says, both trained and served entirely on domestic Chinese AI ASICs with no Nvidia GPUs.
Tencent open-sources Hy3, a lean mixture-of-experts model that punches above its weight News
Tencent released Hy3 under the permissive Apache 2.0 license: a mixture-of-experts model with 295 billion total but only 21 billion active parameters and a 256K context window, which the company says competes with models five times its size.
Chinese open models now handle a third of US enterprise AI traffic News
US companies now route more than 30% of their AI tokens through Chinese open-weight models like DeepSeek and GLM-5.2 every week since February, peaking near 46%, up from an 11% average the year before, according to CNBC's analysis of OpenRouter data.
A startup router is giving away 100 million tokens of Kimi, MiniMax and GLM News
API aggregator Dahl Inference is handing out 100 million free tokens across top open-weight Chinese models like Kimi K2.6 and MiniMax M2.7 - not a price cut from the labs themselves, but a router burning money to win users amid a glut of cheap compute.
A $4-per-million open model is coming for the frontier's 90% margin News
GLM-5.2, an open-weights model priced at under a fifth of Opus, scores as the top open model and 4th overall - and a widely-shared essay argues it is the first real threat to frontier labs' ~90% inference margins.
China's GLM-5.2 Ships as the Top Open-Weight Model, Under MIT License News
Z.ai released GLM-5.2, a 753-billion-parameter model, as open weights under an MIT license, and an independent index ranks it the strongest open-weight model available, close behind the leading closed models at a fraction of the price.
An open model from China beat Claude on a security test -- at a sixth of the cost News
Semgrep ran GLM 5.2 against Claude on a narrow vulnerability-finding task and the free, open-weight model came out ahead for far less money.
Are closed AI models overpriced luxury goods? News
An essay argues open-weight models now undercut the big closed AIs by huge margins, and that 'China fears' are being used to protect those prices.
Anthropic says Alibaba ran the biggest 'copy Claude' campaign yet News
Anthropic told U.S. senators that Alibaba's Qwen team quietly milked Claude for its best skills. Alibaba says nothing back, and the whole fight may be as much about price as theft.
The Model Ban Is Quietly Redrawing the AI Map News
Two weeks after the US pulled its top models off the market, a Chinese open model sits atop the global download charts and the community is busy rebuilding the banned capability in the open.
A Free Model That Splits Your Work Across 300 Helpers News
Moonshot AI's Kimi K2.6 is a frontier-grade model anyone can download, and its headline trick is fanning a single job out to hundreds of helpers working in parallel.
Ring-2.6-1T Tool
Ant Group's trillion-parameter mixture-of-experts reasoning model, activating roughly 63 billion parameters per token, with 128K context extendable to 256K. All checkpoints openly downloadable under the MIT license, with high and xhigh reasoning-effort settings that trade depth against speed and cost. Benchmark claims are vendor-supplied and measured against a previous generation of rivals.
Kimi K3 Tool
Moonshot AI's 2.8-trillion-parameter flagship with a 1M-token context window, tuned for agentic coding and knowledge work; it topped a frontend-coding leaderboard. Usable now via kimi.com chat and an OpenAI-compatible API, with open weights due July 27.