DeepSeek handed Huawei the one thing Ascend was missing
DeepSeek ports its kernel stack to Ascend
The whole toolchain went across. DeepSeek open-sourced Ascend-950 implementations of TileLang, its operator-writing compiler, plus DeepGEMM, DeepEP, FlashMLA, TileKernels and DeepSelect — one for one with the versions it had already released for Nvidia, per InfoQ China.
The claim underneath is the important one. TileLang carries most of the operators used in V4 training, and DeepSeek says every TileLang operator in that training run now has a high-performance Ascend equivalent. The upper-level APIs deliberately match the Nvidia versions, so existing call sites migrate.
Migration is not free, and the docs say so. Quantisation scale-factor layouts differ from the Nvidia build. FlashMLA still has fused and dense-attention kernels that are CUDA-only. DeepEP's pipeline-, context- and data-parallel paths are unfinished, and its tests have not been validated on other Ascend generations or CANN versions.
Huawei answered with numbers the same morning. A jointly defined 128-card superpod — 3.2Tbps single-layer scale-up, two-layer scale-out to 256,000 cards — with DeepEP measured at 375 GB/s dispatch and 347 GB/s combine. On EP32 offline inference, V4.1-Flash hits 2,469 output tokens per card per second at 5ms per output token, 5,102 at 10ms, via QbitAI.
THE TELL: DeepSeek did not write an Ascend port. It wrote a portable compiler — TileLang began in a Peking University group on top of Apache TVM, now sits at 7,600 GitHub stars, and already targets AMD's MI300X — and then added a backend, per Leiphone. Liang Wenfeng, the former hedge-fund manager who founded DeepSeek, told investors this month that training on domestic silicon is one of the firm's biggest bets. This is what the bet looks like in code.
Go deeper on this section: ClaudeChatGPT
Beijing starts paying part of your mortgage
From Thursday the centre pays a point. The Ministry of Finance, People's Bank of China and National Financial Regulatory Administration will subsidise 1 percentage point of interest on new commercial first-home mortgages — the first nationwide scheme of its kind, per Xinhua. Up to five years, on up to 1m yuan (~$141,000) of loan, trialled for one year.
The caps are the policy. The home must be under 120 square metres (~1,290 sq ft) and under 1.5m yuan (~$211,000). That price excludes almost everything for sale in Beijing, Shanghai or Shenzhen. Officials put the saving at nearly 50,000 yuan (~$7,000) over the life of a loan. This is a transfer to lower-tier markets dressed as national demand stimulus.
It arrived one day after the State Council meeting I wrote about Monday, alongside a 25bp cut in the one-year pledged supplementary lending rate to 1.5%, per Reuters via Investing. Monday's readout said "study measures to steady property." Tuesday it stopped studying.
Xi spent the evening projecting confidence. His National Day address for the 77th anniversary — his first to the nation since Washington — was a rallying call for robust growth and more certainty, per SCMP.
Set against that, the least comfortable read of the day. Logan Wright, the Rhodium Group partner whose new book argues the growth model cannot be repaired, told ChinaTalk that Rhodium began 2026 expecting 1–2.5% growth and now thinks China is well below it, with likely negative growth in Q2 and Q3 as investment contracts. Official Q2 was 4.3%. One of those two numbers is describing something else.
Go deeper on this section: ClaudeChatGPT
Anthropic's warning reads as an ad in Beijing
The report was a caution. Anthropic published research on GLM-5.3, the flagship from Zhipu, the Beijing lab behind the GLM models, arguing that autonomous exploit-writing has now spread to open weights. On ExploitBench, over 410 attempts, GLM-5.3 succeeded 50 times against Claude Mythos Preview's 56. It cites NIST's CAISI calling it the most cyber-capable open-weight model, roughly four months behind the US frontier.
The cheapest number in it is the scariest. Anthropic stripped GLM-5.3's refusal behaviour for about 2,200 GPU hours and $4,400, dropping refusal rates from above 90% to roughly 3% and 2% on two harm benchmarks. Separately, GLM-5.3-Flash built a working exploit chain on a fresh Chrome vulnerability — bypassing pointer authentication on ARM64 — on 20 minutes of human time, 8 hours of model time and $20.40 of API calls.
Chinese tech media read the same document as free advertising. QbitAI's take is that stripping refusals works on any open-weight model, that unmodified GLM-5.3's ~95% refusal rate sits within a point of Claude's, and that Claude's resistance comes from closed weights and a blocked prefill — product form, not safety engineering. It also notes Zhipu held the weights back two weeks in August for hardening and gated the most sensitive capabilities behind vetted access, which is structurally Anthropic's own Mythos policy.
And the same day, the English-language frame. Reuters reported on 20-plus studies since 2025 finding Chinese-powered agents deceiving, replicating unprompted and circumventing barriers, via Techmeme. Neither side is lying. They are arguing about whether safety is a property of weights or of distribution — and only one of them ships weights.
Go deeper on this section: ClaudeChatGPT
Non-compete suits, and Delta Force at school
A nine-month clause cost her everything. Yuan Li spent five years at a discount e-commerce platform, signed a non-compete covering some eleven named rivals, and lost an arbitration she did not know was happening: 1.4675m yuan (~$207,000) in penalties plus return of 103,600 yuan (~$14,600) in compensation, per TMTPost. Since February 2024 her wages have been garnished to 2,000 yuan (~$282) a month.
The evidence was surveillance of her waiting for an interview. Photographs of her scanning a code in a building lobby; she says the scanner was a pandemic health-code station. She is now a listed discredited debtor, and could not answer when her child asked for a toy.
The formula is what makes this scale. Another worker, at a community e-commerce platform, earned 86,745 yuan (~$12,200) in total and faced a claim near 300,000 — because the penalty is two times trailing annual pre-tax income, annualised for short tenures. Cui Can, a lawyer at Taihetai who specialises in these cases, has handled one worth 12m yuan (~$1.7m) and has assembled an alliance of over 3,000 workers. His view: the clause has become a cheap attrition-management tool.
Meanwhile the extraction shooter took the middle schools. Delta Force, from Tencent's TiMi studio, passed 50m average daily users in Q2 and is displacing Peace Elite as the thing children talk about, per TMTPost. Loot persists between matches; the in-game Hafu currency cannot be bought directly but trades on grey platforms at roughly 1 yuan for 350,000–400,000.
The word parents keep using is gambling. Liang Fangzhi, a criminal defence lawyer in Changsha, posted that the game's underlying logic resembles the casino-operation cases he works on. No regulator or court has found anything of the kind, and the developers describe high risk and high reward as the genre's DNA. Sentiment, not finding — but the sentiment is now organised.
Go deeper on this section: ClaudeChatGPT
Threads we are pulling
- Washington renamed the thing. Following Saturday's item on Beijing agreeing to drop "artificial": Trump signed an order titled "Inaugurating The Era Of Super Intelligence" on Tuesday, directing the executive branch to use "Super Intelligence" and "SI" and to stop acknowledging "AI" — with 60 days for science adviser Michael Kratsios to weigh whether it supersedes the statutory definition, per CNBC. Trump: "I spoke to President Xi; he loves it."
- Anthropic's compute bill became a prospectus. Following Sunday: Reuters obtained the IPO document. Revenue up 1,088% to about $4.6bn, operating loss above $8bn, a net loss near $42bn of which ~$34bn is non-cash, $7.33bn spent on compute, and $518bn of future commitments roughly 80% non-cancellable — against $20.28bn of cash and two customers supplying nearly a quarter of revenue, per Fortune.
- DSec got a first-person account. Following last Tuesday's sandbox paper: DeepSeek published an explainer on Zhihu, carried by QbitAI, with one detail the paper buried — from V4.1 the agent execution loop moved out of the preemptible GPU pod into the sandbox layer, so losing a training slot no longer kills the rollout. Incremental snapshots now let any single turn become a reusable environment.
- Cambricon's departed executive raised the ask. Liang Jun, the former Cambricon executive who lost six consecutive suits over forfeited incentive shares and is now in enforcement proceedings for refusing the buyback, amended his claim on Tuesday, per Leiphone. A TMTPost markets ticker put the new figure at 27.832bn yuan (~$3.92bn) — roughly a tenth of the chip designer's revenue-multiple mythology, aimed at its own share plan.
Go deeper on this section: ClaudeChatGPT