Skip to content
World clockEU--:--UK--:--USA--:--CN--:--PLDEFRIT中文EN

portal about AI and technologyevents · analysis · interviews · technical background

Search
LIVE
›

AI & models

238 texts · 14/14
Filings near 1,200: two tracks in China's AI governance

Filings near 1,200: two tracks in China's AI governance

China's Cyberspace Administration says 1,112 generative AI services had completed filing as of August 31, 2026. The same week, the AI Safety Governance Framework 3.0 came out during Cybersecurity Week.

AI & models25 September 20266 min readSources 3

Latest

11 texts
Xiaomi trained a model live and spent 3.5 million dollars in six days

Xiaomi trained a model live and spent 3.5 million dollars in six days

MiMo-V2.6-Pro has 1.02 trillion parameters, 42 billion of them active. In an open live training run it scored 46 points on the Artificial Analysis index and took first place among open models.

AI & models23 September 20265 min read
The compute race is no longer about stacking cards: two paths for domestic AI compute

The compute race is no longer about stacking cards: two paths for domestic AI compute

At AICC 2026, an IDC report split the compute gap into two layers: the ceiling on capability and the scale of intelligence. By 2030 the shortfall could reach 380.9 billion dollars, and stacking cards alone can no longer cover the different loads of Prefill and Decode.

AI & models23 September 20267 min read
Six months of open-source LLMs: Xiaomi MiMo, DeepSeek and Chinese models on Hugging Face

Six months of open-source LLMs: Xiaomi MiMo, DeepSeek and Chinese models on Hugging Face

Xiaomi streamed an entire reinforcement learning run live, and MiMo-V2.6-Pro took the top open-source spot with 1.02 trillion parameters and 42 billion active. DeepSeek also released a vision experiment under the MIT license.

AI & models23 September 20266 min read
04

5,000 sandboxes a second: DeepSeek publishes DSec, its agent training infrastructure

A new DeepSeek paper signed by Liang Wenfeng describes DSec, an agent training system that produces more than 5,000 sandboxes a second, reaches 3 million in a single day and peaks at 380,000 running at once.

AI & models23 September 20265 min read
05

Same score, different bill: how to read open-source model benchmarks

On the Artificial Analysis Intelligence Index, Xiaomi's MiMo-V2.6-Pro and the closed-source Grok 4.7 both score 46. But one task costs about 29 times more on Grok 4.7. Cost is turning into a benchmark dimension of its own.

AI & models18 September 20265 min read
06

Interview: when AI learns to write GPU operators, what is left for engineers

Liu Sheng leads the main Attention operator on DeepSeek V4.1. He wrote "I Had to Bury My Talent Yesterday"; the post hit No. 1 on Zhihu's trending list within a day. Drawing on his public writing, we look at how AI coding is changing the job.

AI & models17 September 20266 min read
07

Chinese model Mureka V8 beats Suno V5 and claims a new music genre

Kunlun Tiangong showed off Mureka V8 at a concert in a music club and says the model beat Suno V5 in tests with musicians. The company's CEO talks about a new genre, not just another tool.

AI & models16 September 20265 min read
08

890 bytes per token: how DeepSeek shrinks the KV cache

Long context usually runs into memory, not weights. DeepSeek V4.1 Flash squeezes its KV cache to about 890 bytes per token, cutting HBM demand to a quarter of the previous generation and SSD demand to an eighth.

AI & models12 September 20265 min read
09

DeepSeek keeps V4 Pro API alive after calling off shutdown

V4 Pro was due to go offline on September 14. On September 11, DeepSeek said it would keep serving API calls, with billing unchanged. The earlier plan had routed all V4 Pro requests to V4.1 Flash.

AI & models11 September 20264 min read
10

DeepSeek ships V4.1 Flash: 552 billion parameters, native multimodal, cheaper calls

DeepSeek released V4.1 Flash on September 10 Beijing time. The new MoE model carries 552 billion parameters but activates only 8 billion on input, and it cuts KV Cache demand on HBM to a quarter of what the previous generation needed.

AI & models10 September 20265 min read
11

When agents start spending: WeChat Pay's AI card and the rollout of Chinese agents

WeChat Pay's AI card now works with DeepSeek Harness and OpenClaw. Once a user authorizes it, the card handles the whole flow from recommendation to payment inside a conversation. WeChat stresses main-account isolation and per-order confirmation.

AI & models1 September 20265 min read
page 14 / 14← previous1…121314