Foundation Models
LLMs & AI modelsChina's foundation models — Doubao, DeepSeek, Qwen, Kimi, GLM, Ernie, Hunyuan and more: releases, capabilities, open-source moves and benchmarks.

Qwen-Audio-3.1 Adds Five Models Spanning Audio Understanding, Generation and Interaction
Alibaba's Qwen team released the Qwen-Audio-3.1 series on September 23, adding five models covering speech recognition, synthesis, real-time interaction and audio creation. Prices were cut across the board, with ASR down as much as 95%.

Tencent Hunyuan Launches Hy Image3.5 Preview With Multi-Turn Creation
Tencent's Hunyuan team has launched the Hy Image3.5 preview model for professional-grade image generation and editing, supporting text-to-image, image-to-image, and multi-turn conversational creation. The model is already live across Yuanbao and several partner apps, with API access on Tencent Cloud's TokenHub.

Xiaomi Open-Sources MiMo-V2.6, Scaling Reinforcement Learning for Model Self-Improvement
Xiaomi has released and open-sourced its MiMo-V2.6 series, two full-modality models built around large-scale reinforcement learning to explore recursive self-improvement. The company says MiMo-V2.6-Pro ranks among today's stronger open-weight models while keeping prices unchanged.

Qwen Open-Sources Qwen-Image-2.1, a 7B Unified Image Generation and Editing Model
Qwen has open-sourced Qwen-Image-2.1, a 7B-parameter model that unifies text-to-image generation and image editing, with native RGBA transparent output, up to 10 reference images, and mask-based local editing. The release also adds KV cache reuse for inference efficiency and targets portrait and product consistency.

Shengshu Technology Launches Vidu S2 with Real-Time Video Generation, Editing and Spatial Video
Shengshu Technology has released Vidu S2, a real-time video model that comes in two variants: S2-Avatar for interaction with digital characters, and S2-Editing for editing video streams. The new model supports real-time 720P output, lets users inject reference images mid-generation, and can adjust clothing, subjects, backgrounds and visual style in a video on the fly. The team also demonstrated spatial video generation and editing aimed at VR headsets.

Doubao Mobile Assistant Adds Screen-Based Tasks, Personal Memory and Cross-App Actions
ByteDance has launched the consumer edition of Doubao Mobile Assistant with expanded screen interaction, on-device information retrieval, personal memory and cross-app task execution. The assistant will debut on nubia's NaviX Ultra, while ByteDance says reliability, GUI agents and cross-app integration still need improvement.

Xiaohongshu's AllSpark Releases Iris: Leading Performance Among Open-Source Search Agents in Its Class
Xiaohongshu's AllSpark team open-sourced Iris Search Agent weights and evaluation code, with Iris-mini and Iris-pro topping BrowseComp, BrowseComp-ZH, DeepSearchQA and HLE benchmarks in their respective parameter classes.

Un, the First Large-Scale Generative Model to Use Physics as a Computational Primitive
Over the past decade-plus, GPU-centric digital computing has dominated AI. Bigger clusters, higher bandwidth, more powerful GPUs, and denser data centers have seemed to be the mainstream path toward the next generation of AI.

DeepSeek and Kimi Diverge as Funding Rises
Liang Wenfeng and Yang Zhilin Start Answering Different Questions

ByteDance Lights Up Its Own ChatGPT Moment During Chinese New Year
The real AI offensive is a hard decision to turn itself into a technology company.

Doubao’s 500-yuan-a-month Pro Plan Is Officially Here. Is It Expensive?
Can an Agent-Driven Office Workflow Really Replace Employees?

Huawei’s SpaceMind Tops a Leading Spatial Intelligence Benchmark: A Pure RGB Vision-Language Model Scores 70.6, Setting a New Record on Fei-Fei Li’s Leaderboard
Large models can now hold fluent conversations and recognize objects in images, but a more fundamental question remains unresolved: do they actually “understand” the three-dimensional world we live in?

Huawei Open-Sources 7B Multimodal Model With Strong Visual Grounding and OCR, Bringing a New Ascend Edge “Sweet Spot
Models in the 7B parameter class have long been a favorite for edge deployment and individual developers.

Kimi K3 Sparks Heavy Demand as GPUs Struggle to Keep Up
Kimi K3 is drawing strong attention, with demand apparently outpacing available GPU capacity. The headline points to pressure on compute resources around the model, but the body provided is empty and details cannot be verified.