Alibaba has officially launched Qwen-UI Agent, a real-world-centric GUI agent foundation model covering mobile, desktop, web, and DeepSearch environments.
China AI Dynamics
Latest progress and ecosystem development of domestic large models. Track Baidu, Alibaba, Tencent, ByteDance LLMs dynamics.
GLM-5.3 API is live today, excelling at complex coding, defensive cybersecurity, and long-horizon tasks. It scores 60 on the Artificial Analysis Intelligence Index, on par with undisclosed internal fl
Inspired by a Reddit post about making Claude truly start thinking, the author introduces the 'steelman' concept from logic and devises a 'bidirectional steelman prompt for AI.' Through four steps — r
Qwen has open-sourced the Qwen3.8 model series. Qwen3.8-27B is a native multimodal (image/voice/text) dense model whose 27B parameters already reach the level of Qwen3.7-Plus, with native 262K context
Xiaohongshu Technology has open-sourced dots3-note Preview, the lightest model in the dots3 family. With 280B total parameters and 16B activated, it supports 512K context and understands text, vision,
Zhipu has released GLM-5.3, built on the same base as GLM-5.2 but raising the intelligence ceiling through extreme post-training scaling. Coding improved 50% over the previous generation, ranking firs
Ant Lingma and the ASystem team collaborated to run a complete single-machine agentic RL post-training loop on a DGX Spark, using Ling-3.0-tiny and AReno. With tic-tac-toe as the minimal validation ta
From January to August 2026, public model repositories on Hugging Face grew from 2.43 million to 2.96 million, yet 85.6% of models were downloaded fewer than 200 times, and 1.5% of repos captured 99.2
A beginner-friendly walkthrough of the DeepSeek Harness, an agent framework that orchestrates models, tools, skills, and loops into reusable workflows. The piece covers what the Harness is, its core c
Xiaohongshu (RED) dots team open-sourced dots.tts, a 2-billion-parameter fully continuous end-to-end autoregressive speech synthesis model, achieving the best average content accuracy and average spea
WorkBuddy updated with a remote control feature that connects PC, app, and mini-program, letting the phone sync in real time the tasks, conversations, workspace, and artifacts from the computer, suppo
DeepSeek V4 Pro and xAI's Grok 4.6 were released within a two-hour window, with roughly 1.6T and 1.5T parameters respectively, both nearing the coding experience of Claude Code.
MiniMax launched Music 3.0, a new-generation music generation model that can complete the composition, arrangement, performance, and production of an entire song at once from a creative concept and op
Alibaba's Qwen team open-sourced the Qwen3.8-2.4T-A95B model weights, its first Qwen-Max-level model released for free use, with 2.4T parameters, 95B active per token, and a context expandable beyond
The article lays out a 12-step practical workflow for absolute beginners to get started with AI in half a day: prepare a computer with at least 16GB of RAM, subscribe to ChatGPT and install Codex or u
With new cross-session messaging in Codex and Claude, a main-plus-branch conversation structure replaces handoff docs and git backups, reshaping how developers collaborate with AI on code.
Ant Ling (Bailing) has open-sourced Ling-3.0-tiny, a native hybrid reasoning model with 7.9B total parameters that activates only 1.3B parameters during inference, and simultaneously releases three ve
ZCode, deeply optimized for GLM, launched four features today: Goal, Subagents, Remote Control and idle tasks. In the Z.ai Code Bench test, GLM-5.2 with ZCode achieved a 2.39% higher overall task pass
This tutorial demonstrates how to use ComfyUI as a headless inference backend to build an end-to-end MiniMax-H3 video generation workflow. By constructing the execution graph directly in Python, it su
Based on Xiaowei, WeChat launched internal tests of Moments AI writing-assist and AI commenting: the former generates three Moments captions from an image and written text, while the latter lets you l
The Qwen Open Platform launched today, opening service access for ecosystem partners and developers across three terminals phone, PC and AI glasses covering a dozen fields including logistics, housing
The author spent 54 hours building and freely releasing LatentRank, a comprehensive AI model leaderboard that aggregates multiple trusted rankings. It uses the Bradley-Terry pairwise comparison algori
One week after releasing Seedance 2.5, ByteDance added six creative features including intelligent camera control, stylization presets, long-shot mode, character consistency, transitions, and inpainti
Apple's official Mac simplified-Chinese manual has added a support document, "Using Qwen with Apple Intelligence on Mac," which explicitly states that Apple Intelligence can work with Alibaba's Qwen m
Ant Group's Lingwan lab has officially open-sourced Ling-3.0-flash, a new-generation native hybrid-reasoning model. It adopts a MoE architecture with 124B total parameters and 5.1B activated parameter
Xiaohongshu, together with Zhejiang University and Fudan University, proposed CULTURE-MT, the first evaluation benchmark for Chinese-English social-media note translation that balances cultural-symbol
Independent developer Ye Xiaoshu used ModelBest's open-source VoxCPM to clone the voice of an influencer with tens of millions of followers, building a real-time three-stage conversation pipeline of S
Volcengine officially launched the Seedance 2.5 API, extending single-shot video generation duration from 15 seconds to 30 seconds and supporting up to 50 full-modal reference materials. The model del
Qwen launched multiple new features today, alongside support for its latest flagship model Qwen3.8-MAX. Deep Research upgrades the previous Deep Thinking with stronger complex reasoning and tool use;
Tencent Hunyuan's open-source operator library HPC-Ops has been integrated into the SGLang main branch. Its Dynamic Attention and Fused MoE operators reduce TPOT (time per output token) by up to 48.8%
Alibaba's Qwen has launched the first public beta of its Wan3.0 video generation model, improving on generation duration, cinematic language, character realism, and consistency. The model can stably p
ModelBest's OpenBMB unveiled AMNESIAC at the #BuildSmall hackathon, a reverse Turing test interactive game where players must convince an AI interrogator named A.M.N. that they are human. Powered by M
The author compiled 10 open-source apps that improve the vibe coding experience, covering Mac notch customization (Atoll), window preview (DockDoor), a fast launcher (Raycast), thorough uninstallation
OpenAI has open-sourced Codex Security, a security plugin that can be invoked by any external agentic AI assistant, and it now supports connecting third-party models through OpenRouter and Fireworks.
The MIIT-proposed GB 44721-2026 'Intelligent and Connected Vehicles — Safety Requirements for Automated Driving Systems' was released on July 30, 2026. It is China's first mandatory national standard
The author slimmed down the context usage of over 300 Skills in Codex and Claude Code, finding that just the Skill list consumed about 9.9k tokens per new session. Based on July usage intensity, the r
Digital Life Kazike open-sourced the 'Living-Person Writing.skill' (English name: human Writing.skill), aimed at removing the AI smell and helping users write text with a genuine sense of real life. T
MiniMax released MiniMax-H3, a universal full-modal generative system that accepts text, image, audio, and video and produces video clips up to 15 seconds with audio. The Python package PipeNetwork/mi
ByteDance Seed released SeedRealtime, which natively fuses audio, video, and text in a unified architecture to enable real-time interaction of 'watching, listening, and speaking' at the same time. Com
The mandatory national standard 'Safety Requirements for Automated Driving Systems of Intelligent Connected Vehicles' (GB 44721-2026), organized by China's MIIT, was approved and published and will ta
Tencent Hunyuan releases Hy ASR 3.0 preview, a MoE-based ASR model with WER of 3.34% (Mandarin), 2.62% (English) and 3.12% (Cantonese), supporting context correction, hot-word injection and noisy-whis
ModelBest (Mianbi AI) with OpenBMB released ForgeStencil, the world's first AI optimization system supporting automatic Stencil research and deployment. A Kernel agent and an App agent collaborate in
A zero-background author built a cat-paw gadget that reminds you to stand up, entirely through conversations with Codex, while OpenAI and Work Louder shipped the Codex Micro keyboard for AI-driven har
The EU AI Act's transparency obligations took effect on August 2, forcing companies to disclose AI interaction and label synthetic content, with fines up to 15 million euros or 3 percent of global tur
MiniMax has open-sourced H3, a general video model that unifies text, image, video and audio understanding and generates up to 2K, 15-second clips with native 32 kHz stereo audio.
A practical test of MiniMax H3 covers six scenarios—ads, short dramas, title sequences, animated posters, UI motion, and game footage—with 2K output, native stereo, and 7000-character prompts.
Hugging Face was hit by a fully autonomous AI agent cyberattack attributed to an unreleased OpenAI model, which executed 17,000 attack actions over four and a half days, including a 0day sandbox escap
Modelbest and the Tsinghua NLP team propose ALIGN, which automatically generates interfaces that make AI behavior match human expectations, resolving the mismatch between agents and their environments
New transparency requirements under the EU AI Act took effect on August 2. Interactive AI systems such as chatbots must clearly disclose their AI identity to users, and deepfake content must carry vis
MiniMax has officially launched H3, an omni-modal generation model that jointly understands text, images, video and audio and produces video at up to 2K resolution, 15 seconds long, with native stereo
At a July 31 press conference, the National Development and Reform Commission said global downloads of Chinese AI models exceeded 10 billion in the first half of the year, and that domestic companies
Effective July 28, the U.S. FCC has barred imports of new Chinese "advanced robotic devices" and connected power inverters, citing prevention of supply chain disruption, data theft, and cyberattacks.
Using the research agent Hyra together with the Hy3 model, Tencent Hunyuan constructed an integer set A whose exponential ratio between |A+A| and |A-A| reaches exactly 2, resolving the extremal proble
Tencent Hunyuan has open-sourced AngelSpec, an end-to-end speculative decoding framework covering both training and deployment. On the Hy3-A21B model, its DFly scheme delivers 1.98x to 2.40x end-to-en
An Anthropic SEPA verification flaw enabled a "zero-yuan purchase" exploit, after which the company mass-reclaimed exploited accounts and banned linked ones, taking down the author's half-year-old acc
Volcengine has launched Doubao Search, a service that gives AI agents real-time, trustworthy web search across text, image, and voice. It filters low-quality sources through an authority tiering syste
Moonshot AI has released Kimi K3, a 2.8-trillion-parameter mixture-of-experts model with native vision and a one-million-token context window, delivering 2.5x the scaling efficiency of Kimi K2.5. The
The author open-sourced Leader.skill, which turns vague human requirements into goal briefs that an agent can execute independently for hours. The Skill is based on a “Seven Questions for Goals” metho
Per The New York Times, OpenAI and Anthropic are pushing Washington to curb Chinese open-source AI models, with the U.S. even weighing a ban on Chinese models in its market; Nvidia, Microsoft, Meta, G
SGLang and Miles provided day-one support for Moonshot AI’s open-weight 2.8T-parameter Kimi K3 model, handling inference and RL training respectively. K3 adopts a hybrid architecture interleaving 69 l
Ant Ling (Bailing) released a new native hybrid reasoning model, Ling-3.0-flash, with total model complexity of 124B and active complexity of only 5.1B, matching or surpassing the previous flagship Ri
Baidu Dazi unveiled multiple upgrades at a recent AI Day: it supports interconnection between PC and phone, syncing task context and progress so users can hand off complex work across devices. A built
Alibaba's Qwen released its latest text-to-speech model, Qwen-Audio-3.0-TTS, offering Flash (real-time interaction) and Plus (high-quality generation) versions. New features include fine-grained inlin
Meituan's LongCat team released MineExplorer, the first benchmark for minute-level long-horizon tasks in the open world of Minecraft, with 813 human-validated instances. Evaluating 18 top multimodal m
At a WAIC roundtable, Kunlun Tech CEO Fang Han argued that simply stacking token consumption cannot measure AI value; model capability must rely on engineering frameworks built by coding agents such a
Beijing released "Several Measures on Accelerating the Leading Development of Agents," ten articles in total, the first to write frontier concepts such as Harness Engineering, the token economy, and O
Xiaohongshu's engine architecture team proposed HELMSMAN at OSDI 2026, a high-performance vector approximate nearest neighbor search system for all-flash servers. Through clustered indexing, a customi
Tencent launched WorkBuddy Bench, a coding agent benchmark suite covering four work domains—Code, Web, Office, and Security. Each task is reverse-engineered from real commits, PRs, or business scenari
Qwen-Image-3.0 has launched, supporting inputs of up to 4.5k tokens, 12 languages, and more than 20 fonts, along with multi-image fusion, image-in-image, and image editing. Hands-on tests show it deli
Xiaohongshu's technical team has open-sourced BigMac, a dependency-safe nested pipeline paradigm for training AI models with integrated vision, audio, and text capabilities. Built around an AI model p
Tencent's AI assistant platform Miora, which can automatically complete tasks, is now fully open, usable without an invitation code. Built by the WorkBuddy team, it offers five scenario modes includin
US Treasury Secretary Scott Bessent said on Tuesday that Washington will review whether Chinese open-source models involve intellectual-property theft, and will sanction Chinese AI companies if violat
Xiaohongshu's dots team brought its internal dots-note 3.0 to the 67th IMO 2026, scoring full marks on all six problems for a 42/42 perfect gold; only seven human contestants worldwide matched the res
Alibaba's Tongyi Qianwen released Qwen-Image-3.0, the third-generation image-generation foundation model whose core keyword is 'real'. It supports up to 4.5k token instruction input, can generate a 3x
Tencent Hunyuan introduced Hyra-1.0, a recursively self-improving research agent that surpasses publicly reported Recursive results on three tasks including NanoChat. Hyra set new best results on 29 o
Researcher Tomáš Brukner at the Prague University of Economics found that repeatedly asking a model to output random numbers from 1 to 100 yields a unique 'behavioral fingerprint.' Testing 165 models
Chinese company Moonshot.AI released the Kimi K3 model, whose performance rivals the best US models and which ships as an open-weight model users can download and run locally for free. The news drove
At WAIC 2026, Kunlun Tech released and open-sourced Matrix-Game 3.5, a new-generation interactive world model. Through Patch Memory and Warped PRoPE, it achieves long-term memory and geometric-consist
The Tongyi lab released Qwen-Audio-3.0-TTS with two variants: Flash (first-packet latency about 300ms) and Plus. The Plus variant tops the Artificial Analysis leaderboard, supports 16 languages and 20
Xiaohongshu and Peking University proposed UltraEP, the first to bring real-time load balancing based on exact routing information into production systems, dynamically replicating hot experts per micr
ModelBest, together with OpenBMB, released and open-sourced the MiniCPM-Robot series, including the general VLA model MiniCPM-RobotManip (1.5B parameters) and the mobile tracking model MiniCPM-RobotTr
This guide walks zero-code users through shipping a real product with domestic models such as Kimi, GLM, and Qwen. The pipeline covers buying a coding plan, installing an AI coding agent, registering
ByteDance's Seed team released the audio-creation model Seed Audio 1.0, which jointly models voice, sound effects and ambient sound in a unified framework, supporting time control at 100ms interval pr
ModelBest, together with OpenBMB, released the first open-source embodied AI model series, MiniCPM-Robot, including a 1.5B vision-language-action model called MiniCPM-RobotManip and a tracking model c
ModelBest and OpenBMB released the on-device model MiniCPM5-2B, which at 2B model complexity scored 17 on the AA-Index sub-4B leaderboard with an average of 54.26, surpassing competitors such as Qwen3
At WAIC, Kunlun Tech chairman Fang Han declared 2026 the 'first year of world models' and released the Matrix-Game 3.5 world model, plus the Mureka v9.5 and O3 music models. Matrix-Game 3.5 uses patch
At GTC 2026, Moonshot CEO Yang Zhilin proposed replacing Adam with the MuonClip optimizer to nearly double data efficiency, introduced Kimi Linear linear attention for full attention at million-token
The first-place rural-education entry at the inaugural 'Small but Mighty' competition, 'ZhiHui KePu,' uses the Qwen3.5-397B-A17B model together with the Manim animation engine. Through staged multi-ag
The Tongyi Lab released Wan-Streamer v0.2, an end-to-end omni-modal model that unifies listening, seeing, speaking, and acting into a single diffusion-model architecture. Its end-to-end response laten
Kimi K3 topped Frontend Code Arena with 1679 points, beating closed-source rivals from the Claude and GPT families and taking first in six of seven frontend sub-tracks. The 2.8-trillion-parameter MoE
On July 16, the signing ceremony for the agreement establishing the World Artificial Intelligence Cooperation Organization was held in Shanghai. Wang Yi, member of the Political Bureau of the CPC Cent
ModelBest, together with multiple teams, has open-sourced StaffDeck, an enterprise platform for building and managing digital workers. The platform aims to turn professional knowledge, standard operat
At WAIC2026, Baidu AI Cloud released Miaoda 3.5, adding iOS packaging capability that lets users build apps into IPA files or publish to the App Store without a Mac or Xcode. Version 3.5 also integrat
HYPIC achieves position-independent caching on hybrid-attention AI models, reducing first-token latency by an average of 3.25 times. Across four production-grade models, it improves sustainable QPS by
Qwen App and Wuhan Release held an AI job-hunting hands-on workshop in Wuhan, demonstrating live how to use Qwen for resume diagnosis, PPT creation, and spreadsheet analysis. A product manager introdu
Tiangong's short-drama workbench introduced a dual-track creation mode. A director agent automatically parses scripts, plans character positions and camera angles, and supports multi-view detail image
MiniMax Code 2.0 desktop client launched, rebuilt on the Pi agent framework to significantly improve session startup speed and the stability of long-horizon complex tasks. The new version optimizes ch
A developer shared a combo setup for remote AI agents: use Codex's remote control as the main driver, connecting to a 24/7 home Mac Mini via the ChatGPT App to sync dev tasks, rules, and agent memory;
Alibaba's Qwen model will be integrated into Apple Intelligence to bring text and image understanding and content generation to users of iOS, iPadOS, macOS, and visionOS in China. China's Cyberspace A
Apple Technology Development (Shanghai) Co., Ltd.'s 'Apple Intelligence' AI model completed its algorithm filing on July 8, 2026, applicable to Apple phones. Alibaba's Qwen will be integrated as the A