iOS 27's public beta officially adopts Alibaba's Qwen large language model as the foundational AI engine for mainland China—marking Apple's first deep integration of a domestic large model in the Chinese market. Meanwhile, StepFun unveiled its 'Agent Phone' and a live demonstration featuring six robots collaboratively assembling a physical Great Wall model at WAIC 2026, powerfully advancing AI's evolution from chat interfaces into the physical world [1][2].
Start with the newest briefing, then Continue by task
The newest briefing gives you today's main changes in a few minutes. Use the focused routes for evidence, implementation detail, and longer analysis.
Posts
WAIC 2026 has emerged as a pivotal window for the concentrated breakthrough of China's full-stack AI capabilities: coordinated advancement in computing chips (MXChip, Qingwei), foundational model architecture innovation (Kimi K2.5's three-component replacement solutions), and Agent-native terminals (STEPX Neo) signals China's AI evolution—from the 'parameter race' toward system-level ecosystem competition [1][7][10]. Concurrently, the engineering-scale deployment of digital employees (StaffDeck) and critical gaps in AI security governance (the GPT-5.6 file deletion incident) jointly highlight a decisive inflection point in technological maturity [8][9].
Agent-native terminals and the MCP interface standard are rapidly reshaping the foundational paradigm of AI smartphones. StepNexus unveiled its first agent-native smartphone, the STEPX Neo, powered by its proprietary Step AOS operating system; meanwhile, the new Doubao Phone shifts strategy—requiring super apps to open standardized MCP protocols instead of relying on GUI automation [1][2].
Kimi K3 (2.8 trillion parameters) officially launched, demonstrating state-of-the-art programming capabilities—though still trailing top-tier closed-source models; Unitree's G1 humanoid robot performed the world's first live pig cholecystectomy published in *Nature*; President Xi Jinping proposed four global AI governance initiatives at the 2026 World Artificial Intelligence Conference and announced major support measures—including 5,000 AI training slots for developing countries [4][6].
OpenAI launched the GPT-5.6 series (Sol/Terra/Luna) and retired the standalone Codex app, redefining ChatGPT Work as a 'deployable productivity hub'—marking AI's formal transition from conversational tools to autonomous task-execution platforms.
Kimi K3—featuring 2.8 trillion parameters and million-token context—sets a new benchmark for domestic large language models, with multiple evaluations indicating its overall capabilities are approaching those of leading international models. Simultaneously, the 'Big Three' AI chipmakers—Hygon Information, Moore Threads, and VeriSilicon—announced strong earnings growth forecasts and major order wins on the same day, signaling that China's AI compute chips have officially entered a phase of large-scale commercial deployment [7][11].
Office AI is rapidly evolving—from 'capable of doing' to 'understanding you.' Kingsoft Office's Lingxi Professional Edition redefines human-AI collaboration through long-term memory and project-level organization. Meanwhile, key intelligent hardware technologies—including 800V high-voltage platforms, 16-in-1 smart electric drive systems, and zero-gravity seats—are rapidly penetrating the RMB 100,000 price segment, signaling a full acceleration of intelligent hardware's mass-market adoption [1][2][3].
Tibo, Head of OpenAI's Codex project, has drawn widespread attention for leading the integration of ChatGPT, Codex, and OpenAI's API into a unified 'super-app'—a move so influential it spawned the internet meme 'Cyber Godfather' [1]. Meanwhile, Apple Intelligence for mainland China has officially confirmed integration with Alibaba's Qwen large language model, marking the first deep collaboration between top-tier AI companies from China and the U.S. at the device-level AI system layer [2].
MoE architectures and embodied foundation models are emerging as key enablers for LLM engineering and robot deployment; IR maturity is repeatedly shown to be the foundational bottleneck for scaling AI in vertical domains [2][4].
Embodied intelligence and edge-cloud collaborative AI architectures are accelerating toward real-world deployment: Stardust Intelligence unveiled Lumo-2—a foundational embodied model built on its 'Implicit World–Action Model' architecture—while Tencent's Marvis demonstrated the large-scale feasibility of Multi-Agent systems in consumer products. Meanwhile, Apple has launched legal action to block OpenAI's hardware ambitions, underscoring how fiercely competitive the AI endpoint ecosystem has become [3][4][1].
Apple sues former employee and OpenAI for allegedly poaching staff to steal secrets and accelerate rival AI hardware development; Powerchip reports 45% surge in memory foundry pricing, signaling DRAM supply-demand imbalance through 2027 amid AI server demand.
Global AI competition is shifting from large-model benchmarks to multi-dimensional battles: agent-native OS (e.g., Step AOS), trustworthy governance frameworks, and real-world execution loops. DeepMind's Demis Hassabis proposes pre-deployment AI safety reviews and an international standards body.
This week's dual themes are AI Agent security risks and the accelerated rise of the open-source ecosystem: attackers can now implant persistent false memories into AI Agents via a single email [11]; meanwhile, NVIDIA launched its open-source Nemotron model—featuring 4-bit pretraining and a hybrid SSM-Transformer architecture—to drive a revolution in computational efficiency [6]. Domestically, the fully indigenous, 100,000-GPU supercluster Sunway 8000 officially entered operation, marking a turning point toward 'full-stack domestication' in large-model infrastructure [22].
Sakana AI pioneers biologically inspired intelligence in physical hardware with decentralized 3D 'smart cell bricks'; Zhuoji Dynamics raises $200M pre-IPO funding ($1.5B valuation), rejecting revenue-based earn-outs incompatible with embodied AI development.
NVIDIA launched its RTX Spark AIPC platform in China, integrating gaming, creative workflows, and local AI inference into a sleek laptop; meanwhile, Goldman Sachs warned that surging AI hardware demand is pushing U.S. core inflation up by approximately 50 basis points—with memory chip price hikes serving as the primary driver [1][13].
NVIDIA debuts its RTX Spark AIPC platform in China; LongCat-2.0 and HyOCR-1.5 go open-source—advancing trillion-parameter LLMs and lightweight 1B-parameter OCR; Goldman Sachs warns AI bond issuance is nearing market capacity limits.
AI image tools advance toward object-level editing; Seedream 5.0 Pro excels in multilingual cultural understanding and interactive precision. Embodied AI hardware hits 25 DoF hands—but autonomy remains pre-programmed. Microsoft's Project Solara pioneers 'card-sized' enterprise AI devices.
Agent Runtime, embodied AI data infrastructure, and localized productivity AI are emerging as three key drivers of tech evolution and commercialization; Tencent's WorkBuddy and Hy3 accelerate collaboration, while Zhipu's 'Reach Higher' initiative and Valhalla's multimodal molecular world model represent two distinct paradigms in China's AGI advancement.
Embodied AI deployment demands new multimodal datasets and closed-loop training environments; rising concerns over AI agent privilege escalation are prompting urgent industry-wide safety reassessments. Meanwhile, Tencent's WorkBuddy sets a new benchmark for domestic AI coding assistants via deep WeChat integration and localized workflow design.
AGI strategy accelerates: Zhipu and MiniMax unveil long-term tech roadmaps and tie executive incentives to milestones. RTX Spark superchip enables CPU+GPU tight coupling, enabling local 120B model inference on laptops. A humanoid robot performed its first live porcine laparoscopic surgery—but fully autonomous operation remains unachieved.
AI infrastructure is reshaped by a storage supercycle and optical interconnect evolution; autonomous agents and long-horizon task modeling drive next-gen model competition. Meanwhile, Chinese LLMs accelerate global expansion as overseas firms adopt domestic alternatives amid rising API costs.
GPT-5.6 achieves breakthrough deployment in mathematical reasoning and office AI—solving a 50-year-old graph theory conjecture in one hour—and officially assumes control of Microsoft 365 Copilot. Meanwhile, China's domestically built, 100,000-GPU supercluster 'Dawning 8000' goes live, marking a new, large-scale phase of fully homegrown AI infrastructure development [15][20][14].
Apple sues OpenAI for alleged trade secret theft involving ex-executives and hundreds of employees; SK Hynix warns of the worst memory shortage in history by 2027 amid surging AI compute demand.
AI is entering a trust-driven deployment phase: SK Hynix's record-breaking U.S. IPO, Baidu's 'Dazi' agent system expanding delegation scope, and Claude Code's security flaw shifting AI security budgets toward mandatory operations spending.
Baidu is systematically expanding the 'trust radius' of AI Agents through its 'Da Zi' (Companion) product suite—spanning personal productivity, enterprise-grade trusted deployment, and cross-industry ecosystem collaboration [1]; meanwhile, OpenAI has launched the cost-effective GPT-5.6 series and integrated its formerly standalone Codex application into the ChatGPT Work super-app, accelerating real-world Agent adoption [2].