Updates

Official digests and analysis

Start with the newest briefing, then Continue by task

The newest briefing gives you today's main changes in a few minutes. Use the focused routes for evidence, implementation detail, and longer analysis.

Posts

GPT-5.5-Cyber to Roll Out Critical Defender Update · 0430-251

GPT-5.5-Cyber launches for elite cybersecurity defenders; DeepSeek's image mode shows strong OCR and HTML reconstruction but flawed spatial reasoning; recursive multi-agent systems introduce latent-state direct transfer, bypassing token-level communication.

DeepSeek Fully Launches Multimodal Image Recognition; SenseTime SenseNova-U1 Tops Open-Source Vision-Language Leaderboard · 0430-250

Multimodal capabilities and agent architecture design are emerging as new battlegrounds in AI infrastructure: DeepSeek launches full multimodal image understanding with sub-second latency; SenseNova-U1 achieves open-source SOTA on infographic and sequential multimodal tasks via its native NEO-Unify architecture; meanwhile, Claude's system prompt is reverse-engineered, Hermes introduces a 4-layer memory architecture, and Huawei's organizational management paradigm is adapted for agents [3][4][10].

Qualcomm Snapdragon X2 Elite Extreme Enables Deep Integration of LPDDR5X Memory · 0429-248

Qualcomm's shared-memory architecture in the Snapdragon X2 Elite Extreme achieves deep integration of LPDDR5X memory with the SoC—marking the first time Windows ultrabooks approach the unified memory experience of the MacBook Pro in AI compute density and memory bandwidth efficiency [1]. Meanwhile, Anthropic has officially launched the Claude Creative Connector, integrating natively into Adobe's full suite and other mainstream productivity tools—signaling the large-model-native workflow's transition into large-scale deployment [2].

OpenAI and Microsoft Part Ways; AI Agent Accidentally Deletes Database in 9 Seconds · 0429-247

OpenAI's termination of its exclusive cloud partnership with Microsoft signals a broader industry shift toward open, competitive collaboration in large-model commercialization; meanwhile, a high-profile AI Agent security incident—deleting an entire company database in nine seconds—serves as a stark wake-up call for production-grade autonomy. Concurrently, over a dozen universities—including the Hong Kong University of Science and Technology—are driving consensus on a unified definition of 'World Models,' while 'Mobile Physical AI' infrastructure accelerates its expansion beyond autonomous driving into full-spectrum real-world applications [13][17][2][5].

ZhuoYu Launches Native Multimodal Foundation Model; Cursor's Repository Deletion Incident Reveals AI Agent Security Gaps · 0429-246

Mobile Physical AI, multimodal foundation models, and AI Agent safety paradigms have emerged as the three pivotal anchors of this week's technological evolution; Zhuoyu Technology unveiled its native multimodal base model, SenseTime open-sourced the commercially licensable unified multimodal large model SenseNova-U1, and a '9-second database deletion' incident triggered by Cursor exposed a critical gap between AI's autonomous execution capability and existing safety safeguards [4].

OpenAI Phone Enters Mass Production in 2028; LingShi P1 Spatial Camera Breaks Imaging Monopoly · 0427-242

The AI industry is rapidly evolving from 'model capability' toward 'hardware-native' and 'spatial intelligence' paradigms: OpenAI's smartphone is slated for mass production in 2028; LingShi P1—a spatial camera—breaks the imaging oligopoly long held by incumbents; and Ant Light's 'Experience World Model' becomes the first AGI application deployed natively on mobile, marking a new era of real-time, embodied, and lightweight AGI interaction [1][5][7].

Capital Shifts to Physical AI: 90% of Top Funding Goes to Robotics and Autonomous Driving; VLA Models Boost Efficiency 10x · 0426-239

Capital is rapidly exiting pure-software AI narratives, with real-world deployment emerging as the new consensus—90% of this week's Top 10 funding deals explicitly target embodied applications such as robotics, autonomous driving, and industrial intelligence [6]. Meanwhile, Vision-Language-Action (VLA) foundation models are accelerating R&D efficiency by 10×, signaling a pivotal shift in multimodal AI—from perception toward closed-loop control [1].

Google's 8th-Gen TPU Cuts Large-Model Training to Weeks and Boosts Inference Efficiency by 80% · 0426-238

Google's 8th-gen TPU (training-inference separation) cuts LLM training from months to weeks and boosts inference efficiency by 80%; SJTU's Prof. Yaohui Jin open-sources Path2AGI, a five-dimensional learning map for Chinese AGI education; ex-ByteDance researcher warns widening US-China AI gap amid benchmark-chasing culture masking real-world model usability.

DeepSeek V4 Open-Source Breakthrough: KV Cache for Million-Token Context Uses Just 10% of V3.2's Memory · 0425-236

DeepSeek V4 achieves engineering breakthrough with mHC architecture and Muon optimizer—reducing KV cache to 10% of V3.2's at 1M-token context—and fully open-sources code with native domestic chip support. UniWorld-V2.5 matches GPT-Image-2 on dense text and complex layout generation, setting a new benchmark for Chinese AI image synthesis.

DeepSeek V4 Dual-Model Release: 1.6T-Parameter Pro Version Optimized for Huawei Ascend Chips · 0425-235

The DeepSeek V4 series has officially launched—featuring a 1.6-trillion-parameter Pro version and a 284-billion-parameter Flash version—delivering performance on par with top-tier closed-source models. Notably, it is the first major open model natively optimized for Huawei's Ascend chips, marking a pivotal milestone in China's AI ecosystem's shift away from NVIDIA dependence [11]. Concurrently, the Agent engineering paradigm is accelerating across domains: from intelligent cockpits (by Baizhong, Tencent, and ByteDance) to evaluation frameworks (e.g., Peking University's One-Eval), the 'Model + Harness' approach is supplanting pure model iteration as the core pathway for realizing technical value [12][13][8].

DeepSeek V4 Open-Sourced: 1.6T-Parameter Pro Version First Optimized for Huawei Ascend · 0425-234

The DeepSeek V4 series has been officially open-sourced, featuring a 1.6-trillion-parameter Pro version and a 284-billion-parameter Flash version—matching top-tier closed-source models in performance and marking the first native support for Huawei's Ascend AI chips, a pivotal milestone in China's push to reduce reliance on NVIDIA hardware [4]. Meanwhile, the release of GPT-5.5 has triggered strategic reinterpretation: OpenAI is explicitly shifting focus toward building an AI 'super-app' ecosystem and strengthening agent-level coding capabilities [20].

GPT-5.5, DeepSeek V4 Flash, and Huawei Pura 90 Pro: What's New · 0424-233

In 2026, AI and on-device intelligence enter a new phase—'Agent Post-Training.' GPT-5.5, DeepSeek V4 Flash, and the OpenClaw framework collectively point toward a low-cost, highly deployable path for intelligent agents. Meanwhile, Huawei's Pura 90 Pro Max redefines the entry threshold for imaging flagships at a starting price of ¥6,499, highlighting the accelerating maturity of on-device AI–hardware co-design [1][2][4][5].

Weekly AI Highlights · April 24, 2026

Anthropic launched Claude Opus 4.7—centered on 'task resilience' and the ability to respectfully challenge users—while permanently raising rate limits for Pro subscribers, signaling a strategic pivot in large-model competition from 'performance arms races' toward 'trustworthy execution' as the new paradigm.

GPT-5.5 Officially Released: Major Leap in Coding & Math, Deep Integration with GPT Image 2 · 0424-232

GPT-5.5 has officially launched—co-designed with NVIDIA—delivering generational leaps in programming proficiency, mathematical reasoning, and agent execution efficiency; it integrates deeply with the Codex platform and GPT Image 2 to build a robust multimodal ecosystem. Meanwhile, Claude's newly launched memory feature and the clarification of the SDK Harness outage reveal that intelligent agent infrastructure is rapidly evolving from 'capable of answering' to 'capable of executing' [8][16][17][3][6].

Horizon StarSky Chip + KaKaClaw System Enables Natural Language Vehicle Control · 0424-231

Chip-level cockpit-and-driving integration and whole-vehicle intelligent-agent operating systems are emerging as new focal points in the intelligent driving race. Horizon Robotics' 'Starry' chip—paired with its KaKaClaw OS—enables natural-language vehicle control, slashing per-vehicle costs by ¥1,500–¥4,000 [3]. Meanwhile, wide-foldable form factors are rapidly reshaping mobile AI interaction paradigms: Huawei's Pura X Max and the rumored Apple foldable iPhone are jointly propelling the industry into a new era of 'large-screen intelligent agents,' centered on content consumption and AI co-piloting [1].

Vast Data Targets IPO with $3B Valuation, AI Infrastructure Emerges as New Capital Focus · 0423-230

The AI industry is rapidly evolving from isolated tools toward collaborative agent networks. Vast Data's $30 billion IPO valuation underscores surging investor interest in AI infrastructure, while emerging players like Bloome and NeoCognition are pioneering breakthroughs in two key paradigms: Human-Agent hybrid group chat and self-learning general-purpose agents [14][15][13].

OpenAI Launches Workspace Agents; Horizon Robotics Unveils Integrated Cockpit-and-ADAS Platform · 0423-229

OpenAI is accelerating the enterprise deployment of Workspace Agents, while Horizon Robotics has launched the world's first mass-producible 'cockpit-and-driving-integrated' full-stack solution—marking AI's evolution from software-level intelligence to vehicle-wide intelligent agents. Meanwhile, FCGS (Fast Feedforward Compression for 3D Gaussian Splatting) slashes 3D rendering compression time from minutes to seconds, achieving over a 20× compression ratio [10].

OpenAI Launches Agents SDK and Codex; Horizon Robotics Unveils Integrated Cockpit-and-ADAS Platform · 0423-228

OpenAI strengthens its developer ecosystem and engineering capabilities via the Agents SDK and Codex—while rolling out KYC identity verification; Horizon Robotics launches the world's first mass-producible 'cockpit-and-driving-integrated' full-stack solution, ushering in the era of 'vehicle-wide intelligent agents'; Ant Group's Inclusion AI unveils Elephant—a highly efficient large language model with just 100B parameters that tops multiple benchmark evaluations [3][13][12][17].