## 🔍 Key Insights The AI industry saw a dense wave of updates this week: **OpenAI** strengthened prompt caching for **GPT-6**, with cache reuse cutting input token costs by up to **90%** [1]; **Artificial Analysis** benchmarks show **GPT-6 Sol** and **Luna** matching their predecessors on the intelligence index while halving per-task cost [10]. Meanwhile, **METR** released a pre-deployment evaluation of **Claude Opus 5.5**, concluding its acceleration of AI R&D is a gradual improvement rather than a jump [0], while **Fei-Fei Li** called for AI safety evaluations not to be left solely to the developing companies themselves [3]. ## 🚀 Top Stories - **METR releases pre-deployment evaluation of Claude Opus 5.5** [0]: Concluded it is a gradual improvement and unlikely to fully automate AI R&D. - **OpenAI enhances GPT-6 prompt caching** [1]: Cache reuse cuts input token costs by up to 90%, with new diagnostic and warm-up tools. - **Fei-Fei Li calls for independent oversight of AI safety** [3]: Says increasingly capable systems should not be entirely self-evaluated by the companies developing them. - **GPT-6 Sol and Luna halve costs** [10]: Intelligence index similar to predecessors, token pricing down about 50%. - **Snorkel AI closes $350M Series E** [20]: Valuation rises to $3.5 billion, annualized revenue reaches $375 million. - **Qualcomm Snapdragon Summit focuses on the agent era** [22][24]: Amon says inference and edge-side are key; Xiaomi, vivo, Honor and others are among the first to adopt the new platform. - **arXiv math submissions surge 33.5% in 2026** [11]: 29 of 30 subfields saw growth, not driven by a single hot topic. - **PixVerse R2 real-time world model released** [16]: Supports prompt-based control and editing, character memory, and responsiveness. ## 🔗 Sources [0] METR releases pre-deployment evaluation summary of Claude Opus 5.5 — https://aihot.news/items/cmudbo6sf04ntroggkmjafzax [1] OpenAI enhances GPT-6 prompt caching reliability and controllability — https://aihot.news/items/cmudbhd2b04fdroggn32l8x0w [3] Fei-Fei Li on AI safety: Don't let developers evaluate their own systems — https://aihot.news/items/cmudaybbm03zoroggjn1m7tm5 [10] Artificial Analysis benchmark: GPT-6 Sol and Luna maintain Intelligence Index scores close to the GPT-5.6 series at about half the cost — https://aihot.news/items/cmud9s4pn001groggnqsackao [11] arXiv math submissions surge 33.5% in 2026 — https://aihot.news/items/cmud9c72203xoroju18d2335h [16] PixVerse R2 real-time world model — https://aihot.news/items/cmud8xugv03oirojupk6i91cu [20] Snorkel AI closes $350M Series E, valuation rises to $3.5 billion — https://aihot.news/items/cmud8okok03auroc2etpdsbii [22] Qualcomm CEO Amon: The agent era isn't about letting AI do everything, but making technology more practical — https://aihot.news/items/cmud