2026
21 篇文章AlphaGenome Atlas:把 90 億個 DNA 變異變成可查詢的預測地圖
Google DeepMind 推出 AlphaGenome Atlas,涵蓋人類基因組中所有 90 億個單核苷酸變異的分子效應預測,並提供 AVI 分數與特徵歸因,協助研究人員快速排序與解讀變異。本文從產品建造者角度,解析這份 1PB 資料集的設計、應用與限制。
閱讀文章 ↗Gemini 3.8 Flash 價格 2027 年翻倍:Cyber 版 CWE-Bench 47.2%
第三個 Flash 六週內上線:入門價 2027 年翻倍,Flash Cyber 以 CWE-Bench 47.2% 逼近旗艦,兩小時找到重大漏洞。
閱讀文章 ↗AI Mode 旅遊規劃升級:追蹤票價、用點數訂房、直接完成飯店預訂
Google 在 AI Mode 加入三項旅遊功能:機票價格追蹤、顯示點數或哩程成本,以及在對話中直接預訂飯店。本文整理實際可用範圍與限制,供產品開發者參考。
閱讀文章 ↗把試算表變成互動小工具:Google Sheets canvas 的實用切入點
Google 推出 Sheets canvas,讓使用者用自然語言把試算表資料轉成可互動的儀表板、座位表或學習追蹤器,並與原始資料即時同步。
閱讀文章 ↗Hassabis 轉任 DeepMind 主席,Jeff Dean 離職創辦 Discovery Loop
2026 年 8 月 5 日 Google 重組 DeepMind:Hassabis 轉任 GDM 主席兼 Alphabet 首席科學家,Kavukcuoglu 接掌日常營運,Jeff Dean 在職 27 年後離職創辦 Discovery Loop。
閱讀文章 ↗Gemini Robotics 2 進場:從腳到指尖的機器人全身智能
2026 年 7 月 30 日 Google DeepMind 發表 Gemini Robotics 2,以 VLA、ER 與 On-Device 三個模型讓人形機器人全身控制、綁繩結、組隊合作,並公布 ASIMOV-Agentic 安全基準。
閱讀文章 ↗Google Images 25 週年:從文字搜尋到視覺生成
Google Images 迎來 25 週年,推出可瀏覽的首頁與 AI Overviews 影像生成功能。本文回顧視覺搜尋的演進,並探討對產品建構者的啟示。
閱讀文章 ↗DeepMind GenCeption:影片生成模型就是通用視覺學習器
Google DeepMind 提出 GenCeption,把文字轉影片擴散模型改造成可用文字指令驅動的通用前饋感知模型,在深度、分割、相機姿態等任務達到 SOTA,所需訓練數據最多減少 500 倍。
閱讀文章 ↗Nano Banana 2 Lite 登場:4 秒生圖、單張 3.4 美分的高速影像路線
2026 年 6 月 30 日,Google 發表 Nano Banana 2 Lite:約 4 秒完成文字生圖、每張 0.034 美元,同場推出 Gemini Omni Flash 影片模型,用「快出圖、再動圖」的串接工作流把低成本高速生成推向主流。
閱讀文章 ↗全端 AI 的真正意義:從 TPU 到 Gmail 的整合思維
Google 的 Richard Seroter 解釋全端 AI:從 TPU 運算、Gemini 模型、協調平台到日常介面四層全部自研整合,減少跨團隊交接與拼裝成本,讓開發者用更少的整合工打造可靠且具成本效益的產品。
閱讀文章 ↗Gemini 3.5 Flash 內建電腦操作:OSWorld 78.4 追平 Sonnet
Google DeepMind 把電腦操作能力直接內建到 Gemini 3.5 Flash,OSWorld 拿下 78.4 分追平 Claude Sonnet 4.6,並配上對抗訓練與兩項企業級安全防護,透過 Gemini API 與企業代理平臺開放。
閱讀文章 ↗諾貝爾獎得主 John Jumper 宣布離開 DeepMind 轉戰 Anthropic
2024 年諾貝爾化學獎共同得主 John Jumper 宣布結束近九年的 DeepMind 生涯轉投 Anthropic,同週 Transformer 論文共同作者 Noam Shazeer 也離開 Google 前往 OpenAI,前沿實驗室的人才爭奪白熱化。
閱讀文章 ↗Google 開源 DiffusionGemma:擴散式生成快 4 倍,單卡 H100 破千 token/秒
Google 於 6 月 10 日開源 DiffusionGemma:26B-A4B 擴散語言模型,Apache 2.0 釋出,單卡 H100 每秒生成逾千 token、比自回歸快 4 倍,但多數基準仍輸 Gemma 4,官方明言最適合本地與低併發場景。
閱讀文章 ↗David Silver 創辦 Ineffable Intelligence:11 億美元豪賭「超級學習者」
2026 年 4 月,DeepMind 強化學習主管 David Silver 創辦的 Ineffable Intelligence 以 51 億美元估值募得 11 億美元,目標是打造純靠強化學習試錯、不依賴人類資料的「超級學習者」。本文解析資金、技術賭注與倫敦 AI 樞紐的崛起。
閱讀文章 ↗Gemma 4 開源發布:Apache 2.0、MoE 與 256K 上下文
2026 年 4 月 2 日 Google DeepMind 發布 Gemma 4:全系列改用 Apache 2.0 授權,四種尺寸從 2B 端側到 31B Dense,支援 140+ 語言、256K 上下文與影像語音輸入,31B 躍上 Arena 開源模型第三名。
閱讀文章 ↗Google Lyria 3 Pro 登場:AI 音樂生成從 30 秒片段走向三分鐘完整曲目
3 月 25 日 Google DeepMind 推出 Lyria 3 Pro,可生成長達 3 分鐘、含完整段落結構的曲目,較 Lyria 3 的 30 秒大幅躍進。本文看它進駐 Gemini、Vids、Vertex AI 與 ProducerAI 的布局、SynthID 浮水印與訓練資料立場。
閱讀文章 ↗Gemini Embedding 2 公開預覽:文字、影像、音訊、影片共用一個向量空間
Google DeepMind 於 2026 年 3 月 10 日推出 Gemini Embedding 2 公開預覽:第一個原生多模態嵌入模型,把文字、影像、影片、音訊與 PDF 映射到單一 3072 維向量空間,支援 Matryoshka 截斷。本文解析輸入限制、成本槓桿與對 RAG 管線的實際影響。
閱讀文章 ↗Gemini 3 Deep Think 更新:從奧賽金牌走向研究級數學
2026 年 2 月 11 日,Google DeepMind 更新 Gemini 3 Deep Think:IMO-ProofBench Advanced 最高可達 90%,數學代理人 Aletheia 能承認失敗,並在 18 個研究問題上產出論文。本文解析推理縮放、Erdős 猜想實績與取得方式。
閱讀文章 ↗Waymo World Model 登場:用 Genie 3 生成超擬真自駕模擬世界
2026 年 2 月 6 日,Waymo 發表建構在 Genie 3 之上的 World Model,同時生成相機影像與光達點雲,支援路線重模擬與罕見情境測試。本文解析三種控制機制、預訓練的槓桿,與對自駕安全驗證的意義。
閱讀文章 ↗Google Project Genie 上線:世界模型即時生成可玩世界,遊戲股應聲重挫
Google 於 1 月 29 日向 AI Ultra 訂閱者開放 Project Genie,首個基於 Genie 3 世界模型的公開產品,可從文字與圖片即時生成可互動 3D 世界。次日 Unity 一度暴跌近 24%,遊戲股全面下挫。
閱讀文章 ↗Boston Dynamics 量產版 Atlas 現身 CES:全電動人形機器人開始出貨
CES 2026 上 Boston Dynamics 發表全電動量產版 Atlas:56 自由度、可舉 50 公斤、自主換電池,2026 年部署名額全數給了現代汽車與 Google DeepMind,2027 年初才開放外部客戶。人形機器人從示範影片走進工廠。
閱讀文章 ↗
2026
21 ARTICLESAlphaGenome Atlas: A Precomputed Map of 9 Billion DNA Variants Changes How We Prioritize Genetic Research
Google DeepMind's AlphaGenome Atlas turns variant interpretation from a per-variant query into a precomputed, rankable dataset.
READ POST ↗Gemini 3.8 Flash Intro Price Doubles on Jan 1: 54.9% HLE
Gemini 3.8 Flash intro pricing doubles on Jan 1, 2027; Flash Cyber scores 47.2% on CWE-Bench and found a critical vulnerability in under two hours.
READ POST ↗Google AI Mode Adds Flight Price Tracking, Points/Miles Rates, and Hotel Booking
Google's AI Mode in Search now supports flight price tracking, points/miles cost display, and hotel booking. Here's what changed and what it means for product builders.
READ POST ↗Google Sheets Canvas: Turning Spreadsheets into Interactive Mini-Apps
Google's new Sheets canvas turns spreadsheet data into interactive mini-apps with natural language prompts. Learn how it works, use cases, and limitations.
READ POST ↗Hassabis Moves to Chair as Jeff Dean Leaves Google DeepMind
Google reshuffled DeepMind on August 5: Hassabis becomes chair and Alphabet chief scientist, Kavukcuoglu takes the day job, Jeff Dean exits for Discovery Loop.
READ POST ↗Gemini Robotics 2 Gives Robots Whole-Body Intelligence
Google DeepMind's Gemini Robotics 2 adds whole-body humanoid control, fine dexterity, multi-robot teamwork, and an ASIMOV-Agentic safety benchmark to its robotics stack.
READ POST ↗Google Images at 25: From Text Links to AI Image Generation in Search
Google Images turns 25 with a new browseable homepage and AI image generation in AI Overviews. Here's what product builders need to know.
READ POST ↗DeepMind GenCeption: Video Generation as Vision Pretraining
DeepMind's GenCeption turns a text-to-video diffusion backbone into a text-steered perception model, hitting SOTA on depth, segmentation, and pose with up to 500x less data.
READ POST ↗Google's Nano Banana 2 Lite: 4-Second, 3.4-Cent Images
On June 30, 2026, Google launched Nano Banana 2 Lite — images in about 4 seconds at $0.034 each — plus the Gemini Omni Flash video model, pairing fast stills with cheap video.
READ POST ↗Full-Stack AI Explained: What It Means and How to Get Started
Google's Richard Seroter explains the full-stack AI approach, its benefits, and three ways to start building with it.
READ POST ↗Gemini 3.5 Flash Gets Built-in Computer Use
Google built computer use directly into Gemini 3.5 Flash: 78.4 on OSWorld ties Claude Sonnet 4.6, with adversarial training and enterprise guardrails, available via the Gemini API.
READ POST ↗Nobel Laureate John Jumper Leaves DeepMind for Anthropic
AlphaFold co-creator John Jumper, a 2024 Nobel chemistry laureate, is leaving DeepMind for Anthropic after nearly nine years, as Noam Shazeer exits Google for OpenAI.
READ POST ↗Google Open-Sources DiffusionGemma: 4x Faster Generation
Google open-sources DiffusionGemma: an Apache 2.0 26B-A4B diffusion LLM hitting 1000+ tokens/sec on one H100 — 4x faster than autoregressive, at a quality cost vs Gemma 4.
READ POST ↗Ineffable Intelligence: $1.1B to Learn Without Human Data
David Silver's lab Ineffable Intelligence raised $1.1B at a $5.1B valuation to build a 'superlearner' trained by reinforcement learning alone — no human data. Inside the round.
READ POST ↗Gemma 4 Ships Under Apache 2.0: Google's Open Model Reset
Google DeepMind released Gemma 4 on April 2, 2026 under Apache 2.0 — four sizes from 2B edge to 31B dense, 140+ languages, 256K context, Arena top-3. What it changes for builders.
READ POST ↗Google Lyria 3 Pro: AI Music Moves From Clips to Full Tracks
Google DeepMind's Lyria 3 Pro generates full three-minute songs with verse-chorus structure, rolling into Gemini, Vids, Vertex AI, and ProducerAI with SynthID watermarking.
READ POST ↗Gemini Embedding 2: One Vector Space for Every Modality
Google's Gemini Embedding 2 hits Public Preview on March 10, 2026 — the first natively multimodal embedding model, mapping text, images, video, audio, and PDFs into one 3072-dim vector space.
READ POST ↗Gemini 3 Deep Think: From Olympiad Gold to Research Math
Google DeepMind's February 11, 2026 update pushes Gemini 3 Deep Think into research territory: up to 90% on IMO-ProofBench Advanced, a math agent that admits failure, and papers across 18 problems.
READ POST ↗Waymo's World Model: Genie 3 Powers Driving Simulation
Waymo's World Model, built on Google DeepMind's Genie 3, generates camera and lidar data for hyper-realistic driving simulation — re-simulating routes and testing rare events at scale.
READ POST ↗Project Genie: Google's World Model Rattles Gaming Stocks
Google opened Project Genie to AI Ultra users on Jan. 29 — the first public product built on Genie 3, turning prompts into playable 3D worlds in real time. Unity plunged nearly 24% the next day.
READ POST ↗Boston Dynamics Electric Atlas Goes Into Production at CES
At CES 2026, Boston Dynamics unveiled the production all-electric Atlas: 56 DoF, 50 kg payload, self-swapping batteries — every 2026 deployment already committed to Hyundai and Google DeepMind.
READ POST ↗