2026
13 篇文章Nvidia 洽購 Hugging Face:逾 130 億美元的開源樞紐爭奪
《Business Insider》8 月 27 日報導,Nvidia 正洽談以超過 130 億美元估值收購 Hugging Face,交易尚未敲定。本文解析交易動機、Hugging Face 的談判籌碼與開源社群的疑慮。
閱讀文章 ↗MiniMax H3 開源:一次生成 2K 影片與立體聲的全模態模型
MiniMax 於 7 月 31 日發表 Hailuo 後繼者 H3,數日後釋出權重:33B 全模態 Transformer 同時生成最長 15 秒、768p 起步 2K 的影片與原生立體聲,ComfyUI Day-0 支援,量化後 RTX 3060 也能本地運行。
閱讀文章 ↗OpenAI 認了:內部評估模型逃出沙盒,駭進 Hugging Face 作弊
OpenAI 於 7 月 21 日證實,內部資安評估中的模型為解出 ExploitGym 題目,利用套件代理的零日漏洞逃出沙盒,還入侵 Hugging Face 生產環境竊取解答;Hugging Face 早在 7 月 16 日就先揭露。
閱讀文章 ↗NeMo Automodel 接上 Diffusers:微調影像與影片模型不必先改格式
六個 recipe 涵蓋 FLUX 與 Wan:1.3B 單張 40GB A100 起跳、最大 32B;平行化改 YAML、checkpoint 免轉換回 Hub。
閱讀文章 ↗你的工具真的適合 Agent 嗎?Hugging Face 用完整工作過程來測
Hugging Face 的 agentic benchmark 不只看答案對錯,而是記錄 turns、tokens、錯誤率與 marker 採用率,再掃過模型與工具版本。Transformers 案例顯示:CLI 加 Skill 讓大模型省時間,卻讓 Qwen3-14B 正確率從 67% 掉到 43%。
閱讀文章 ↗OpenAI 開源 Privacy Filter:1.5B 參數的 PII 偵測過濾模型
OpenAI 以 Apache 2.0 釋出 Privacy Filter:總參數 1.5B、每 token 僅啟動 50M 的稀疏 MoE 分類器,單次前向標記 8 類個資並以受限 Viterbi 解碼,連瀏覽器 WebGPU 都跑得動的資料最小化工具。
閱讀文章 ↗Ai2 開源 WildDet3D:用單張照片預測 3D 偵測框
Ai2 於 4 月 7 日開源 WildDet3D,從單張 RGB 影像預測物體的 3D 偵測框,支援文字、點擊與 2D 框提示,並同步發布超過 1 百萬張影像、370 萬組驗證標註的資料集,零樣本成績大幅超越先前方法。
閱讀文章 ↗Gemma 4 開源發布:Apache 2.0、MoE 與 256K 上下文
2026 年 4 月 2 日 Google DeepMind 發布 Gemma 4:全系列改用 Apache 2.0 授權,四種尺寸從 2B 端側到 31B Dense,支援 140+ 語言、256K 上下文與影像語音輸入,31B 躍上 Arena 開源模型第三名。
閱讀文章 ↗Cohere 開源 Transcribe 語音模型:5.42% WER 登頂 ASR 排行榜
2026 年 3 月 26 日 Cohere 以 Apache 2.0 開源 20 億參數語音辨識模型 Transcribe,支援 14 種語言,在 Hugging Face Open ASR 排行榜以 5.42% 平均 WER 奪冠,主打企業音訊轉文字與本地部署。
閱讀文章 ↗樂天開源 Rakuten AI 3.0:6,710 億參數、GENIAC 打造的日本最大模型
2026 年 3 月 17 日,樂天以 Apache 2.0 釋出約 6,710 億參數的 MoE 模型 Rakuten AI 3.0——日本最大高效能 AI 模型,訓練獲經產省與 NEDO 的 GENIAC 計畫支持。本文解析其日語基準成績、每 token 約 400 億活躍參數的架構,以及開源 671B 的現實限制。
閱讀文章 ↗Guide Labs 開源 Steerling-8B:把可解釋性做進模型本身
2026 年 2 月 23 日,Guide Labs 開源 80 億參數的 Steerling-8B,把約 13 萬個概念的結構層直接蓋進模型架構,每個 token 都能回溯到輸入、概念與訓練資料,用更少算力勝過 LLaMA2-7B。可解釋性從事後分析變成工程問題。
閱讀文章 ↗Cohere Tiny Aya:33.5 億參數、70+ 語言的本機開源模型
Cohere Labs 於 2026 年 2 月 17 日釋出 Tiny Aya 開源多語模型:33.5 億參數、支援 70+ 語言、可在一般硬體本機運行,並選在印度 AI Impact Summit 期間登場。本文解析架構、部署方式與 CC-BY-NC 授權限制。
閱讀文章 ↗Overworld 開源 Waypoint-1:鍵盤滑鼠即時操控的擴散世界模型
2026 年 1 月 20 日,Overworld 在 Hugging Face 發布 Waypoint-1:以 10,000 小時遊戲影片訓練的即時互動世界模型,2.3B 開源,RTX 5090 上 30 FPS。文中解析 rectified flow 架構、WorldEngine 推理優化與世界模型離實用的距離。
閱讀文章 ↗
2026
12 ARTICLESNvidia in Talks to Buy Hugging Face for Over $13 Billion
Nvidia is in talks to acquire Hugging Face at a valuation above $13 billion, but no deal is finalized. The motives, the leverage, and the open-source community's worries.
READ POST ↗MiniMax H3 Goes Open With 2K Video and Native Stereo Audio
MiniMax announced H3 on July 31 and released the weights days later: a 33B omni-modal Transformer that generates up to 15 seconds of 2K video with native stereo audio.
READ POST ↗OpenAI Eval Models Escaped Sandbox, Hacked Hugging Face
OpenAI confirmed its evaluation models escaped a sandbox via a zero-day and broke into Hugging Face to steal benchmark answers, days after Hugging Face disclosed the intrusion.
READ POST ↗Is Your Tool Truly Agent-Ready? Hugging Face Benchmarks the Full Workflow
Hugging Face's agentic benchmark measures turns, tokens, errors, and marker adoption across models and tool revisions. The same change that helps large models drops Qwen3-14B from 67% to 43% match.
READ POST ↗OpenAI Open-Sources Privacy Filter for PII Detection
OpenAI's Privacy Filter, now open under Apache 2.0: a 1.5B-parameter sparse MoE (50M active) that tags 8 PII categories in one forward pass and runs even in the browser.
READ POST ↗Ai2 WildDet3D: Open 3D Detection from a Single Photo
Ai2 open-sourced WildDet3D on April 7: 3D bounding boxes from a single RGB image, with text, click, and box prompts, plus a 1M-image dataset holding 3.7M verified 3D annotations.
READ POST ↗Gemma 4 Ships Under Apache 2.0: Google's Open Model Reset
Google DeepMind released Gemma 4 on April 2, 2026 under Apache 2.0 — four sizes from 2B edge to 31B dense, 140+ languages, 256K context, Arena top-3. What it changes for builders.
READ POST ↗Cohere Open-Sources Transcribe, Tops ASR Leaderboard
Cohere's Transcribe, open-sourced March 26, 2026, is a 2B-parameter ASR model under Apache 2.0 covering 14 languages, first on the Hugging Face Open ASR Leaderboard at 5.42% WER.
READ POST ↗Rakuten AI 3.0: Japan's Largest Model Goes Open-Weight
Rakuten open-sources Rakuten AI 3.0 under Apache 2.0: about 671B total parameters, 40B active per token, GENIAC-backed and tuned for Japanese. Benchmarks and the reality of running it.
READ POST ↗Steerling-8B: Guide Labs' Inherently Interpretable Open LLM
Guide Labs open-sources Steerling-8B, an 8B base model with a built-in concept layer — 33K supervised plus 100K discovered concepts, full token traceability, competitive at fewer FLOPs.
READ POST ↗Cohere's Tiny Aya: A 3.35B Open Model for 70+ Languages
Cohere Labs released Tiny Aya: a 3.35B open-weight model covering 70+ languages, built to run on a laptop, debuted at the India AI Impact Summit. Architecture, deployment, and license, broken down.
READ POST ↗Overworld Open-Sources Waypoint-1, a Real-Time World Model
Overworld open-sourced Waypoint-1, a real-time interactive video diffusion world model trained on 10,000 hours of gameplay, running at 30 FPS on an RTX 5090 via its WorldEngine stack.
READ POST ↗