2026
87 篇文章當瀏覽器內建 AI:Mistral 與 Mozilla 把主權與隱私放進 Firefox Smart Window
Mistral 與 Mozilla 合作,讓 Firefox Smart Window 由 Mistral 模型驅動,主打零資料保留與在地語言微調。
閱讀文章 ↗Hermes Agent:記憶留在本機、會自己長出技能的開源助理
Nous Research 的開源 Hermes Agent 把記憶留在本機、會自動生成技能,2026 年 2 月推出後登上 OpenRouter 累積用量第一。本文整理它的設計賭注、Desktop 公測與 15 億美元估值融資,並看用量數字該怎麼讀。
閱讀文章 ↗當實驗數據多到看不完:SAM 3 與 DINOv3 在 Genesis Mission 裡的角色
美國國家實驗室把 Meta 開源視覺模型微調後部署到 300 張 A100,將一個月的標註工作壓到 15 分鐘。
閱讀文章 ↗Ox Alpha 就是 GLM-5.3-Flash:MIT 開源,OpenRouter 週流量佔 31%
智譜揭曉匿名冠軍 Ox Alpha 就是 GLM-5.3-Flash 並以 MIT 授權開源:預覽期在十萬顆國產晶片上處理 62 兆 tokens,OpenRouter 週流量佔 31%、登頂程式碼模型榜,消息公布當日港股大漲逾 12%。
閱讀文章 ↗Mistral Forge:不微調、不 RAG,從零訓練企業專屬模型的另類路線
Mistral 在 NVIDIA GTC 2026 推出 Forge 平台,讓企業與政府用自有資料從零訓練客製模型,隨平台推出 Mistral Small 4。本文解析這條有別於 OpenAI/Anthropic 微調路線的策略、前進部署工程師模式,以及誰真的需要自訓模型。
閱讀文章 ↗加州年齡驗證法修法豁免開源系統:AB 1856 解讀
加州參議院以 39 票對 0 票通過 AB 1856 修正案,把 GPL、MIT、BSD、Apache 授權的作業系統排除在年齡驗證義務外,全案已送交州長紐森簽署。本文解讀條文與影響。
閱讀文章 ↗Nvidia 洽購 Hugging Face:逾 130 億美元的開源樞紐爭奪
《Business Insider》8 月 27 日報導,Nvidia 正洽談以超過 130 億美元估值收購 Hugging Face,交易尚未敲定。本文解析交易動機、Hugging Face 的談判籌碼與開源社群的疑慮。
閱讀文章 ↗Haiku R1/beta6 出爐:25 週年後的回歸之作
距離 beta5 約兩年,Haiku 於 25 週年後一週發布 R1/beta6:解決逾 530 張票券,Firefox 正式品牌進駐、QEMU 獲 NVMM 硬體虛擬化支援,git status 熱快取案例從 15 秒降到 2.5 秒。
閱讀文章 ↗AWS 收購 DuckLabs:DuckDB 母公司加入 Amazon,開源授權不變
2026 年 8 月 26 日,AWS 宣布收購 DuckDB、DuckLake 與 Quack 背後的阿姆斯特丹公司 DuckLabs;產品維持 MIT 授權,DuckDB 基金會繼續持有 IP,三十多人團隊原地留任。
閱讀文章 ↗DuckDB v2.0 預覽:新解析器、新儲存格式與伺服器模式
DuckDB 團隊於 8 月 17 日預覽今秋的 v2.0「Cyanoptera」:原生伺服器模式、PEG 新解析器、儲存格式 2.0、非同步 I/O、觸發器,遞迴查詢快了約 40 倍。
閱讀文章 ↗Nari Labs 把 Qwen3-TTS 壓進 50 毫秒內開口
Nari Labs 於 8 月 19 日公開 Qwen3-TTS 1.7B 優化成果:單張 H100 上達到每秒 10 次請求、p95 首音延遲低於 50 毫秒,並開源整套服務實作與基準測試。
閱讀文章 ↗Go 1.27 登場:泛型方法、json v2 與後量子加密
Go 1.27 於 8 月 19 日發布:方法終於能宣告型別參數,encoding/json 底層改由 v2 實作驅動,ML-DSA 後量子簽章進入 TLS,小物件配置成本最多降 30%。
閱讀文章 ↗DeepSeek 開源 Agent Harness:一切皆外掛的開發者預覽
2026 年 8 月 13 日,DeepSeek 以 MIT 授權公開 DeepSeek Harness 開發者預覽:Cordis 核心做到一切皆外掛,附錄影重播的會話日誌與四種執行模式,原始碼與外掛生態同步開放。
閱讀文章 ↗Firefox 成最後防線:uBlock Origin 的 MV2 保衛戰
Edge 開始停用 Manifest V2 後,Firefox 成為唯一仍支援完整版 uBlock Origin 的主流瀏覽器;Mozilla 表態支援不會中斷,獨立引擎成為最清晰的差異化。
閱讀文章 ↗Mojo 1.0 正式登場:給 AI 時代的穩定系統語言
Modular 於 2026 年 8 月 11 日宣布 Mojo 語言正式抵達 1.0:以穩定為承諾的 1.x 系列,讓開發者能長期建置;標準庫開源以來已有近 200 位貢獻者與超過 1,100 個 PR。
閱讀文章 ↗美國能源部啟動 Genesis 開放模型計畫:開放權重做科學
美國能源部在 Genesis Mission 之下啟動開放權重科學模型計畫,與 Arcee AI 合作開發首個模型 Genesis-Science-1,並開放全國貢獻門戶,首輪申請 8 月 14 日截止。
閱讀文章 ↗Oracle 劃紅線:AI 生成的內容不得進入 OpenJDK
OpenJDK 發布生成式 AI 臨時政策:貢獻不得包含 LLM 或擴散模型生成的任何內容,理由是審查負擔、安全與 IP 風險;GitHub PR 也將加上合規勾選框。
閱讀文章 ↗MiniMax H3 開源:一次生成 2K 影片與立體聲的全模態模型
MiniMax 於 7 月 31 日發表 Hailuo 後繼者 H3,數日後釋出權重:33B 全模態 Transformer 同時生成最長 15 秒、768p 起步 2K 的影片與原生立體聲,ComfyUI Day-0 支援,量化後 RTX 3060 也能本地運行。
閱讀文章 ↗FFmpeg 9.0「Lei」發布:swscale 重寫與更安全的預設
2026 年 8 月 4 日,FFmpeg 9.0「Lei」亮相:七個函式庫同步升版打破 ABI、swscale 多年重寫落地、動畫 WebP 解碼補上、TLS 憑證驗證預設開啟,開發主場遷往自家 Forgejo。
閱讀文章 ↗在 8GB Mac 上跑 26B 模型:TurboFieldfare 的 SSD 串流推理
開源專案 TurboFieldfare 以 Swift 與 Metal 打造推理引擎:只常駐約 2GB 記憶體,把 Gemma 4 26B-A4B 的 MoE 專家權重從 SSD 逐 token 串流載入,在 8GB MacBook Air 上跑出每秒 5 到 6 個 token。
閱讀文章 ↗Keychron 開源滑鼠韌體 ZGM:把 QMK 的開源文化帶進遊戲滑鼠
Keychron 於 7 月底公開 ZGM——以 Zephyr RTOS 為基礎的開源遊戲滑鼠韌體專案,採 GPL-3.0 授權,要把 QMK 與 ZMK 建立的開源輸入裝置文化延伸到滑鼠。目前倉庫仍屬早期架構階段,原始碼尚未到位。
閱讀文章 ↗Debian 公開決議:LLM 貢獻的去留之爭
Debian 於 7 月 23 日起為「LLM 使用」啟動一般決議討論:八個提案從全面禁止到有條件開放,8 月 15 日開始投票,將為大型開源專案的 AI 貢獻治理立下先例。
閱讀文章 ↗印度下令 GitHub 下架 Bitchat:藍牙通訊遇上審查
印度內政部網路犯罪協調中心 I4C 於 2026 年 7 月 23 日發函,要求 GitHub 在三小時內停用 Jack Dorsey 的藍牙網狀通訊 app Bitchat 的三個儲存庫。通知書由 Dorsey 公開,事件凸顯審查權力伸進開源程式碼的新界線。
閱讀文章 ↗機場一組密碼清空手機:GrapheneOS 用戶遭聯邦起訴
亞特蘭大男子 Sam Tunick 在機場邊檢輸入 GrapheneOS 壓力碼清空手機,遭美國司法部以「為防止扣押而毀損財產」起訴,安全功能本身首次成為起訴焦點。
閱讀文章 ↗Ruff 0.16:預設規則從59條拉到413條
Ruff v0.16.0 於 2026 年 7 月 23 日發布:預設啟用規則從 59 條增至 413 條,新增 Markdown 內 Python 程式碼格式化與 ruff: ignore 抑制語法,並有 18 條意見型規則退出預設集。本文整理版本重點與升級注意事項。
閱讀文章 ↗Block 開源 Buzz:團隊聊天、AI 代理與 Git 託管共用一條事件流
Block 於 2026 年 7 月 21 日開源工作區 Buzz,把團隊聊天、AI 代理人與 Git 託管放進同一套簽名事件系統,明言要減少對 Slack 與 GitHub 的依賴。
閱讀文章 ↗Firefox 153 推出:Vulkan 影片解碼與 JPEG XL 終於上船
Firefox 153 開始向 Release 頻道推出,帶來 Vulkan 硬體影片解碼、Firefox Labs 實驗性 JPEG XL 支援與 Windows HDR 播放,並把本機網路存取改為預設需詢問。
閱讀文章 ↗Grok Build 開源:Coding Agent 最值得讀的是 Harness,不是 UI
xAI 把 Grok Build 的完整 harness 開源:agent loop、工具、terminal UI 與擴充系統四大塊一次公開。本文整理開源範圍、「原始碼才是 definitive reference」的論點,以及 local-first 設定對 builder 的意義。
閱讀文章 ↗Reflection 與 Nebius 簽 10 億美元算力合約
開源模型公司 Reflection AI 與 GPU 雲 Nebius 簽下 10 億美元算力合約,取得 NVIDIA 最新晶片;這是它三週內第二筆超大單,顯示前沿算力已是新創最貴的入場券。
閱讀文章 ↗LangChain 攜 NVIDIA 推 NemoClaw:治理優先的 Deep Agents 藍圖
LangChain 與 NVIDIA 推出 NemoClaw Deep Agents 藍圖,以 Nemotron 3 Ultra 開放模型、Deep Agents Code 代理框架與 OpenShell 沙盒構成可治理的三層堆疊,評測成本僅最接近對手的十分之一。
閱讀文章 ↗Bun 1.4 改寫成 Rust:64 個 Claude 代理 11 天完成的大搬遷
JavaScript 執行環境 Bun 釋出 1.4.0,首個以 Rust 重寫的版本:Claude Code 動態工作流程搭配 Claude Fable 5,以 64 個並行代理在 11 天內移植約 53.5 萬行 Zig 程式碼,API 成本約 16.5 萬美元。
閱讀文章 ↗Mistral 開源 Leanstral 1.5:miniF2F 滿分、每題 4 美元的 Lean 證明
Mistral 於 2026 年 7 月 2 日釋出 Apache-2.0 授權的 Leanstral 1.5:119B 總參數、約 6B 活躍,在 miniF2F 拿下滿分、PutnamBench 解出 587 題,把 Lean 4 證明成本壓到每題約 4 美元。
閱讀文章 ↗Meta Brain2Qwerty v2:腦波轉文字達 61%
Meta 發表 Brain2Qwerty v2:以腦磁圖搭配端到端深度學習與微調語言模型,將非侵入式腦波解碼文字的字準確率從過去的 8% 推升至 61%,並開源 v1 與 v2 的訓練程式碼。
閱讀文章 ↗Patch the Planet:用 AI 幫開源維護者補洞,而不是增加負擔
OpenAI 推出 Patch the Planet 計畫,結合 AI 與人工審查,協助開源專案修補漏洞。本文整理其運作方式、初步成果,以及對維護者與產品開發者的啟示。
閱讀文章 ↗GLM-5.2 開源發布:百萬上下文長任務表現緊咬 Opus 4.8
Z.AI 於 2026 年 6 月 17 日開源 753B 參數的 GLM-5.2:MIT 授權、穩定 1M token 上下文,長任務與程式基準緊追 Claude Opus 4.8,成為開源陣營排名最高的模型。
閱讀文章 ↗Baseten 傳以 130 億美元估值募 15 億美元:五個月估值跳 160%
WSJ 報導 Baseten 接近完成 15 億美元輪次,估值最高 130 億美元,距 1 月以 50 億美元估值完成的 3 億美元 Series E 僅五個月。本文拆解雙軌定價設計、開源推理路由生意,與推理淘金熱背後的毛利風險。
閱讀文章 ↗Sarvam AI 估值 15 億美元:HCLTech 領投下的印度 AI 獨角獸
6 月 15 日,班加羅爾的 Sarvam AI 以 15 億美元估值募得 2.34 億美元,成為印度最新 AI 獨角獸。HCLTech 出資 1.5 億美元領投,B 輪目標 3 億美元。本文解析這筆主權 AI 交易的結構,與對印度模型生態的意義。
閱讀文章 ↗Vercel 開源 eve:agent 框架終於有了一個標準形狀
Vercel 開源 eve agent 框架:目錄即 agent、檔案即能力,durable session、sandbox、human-in-the-loop 與 evals 全部預裝。本文拆解其檔案系統設計、安全邊界、Vercel 內部一百個 agent 的自述戰績,以及深度綁定 Vercel 生態的取捨。
閱讀文章 ↗Mistral 傳募 30 億歐元、估值 200 億:歐洲主權 AI 的加注與鴻溝
彭博 6 月 12 日報導 Mistral 洽談以約 200 億歐元估值募集 30 億歐元,較去年 9 月的 117 億估值近乎翻倍。本文拆解這家歐洲主權 AI 旗手的籌碼:巴黎資料中心、國防合作,以及與美國實驗室之間巨大的資金鴻溝。
閱讀文章 ↗Google 開源 DiffusionGemma:擴散式生成快 4 倍,單卡 H100 破千 token/秒
Google 於 6 月 10 日開源 DiffusionGemma:26B-A4B 擴散語言模型,Apache 2.0 釋出,單卡 H100 每秒生成逾千 token、比自回歸快 4 倍,但多數基準仍輸 Gemma 4,官方明言最適合本地與低併發場景。
閱讀文章 ↗TensorZero 收攤:730 萬美元種子輪的開源 LLM 閘道一夜封存
6 月 13 日,開源 Rust LLM 閘道 TensorZero 的 GitHub repo 無預警轉為封存。共同創辦人證實公司收攤:兩年半僅花用不到一半的 730 萬美元種子輪資金,剩餘資本退還投資人,程式碼以 Apache 2.0 保留但不再維護。
閱讀文章 ↗Hello Robot Stretch 4 開賣:3 萬美元家用機器人的務實路線
Hello Robot 第四代家用助理機器人 Stretch 4 售價 29,950 美元,首波 200–300 台售罄。它刻意保留人類在環路操控,四肢癱瘓使用者已靠語音 App 用它完成日常任務。
閱讀文章 ↗Miasma 蠕蟲再襲 Microsoft:73 個儲存庫停用,AI 編碼代理成靶
被劫持的帳號六月五日把惡意提交推入 Azure durabletask 儲存庫,植入 .claude、.gemini 與 .cursor 設定檔,開檔即執行竊憑證 payload;GitHub 在 105 秒內停用 73 個儲存庫,官方 functions-action 工作流應聲斷裂。
閱讀文章 ↗微軟開源 ASSERT:把文字規格變成 AI 行為測試套件
2026 年 6 月 2 日,微軟開源 ASSERT 框架:開發者以自然語言描述代理應有行為,它自動生成測試情境、用 LLM 評審計分,輸出傷害與權衡兩類指標,並可掛進 CI 做回歸把關。
閱讀文章 ↗Runtime(YC P26)上線:把沙盒化 coding agents 開放給整個團隊
YC P26 團隊 Runtime 於 5 月 21 日在 Hacker News 發布:讓非工程同事也能安全使用 Claude Code 與 Codex,以環境快照、沙盒編排、密鑰代理與 RBAC 控制風險,核心開源、不抽 token 加價。
閱讀文章 ↗NanoClaw 走紅之後:拒絕 2,000 萬美元收購,把 agent 關進容器
NanoClaw 為每個 agent session 配一個獨立 Docker 容器隔離執行,走紅後拒絕約 2,000 萬美元收購,改募 Valley Capital 領投、Docker 與 Vercel 參與的 1,200 萬美元種子輪,企業客戶已上門。
閱讀文章 ↗Google ERA 登上 Nature:Gemini 寫出專家級科學程式碼
Google 的 ERA 用 Gemini 撰寫並最佳化科學程式碼,5 月 19 日登上 Nature:六個領域達專家級水準,五份新應用論文與開源程式碼同步釋出,並透過 Gemini for Science 逐步開放。
閱讀文章 ↗NHS 畏懼 AI 漏洞挖掘大舉關閉開源庫,GDS 發布指引唱反調
AI 尋找漏洞的能力躍升後,NHS England 內部指示關閉幾乎所有開源儲存庫;GDS 與 DSIT 於 5 月 14 日發布指引唱反調:預設保持開放,關庫無助修補根本弱點,只會增加成本。
閱讀文章 ↗DuckDB 推出 Quack 協定:內嵌資料庫補上主從架構缺口
DuckDB 釋出實驗性 Quack 遠端協定,讓 DuckDB 同時扮演 client 與 server:走 HTTP、查詢單次往返,6,000 萬列 4.94 秒傳完,正式版將隨 v2.0 於 2026 年秋季推出。
閱讀文章 ↗TanStack npm 供應鏈攻擊解析:三個漏洞串出 84 個惡意版本
2026 年 5 月 11 日,攻擊者利用 pull_request_target、Actions cache 中毒與 OIDC 記憶體竊取,從 TanStack 的合法發佈管線推送 84 個惡意版本,竊取雲端憑證與 SSH 金鑰。本文拆解攻擊鏈與修補對策。
閱讀文章 ↗Config 募 2,700 萬美元種子輪:韓國製造業押注機器人資料
2026 年 5 月,Config 以逾 2 億美元估值完成超額認購的 2,700 萬美元種子輪,三星創投領投,現代、LG、SK 跟投。它不造機器人,而是幫所有人訓練機器人模型,已累積逾 10 萬小時人類動作資料,自比機器人資料的台積電。
閱讀文章 ↗DeepSeek 首度對外募資:估值傳從 200 億美元跳到 450 億
DeepSeek 傳出成立以來首次外部融資:國家集成電路產業投資基金據傳領投,估值幾週內從 200 億美元升至 450 億,騰訊與阿里巴巴洽談參投,募資首要目的是留住被挖角的研究人才。
閱讀文章 ↗Mini Shai-Hulud 蠕蟲襲捲 npm:連 Mistral SDK 與 SLSA 證明都淪陷
2026 年 5 月 11 日,自傳播的 Mini Shai-Hulud 蠕蟲透過被劫持的發布管線感染 170 多個 npm 套件,連 Mistral 官方 SDK 也中鏢。惡意版本帶著有效 SLSA 證明上架,專偷 AI 開發者的憑證與 Claude Code 設定。本文拆解攻擊鏈與對策。
閱讀文章 ↗Redis 之父 antirez 開源 ds4:跑得動前沿模型的本地推理引擎
Redis 作者 antirez 用一週寫出 ds4(DwarfStar):MIT 授權、C 語言的本地推理引擎,以 2/8-bit 非對稱量化讓 DeepSeek V4 Flash 與 GLM 模型跑進 96–128GB 的 Mac,內建原生編碼代理,登上 HN 首頁。
閱讀文章 ↗SAP 收購 Prior Labs:四年 11 億歐元打造結構化資料 AI 實驗室
2026 年 5 月 4 日,SAP 宣布收購成立 18 個月的德國新創 Prior Labs,並計畫四年投入逾 11 億歐元,以 TabPFN 表格基礎模型攻企業結構化資料;同時收緊代理政策:封殺 OpenClaw、僅認可 NemoClaw 等受背書架構。
閱讀文章 ↗Ghostty 宣布撤離 GitHub:一本停機日誌終結 18 年依賴
2026 年 4 月 28 日,Ghostty 終端機作者 Mitchell Hashimoto 宣布把專案撤離 GitHub:一個月的停機日誌幾乎天天打 X,Actions 故障擋住 PR 流程。本文解析這場 3,521 分 HN 熱議背後的平台風險。
閱讀文章 ↗IBM Granite 4.1 登場:8B 稠密模型追平 32B MoE 的開源算盤
IBM 於 2026 年 4 月 29 日發布 Granite 4.1:3B/8B/30B 全稠密模型以 Apache 2.0 開源,8B 追平上一代 32B MoE,並刻意拿掉推理模式換取可預測延遲。本文拆解 15T tokens 預訓練與四階段 RL 的工程細節。
閱讀文章 ↗DeepSeek V4 預覽上線:宣稱追平前沿模型,開源陣營再掀波
2026 年 4 月 24 日,DeepSeek 以預覽形式推出 V4,宣稱 V4-Pro-Max 追平前沿模型,在推理評測超越開源同級、勝過 GPT-5.2。CNBC 定調為 R1 之後又一次開源佈局擴張。本文解析宣稱的讀法與對開發者的意義。
閱讀文章 ↗同一顆模型、兩倍差距:四款 CLI 編碼 Agent 腳手架實測
開發者 Charles Azam 讓四款開源 CLI 編碼 Agent 接上同一顆 GLM-4.7 跑 Terminal-Bench 2.0:Mistral Vibe 拿 0.35、Codex 只有 0.15。結論是腳手架主導成績,模型之外的程式碼決定了兩倍以上的差距。
閱讀文章 ↗OpenAI 開源 Privacy Filter:1.5B 參數的 PII 偵測過濾模型
OpenAI 以 Apache 2.0 釋出 Privacy Filter:總參數 1.5B、每 token 僅啟動 50M 的稀疏 MoE 分類器,單次前向標記 8 類個資並以受限 Viterbi 解碼,連瀏覽器 WebGPU 都跑得動的資料最小化工具。
閱讀文章 ↗Android CLI 與官方 Skills:Google 把代理開發帶進終端機
Google 於 2026 年 4 月 16 日推出 Android CLI、Skills 與 Knowledge Base,讓 Gemini CLI、Claude Code、Codex 等任何代理都能在 Android Studio 外高效開發;token 消耗減少逾 70%、任務快 3 倍。
閱讀文章 ↗內省式擴散語言模型 I-DLM:首次追平同規模自回歸模型
Together AI 與 UIUC、Stanford 等團隊提出 I-DLM,用內省式跨步解碼讓擴散語言模型邊生成邊驗證,8B 版在 15 項基準追平 Qwen3-8B,AIME-24 大勝 LLaDA-2.1-mini,還能直接跑在 SGLang 上。
閱讀文章 ↗NVD 棄守 CVE 積壓:AI 讓漏洞洪流沖垮人工管線
NIST 宣布 NVD 只富化 KEV、聯邦政府與關鍵軟體三類 CVE,3 月 1 日前積壓全改標「Not Scheduled」。2020–2025 年 CVE 提交量增 263%,2025 年達 49,458 筆創新高,人工分析管線正式棄守。
閱讀文章 ↗A2A 協定滿一週年:150 家組織、五種 SDK 與 v1.0 穩定規格
2026 年 4 月 9 日 A2A 協定滿一週年:Linux Foundation 宣布支持組織從 50+ 增至 150+,SDK 擴至五種語言,v1.0 加入 Signed Agent Cards,並已落地 Azure AI Foundry 與 Bedrock AgentCore。
閱讀文章 ↗Gemma 4 開源發布:Apache 2.0、MoE 與 256K 上下文
2026 年 4 月 2 日 Google DeepMind 發布 Gemma 4:全系列改用 Apache 2.0 授權,四種尺寸從 2B 端側到 31B Dense,支援 140+ 語言、256K 上下文與影像語音輸入,31B 躍上 Arena 開源模型第三名。
閱讀文章 ↗PrismML 推出 1-bit Bonsai:把 8B 模型壓進 1.15 GB 的端側 LLM
2026 年 3 月 31 日,Caltech 衍生新創 PrismML 走出 stealth,發表號稱首款可商業化的端到端 1-bit LLM 家族 Bonsai:8B 旗艦僅 1.15 GB、比 16 位元同級小 14 倍,以 Apache 2.0 開源上架 Hugging Face。
閱讀文章 ↗Meta 開源 TRIBE v2:預測大腦如何回應影像、聲音與語言
Meta FAIR 釋出 TRIBE v2,以 700 多名受試者的 fMRI 資料訓練,能 zero-shot 預測新受試者、新語言與新任務的大腦反應,解析度較同類模型高約 70 倍,模型、程式碼與展示皆以 CC BY-NC 4.0 釋出。
閱讀文章 ↗Cohere 開源 Transcribe 語音模型:5.42% WER 登頂 ASR 排行榜
2026 年 3 月 26 日 Cohere 以 Apache 2.0 開源 20 億參數語音辨識模型 Transcribe,支援 14 種語言,在 Hugging Face Open ASR 排行榜以 5.42% 平均 WER 奪冠,主打企業音訊轉文字與本地部署。
閱讀文章 ↗Cursor Composer 2 被抓包以 Kimi K2.5 為底:開源權重的署名難題
2026 年 3 月 20 日 Cursor 發表 Composer 2,數小時後即有開發者從 API 流量裡的 model ID 讀出底模為 Moonshot 的 Kimi K2.5。Cursor 兩天後承認,引發開源授權合規與行銷透明的產業論戰。
閱讀文章 ↗USCC 報告:中國開源 AI 的「雙循環」正在強化工業主導地位
2026 年 3 月 23 日,美中經濟與安全審查委員會發布《Two Loops》研究報告:中國全面押注開源 AI,透過數位與實體兩個互相強化的循環鞏固工業主導地位。Qwen 衍生模型已超過十萬個,而美國出口管制主要只打到數位循環。
閱讀文章 ↗樂天開源 Rakuten AI 3.0:6,710 億參數、GENIAC 打造的日本最大模型
2026 年 3 月 17 日,樂天以 Apache 2.0 釋出約 6,710 億參數的 MoE 模型 Rakuten AI 3.0——日本最大高效能 AI 模型,訓練獲經產省與 NEDO 的 GENIAC 計畫支持。本文解析其日語基準成績、每 token 約 400 億活躍參數的架構,以及開源 671B 的現實限制。
閱讀文章 ↗Mistral 開源 Leanstral:專為 Lean 4 證明工程打造的 120B 稀疏模型
Mistral 以 Apache 2.0 釋出 Leanstral-120B-A6B,第一個專為 Lean 4 設計的開源程式代理。單次推理 18 美元,pass@2 便以約 15 分之一的成本超越 Claude Sonnet。本文解析其架構、FLTEval 成績與三種部署方式。
閱讀文章 ↗LeCun 的 AMI Labs 籌得 10.3 億美元種子輪:押注世界模型
Yann LeCun 創辦的 AMI Labs 完成 10.3 億美元種子輪,估值 35 億美元,寫下歐洲史上最大種子輪紀錄。團隊押注 JEPA 世界模型路線,主張 LLM 的幻覺問題難以根除,首個合作夥伴是數位醫療公司 Nabla,並承諾開源程式碼。
閱讀文章 ↗Sakana AI 發布 Doc-to-LoRA 與 Text-to-LoRA:一次前向傳遞生成 LoRA 配接器
Sakana AI 開源 Doc-to-LoRA 與 Text-to-LoRA 兩個超級網路研究:前者把整份文件壓進不到 50 MB 的 LoRA 配接器,後者用一句任務描述即時產生配接器。本文解析架構、實測數字與限制。
閱讀文章 ↗Guide Labs 開源 Steerling-8B:把可解釋性做進模型本身
2026 年 2 月 23 日,Guide Labs 開源 80 億參數的 Steerling-8B,把約 13 萬個概念的結構層直接蓋進模型架構,每個 token 都能回溯到輸入、概念與訓練資料,用更少算力勝過 LLaMA2-7B。可解釋性從事後分析變成工程問題。
閱讀文章 ↗阿里巴巴開源 Qwen3.5:397B 參數、201 種語言的 Agent 時代模型
2026 年 2 月 16 日除夕,阿里巴巴開源 Qwen3.5:首發 Qwen3.5-Plus 為 397B-A17B 稀疏 MoE 原生多模態模型,支援 201 種語言、內建可操作手機與電腦的視覺 agent,長情境解碼吞吐量達 Qwen3-Max 的 8.6 倍。本文解析規格、開源策略與中國 Agent 競賽。
閱讀文章 ↗Cohere Tiny Aya:33.5 億參數、70+ 語言的本機開源模型
Cohere Labs 於 2026 年 2 月 17 日釋出 Tiny Aya 開源多語模型:33.5 億參數、支援 70+ 語言、可在一般硬體本機運行,並選在印度 AI Impact Summit 期間登場。本文解析架構、部署方式與 CC-BY-NC 授權限制。
閱讀文章 ↗智利領軍推出 Latam-GPT:拉美首個本土開源大模型
2026 年 2 月 10 日,智利總統博里奇為 Latam-GPT 揭幕。這個由 CENIA 主導、集結 15 國力量的 700 億參數開源模型,以 8 TB 區域資料訓練,要對抗美國中心偏見。本文解析其規格、資金現實與主權意涵。
閱讀文章 ↗智譜開源 GLM-5:從 vibe coding 到 agentic engineering
2026 年 2 月 11 日,智譜(Z.ai)開源發布旗艦模型 GLM-5,主打更強的 coding 能力與長時程 agent 任務,官方標題寫著「From Vibe Coding to Agentic Engineering」,技術報告同步上架 arXiv。本文解析開源旗艦定位與中國模型軍團的發布節奏。
閱讀文章 ↗Mistral 開源 Voxtral Transcribe 2:即時轉錄挑戰雲端大廠
2026 年 2 月 4 日,Mistral 發表 Voxtral Transcribe 2:批次與串流兩款轉錄模型,FLEURS 詞錯誤率約 4%,即時版以 Apache 2.0 開源、4B 參數可跑邊緣裝置,API 每分鐘 0.003 美元起。本文解析其延遲與準確度取捨和生態定位。
閱讀文章 ↗從 Clawdbot 到 OpenClaw:爆紅開源代理一週二改名,安全疑慮升高
2026 年 1 月 30 日,爆紅開源 AI 代理 Clawdbot 因 Anthropic 商標關切先改名 Moltbot,三日內再更名 OpenClaw。本文解析兩次改名始末、代理架構,以及公網暴露的控制介面、詐騙與企業影子採用等安全風險。
閱讀文章 ↗AI 一次找齊 OpenSSL 全部 12 個零日漏洞:資安研究的分水嶺
2026 年 1 月 27 日,OpenSSL 協調修補 12 個零日漏洞,全數由 AISLE 的自主 AI 分析器發現,其中最高風險者 CVSS 9.8、不需有效金鑰即可能遠端觸發,最老的程式碼可追溯到 1998 年 SSLeay 時代。AI 漏洞發現正在改寫攻防規則。
閱讀文章 ↗Kimi K2.5 開源釋出:原生視覺加上 Agent Swarm 多代理協作
Moonshot AI 於 2026 年 1 月 27 日發表 Kimi K2.5:MIT 授權的開源 MoE model,具備原生視覺與 Agent Swarm 多代理協作。本文解析它對自架多模態 agent 堆疊與多代理系統生態的意義。
閱讀文章 ↗NVIDIA Earth-2 開放氣象模型全家桶:整條 AI 天氣預報管線開源
2026 年 1 月 26 日,NVIDIA 在美國氣象學會年會發表 Earth-2 開放模型家族:Atlas 做 15 天中期預報、StormScope 做 0 至 6 小時臨近預報、HealDA 用 GPU 秒級生成初始場,稱為首個全開源的加速氣象 AI 堆疊。
閱讀文章 ↗Overworld 開源 Waypoint-1:鍵盤滑鼠即時操控的擴散世界模型
2026 年 1 月 20 日,Overworld 在 Hugging Face 發布 Waypoint-1:以 10,000 小時遊戲影片訓練的即時互動世界模型,2.3B 開源,RTX 5090 上 30 FPS。文中解析 rectified flow 架構、WorldEngine 推理優化與世界模型離實用的距離。
閱讀文章 ↗憶阻器訓練新法 EaPU:AI 訓練能耗比 GPU 低近百萬倍
中國研究團隊在 Nature Communications 提出 EaPU 訓練法,把憶阻器的隨機切換特性轉為機率式權重更新,寫入次數減少逾 99%,訓練能耗比 GPU 低近六個數量級,裝置壽命延長約千倍。本文解析原理、實測結果與距離 LLM 的路還有多遠。
閱讀文章 ↗Ultralytics 推出 YOLO26:去 NMS、更快的邊緣視覺模型
2026 年 1 月 14 日,Ultralytics 發布 YOLO26:原生 NMS-free 端到端偵測、移除 DFL、Nano 版 CPU 推理提速最多 43%,五種尺寸對應邊緣到伺服器,採 AGPL-3.0 與企業雙授權。
閱讀文章 ↗
2025
2 篇文章Hugging Face 釋出 450M 開源機器人模型 SmolVLA
2025年6月初,Hugging Face 開源 450M 參數的視覺-語言-動作模型 SmolVLA,可跑在 MacBook 或單一消費級 GPU 上,並以非同步推論堆疊把任務完成時間縮短約30%,為低成本開源機器人生態提供基礎模型。
閱讀文章 ↗Google推出AI Edge Gallery:手機離線跑開源模型
2025年5月底Google低調釋出實驗性Android應用AI Edge Gallery,可從Hugging Face下載Gemma等開源模型,在手機上完全離線推理,採Apache 2.0授權並開源於GitHub;本文整理功能、應用場景與效能限制。
閱讀文章 ↗
2026
87 ARTICLESMistral Powers Firefox's Smart Window: What Zero Data Retention Changes for Browser AI
Mistral models now power Firefox Smart Window in France and North America, with zero data retention by default.
READ POST ↗Hermes Agent: Self-Hosted AI That Writes Its Own Skills
Nous Research's open-source Hermes Agent keeps its memory on your machine and writes its own skills. Its design bets, the Desktop beta, $1.5B valuation talks, and how to read its OpenRouter lead.
READ POST ↗SAM 3 and DINOv3 Cut Beamline Segmentation From a Month to 15 Minutes
Meta's open vision models let Berkeley Lab's SYNAPS-I label 3D beamline volumes in about 15 minutes instead of a month.
READ POST ↗Ox Alpha Was GLM-5.3-Flash: 31% of OpenRouter Traffic
The anonymous Ox Alpha that topped OpenRouter is GLM-5.3-Flash: 62T tokens on 100,000 Chinese chips, 31% of weekly traffic, MIT-licensed.
READ POST ↗Mistral Forge: No Fine-Tuning, No RAG — Training Enterprise Models From Scratch
Mistral's Forge trains enterprise models from scratch on proprietary data — no fine-tuning, no RAG. Embedded engineers, Mistral Small 4, and who actually needs a custom model.
READ POST ↗California Exempts Linux from Age-Verification Law
California passed AB 1856 unanimously, exempting operating systems under GPL, MIT, BSD, and Apache licenses from age-verification duties. The bill now awaits Governor Newsom.
READ POST ↗Nvidia in Talks to Buy Hugging Face for Over $13 Billion
Nvidia is in talks to acquire Hugging Face at a valuation above $13 billion, but no deal is finalized. The motives, the leverage, and the open-source community's worries.
READ POST ↗Haiku R1/beta6 Arrives: Two Years of Work After Beta5
Two years after beta5, Haiku ships R1/beta6 in the project's 25th anniversary week: 530+ tickets resolved, official Firefox branding, NVMM virtualization for QEMU, and a Go port.
READ POST ↗AWS acquires DuckLabs: DuckDB stays open source under MIT
AWS is acquiring DuckLabs, the Amsterdam company behind DuckDB, DuckLake, and Quack. The projects stay MIT-licensed, with the DuckDB Foundation keeping the IP.
READ POST ↗DuckDB v2.0 Preview: Server Mode, New Parser, New Storage
DuckDB v2.0 lands this fall: server mode, a new PEG-based SQL parser, storage format 2.0, asynchronous I/O, triggers, and 40x faster recursive queries.
READ POST ↗Nari Labs Gets Qwen3-TTS Talking in Under 50 ms
Nari Labs open-sourced a Qwen3-TTS 1.7B serving stack that hits 10 requests per second with sub-50 ms p95 time-to-first-audio on a single H100.
READ POST ↗Go 1.27 Lands: Generic Methods, json/v2, Post-Quantum TLS
Released August 19, Go 1.27 adds generic methods, moves encoding/json onto the v2 implementation, brings ML-DSA signatures to TLS, and speeds small allocations.
READ POST ↗DeepSeek Open-Sources Agent Harness in Developer Preview
DeepSeek's agent harness enters an MIT-licensed developer preview: a Cordis plugin kernel, a replayable append-only session log, and four runtime modes, open from day one.
READ POST ↗Firefox, Last Major Browser Still Supporting uBlock Origin
With Edge phasing out Manifest V2, Firefox is now the only major browser still running full uBlock Origin — and Mozilla says that support isn't going anywhere.
READ POST ↗Mojo 1.0: A Stable Foundation for AI Systems Programming
Mojo officially reached 1.0 on August 11, 2026: a stable, production-ready foundation that Modular itself relies on daily, with a 1.x path of mostly additive changes.
READ POST ↗DOE Launches Genesis Open Models for Open-Weight Science
The U.S. Department of Energy opened a Genesis open-weight models program under its Genesis Mission, partnering with Arcee AI on the first model, Genesis-Science-1.
READ POST ↗Oracle Bans AI-Generated Code from OpenJDK Contributions
OpenJDK's interim policy bars LLM- or diffusion-generated content from contributions, citing reviewer burden, security, and IP risk. Rust took a narrower path the same week.
READ POST ↗MiniMax H3 Goes Open With 2K Video and Native Stereo Audio
MiniMax announced H3 on July 31 and released the weights days later: a 33B omni-modal Transformer that generates up to 15 seconds of 2K video with native stereo audio.
READ POST ↗FFmpeg 9.0 'Lei': swscale Rewrite and Safer Defaults
FFmpeg 9.0 'Lei' arrived August 4: an ABI break across all seven libraries, the multi-year swscale rewrite, animated WebP decoding, and TLS verification on by default.
READ POST ↗Running Gemma 4 26B in 2 GB of RAM: TurboFieldfare
Open-source Swift and Metal engine TurboFieldfare keeps about 2 GB resident, streaming Gemma 4 26B-A4B MoE experts from SSD per token — 5-6 tok/s on an 8GB MacBook Air.
READ POST ↗Keychron's ZGM Brings Open-Source Firmware to Gaming Mice
Keychron opened ZGM, a GPL-3.0 gaming mouse firmware project built on Zephyr RTOS, extending the QMK and ZMK open-source culture to mice. The repo is still scaffolding.
READ POST ↗Debian's General Resolution on LLM Contributions
Debian opened discussion on July 23 for a General Resolution on LLM usage: eight proposals spanning an outright ban to conditional acceptance, with voting from August 15 to 28.
READ POST ↗India Orders GitHub to Take Down Dorsey's Bitchat
India's I4C ordered GitHub to disable Jack Dorsey's Bluetooth mesh app Bitchat within three hours. The leaked notice tests how takedown powers reach into open-source code.
READ POST ↗GrapheneOS Duress PIN Wipe Leads to Federal Prosecution
Sam Tunick entered a GrapheneOS duress code during an airport border search and his phone wiped itself. The DOJ now prosecutes him for destroying property to prevent seizure.
READ POST ↗Ruff 0.16 Ships 413 Default Rules, Up From 59
Ruff v0.16.0 turns on 413 rules by default (up from 59), formats Python blocks in Markdown, and adds ruff: ignore comments. What changed and how to upgrade.
READ POST ↗Block Open-Sources Buzz: Chat, AI Agents, Git in One Place
Block open-sourced Buzz on July 21, 2026: team chat, AI agents as workspace members, and Git hosting on one signed-event stream, to cut reliance on Slack and GitHub.
READ POST ↗Firefox 153 Ships Vulkan Video Decoding and JPEG XL
Firefox 153 is rolling out with Vulkan hardware video decoding, experimental JPEG XL in Firefox Labs, HDR playback on Windows, and local network access now asking by default.
READ POST ↗Grok Build Open Source: The Harness Is the Part Worth Reading, Not the UI
xAI open-sourced Grok Build's complete harness: agent loop, tools, terminal UI, and extension system — and why the source itself is the definitive reference for extension authors.
READ POST ↗Reflection Signs $1B Compute Deal With Nebius
Open-weight lab Reflection AI signed a $1B compute deal with GPU cloud Nebius for Nvidia's latest chips — its second mega-deal in three weeks, as frontier compute turns scarce.
READ POST ↗NemoClaw: LangChain and NVIDIA's Governed Deep Agents Stack
LangChain and NVIDIA's NemoClaw blueprint pairs Nemotron 3 Ultra with the Deep Agents Code harness and OpenShell sandbox — governed agents at a tenth of the eval cost.
READ POST ↗Bun 1.4 Is Rust Now: 64 Claude Agents Ported It in 11 Days
Bun v1.4.0 is the first Rust-written release of the JS runtime: Claude Code workflows plus Claude Fable 5 ported 535k lines of Zig in 11 days with 64 agents, for about $165k.
READ POST ↗Leanstral 1.5: Mistral Saturates miniF2F at $4 a Proof
Mistral's Apache-2.0 Leanstral 1.5 saturates miniF2F, solves 587 PutnamBench problems at roughly $4 each, and turns Lean 4 proof engineering into a cheap, repeatable routine.
READ POST ↗Meta Brain2Qwerty v2: 61% Word Accuracy Without Surgery
Meta's Brain2Qwerty v2 decodes sentences from non-invasive MEG recordings at 61% word accuracy, up from 8% for prior methods, using end-to-end deep learning and fine-tuned LLMs.
READ POST ↗Patch the Planet: How OpenAI and Trail of Bits Are Using AI to Help Open Source Maintainers
OpenAI's Patch the Planet initiative uses AI models and human review to help open source maintainers find and fix vulnerabilities without adding to their burden.
READ POST ↗GLM-5.2: Open Weights, 1M Context, Long-Horizon Gains
Z.AI open-sources GLM-5.2, a 753B MIT-licensed model with a stable 1M-token context that trails Claude Opus 4.8 by about one percent on long-horizon coding benchmarks.
READ POST ↗Baseten's $1.5B Round Bets Big on Open-Source Inference
Baseten is reportedly raising $1.5B at up to a $13B valuation, five months after a $300M Series E at $5B. Inside the split-priced round and the open-source inference bet.
READ POST ↗Sarvam AI Turns Unicorn as HCLTech Leads $234M Round
Sarvam AI raised $234 million at a $1.5 billion valuation, becoming India's newest AI unicorn, as HCLTech put $150 million into its Indian-language model stack.
READ POST ↗eve: Vercel's Open-Source Agent Framework Finally Gives Agents a Standard Shape
Vercel open-sourced eve, the filesystem-first framework behind its hundred-plus production agents: directory-as-agent design, preloaded primitives, a self-reported roster, and Vercel-bound trade-offs.
READ POST ↗Mistral Rumored to Raise €3B at €20B Valuation
Mistral is in early talks to raise ~€3B at a ~€20B valuation, nearly double its September 2025 Series C — yet its ~$4B raised to date is a fraction of US labs' war chests.
READ POST ↗Google Open-Sources DiffusionGemma: 4x Faster Generation
Google open-sources DiffusionGemma: an Apache 2.0 26B-A4B diffusion LLM hitting 1000+ tokens/sec on one H100 — 4x faster than autoregressive, at a quality cost vs Gemma 4.
READ POST ↗TensorZero Archives Repo After $7.3M Seed, Winds Down
Open-source LLM gateway TensorZero archived its repo on June 13. Its founder says the $7.3M-seed company is winding down; capital goes back to investors, code stays Apache 2.0.
READ POST ↗Hello Robot Stretch 4: A $30,000 Bet on Home Robots
Hello Robot's $29,950 Stretch 4 home robot sold out its first 200-300 units. It ships by courier, keeps humans in the loop, and quadriplegic users already depend on it daily.
READ POST ↗Miasma Worm Hits Microsoft Repos, Targeting AI Coding Agents
A hijacked account pushed a malicious commit into Azure's durabletask repo, planting files that run a credential stealer when Claude Code or Cursor opens it. 73 repos went dark.
READ POST ↗Microsoft ASSERT Turns Text Specs into AI Behavior Tests
Microsoft open-sourced ASSERT, a framework that turns plain-language behavior specs into generated test suites with LLM-judge scoring and CI regression gates for AI agents.
READ POST ↗Runtime (YC P26): Sandboxed Coding Agents for Whole Teams
YC P26 startup Runtime launched May 21 on Hacker News: sandboxed Claude Code and Codex sessions for whole teams, with env snapshots, secret proxy, and RBAC. Open core.
READ POST ↗NanoClaw Turned Down $20M to Keep Building Sandboxed Agents
NanoClaw runs each agent session in its own Docker container. After viral growth it turned down a $20M buyout and raised a $12M seed with Docker, Vercel, and HF's CEO on board.
READ POST ↗Google's ERA Lands in Nature: Expert-Level Scientific Code
Google's ERA uses Gemini to write and optimize scientific code; a May 19 Nature paper reports expert-level results across six domains, with five new applications and open source.
READ POST ↗NHS Retreats from Open Source; GDS Says Keep Code Open
After AI bug-hunting jumped, NHS England moved to close nearly all its open source repos; May 14 GDS and DSIT guidance pushes back: keep code open by default and fix weaknesses.
READ POST ↗DuckDB Ships Quack, Its Own Client-Server Protocol
DuckDB's Quack protocol lets DuckDB act as client and server over HTTP in one round trip: 60M rows in 4.94s vs 17.4s for Arrow Flight SQL. Production ships with DuckDB v2.0.
READ POST ↗Inside the TanStack npm Supply-Chain Compromise
A pwn request, a poisoned Actions cache, and an OIDC token dumped from runner memory let attackers publish 84 malicious versions of 42 TanStack npm packages on May 11, 2026.
READ POST ↗Config Raises $27M to Be the TSMC of Robot Training Data
Config closed an oversubscribed $27M seed at a $200M+ valuation, led by Samsung with Hyundai, LG, and SK joining — a neutral data layer for robot foundation models.
READ POST ↗DeepSeek's First Outside Round: Valuation Talk Hits $45B
DeepSeek is negotiating its first-ever outside round per FT and Bloomberg — a state chip fund to lead, Tencent and Alibaba in talks, valuation up from $20B to $45B.
READ POST ↗Shai-Hulud npm Worm Hits Mistral SDK; Provenance No Defense
A self-spreading npm worm infected 170+ packages including Mistral's SDK on May 11, 2026, published with valid SLSA provenance and stealing Claude Code configs and cloud creds.
READ POST ↗antirez's ds4: A Local LLM Inference Engine Built for Metal
Redis creator antirez open-sourced ds4: a C local inference engine tuned for DeepSeek V4 Flash and GLM on Metal, CUDA and ROCm, with a coding agent. MIT licensed.
READ POST ↗SAP Buys Prior Labs: €1B Bet on Tabular Foundation Models
SAP agreed to buy Prior Labs, the 18-month-old Freiburg lab behind the TabPFN model, and will invest over €1 billion over four years to attack enterprise structured data.
READ POST ↗Ghostty Quits GitHub After a Month of Daily Outages
Ghostty creator Mitchell Hashimoto is moving the terminal off GitHub after a month of near-daily outages. What an 18-year user's exit says about platform risk for dev teams.
READ POST ↗IBM Granite 4.1: An 8B Dense Model Matching a 32B MoE
IBM's Granite 4.1 ships 3B/8B/30B dense models under Apache 2.0; the 8B matches the old 32B MoE, reasoning is deliberately removed, and context reaches 512K.
READ POST ↗DeepSeek V4 Arrives in Preview: Claiming Frontier Parity, Rattling Open Source Again
On April 24, 2026, DeepSeek launched V4 as a preview, claiming V4-Pro-Max closes the gap with frontier models — beating open-source peers on reasoning and outstripping GPT-5.2. How to read the claims.
READ POST ↗Same Model, 2x Gap: Benchmarking Four CLI Coding Agents
Four open-source CLI coding agents running the same model (GLM-4.7) on Terminal-Bench 2.0: Mistral Vibe scored 0.35, Codex 0.15. The scaffolding decides, not the model.
READ POST ↗OpenAI Open-Sources Privacy Filter for PII Detection
OpenAI's Privacy Filter, now open under Apache 2.0: a 1.5B-parameter sparse MoE (50M active) that tags 8 PII categories in one forward pass and runs even in the browser.
READ POST ↗Android CLI: Google's Terminal-First Bet on Coding Agents
Google's Android CLI, open-source Skills, and Knowledge Base let any agent — Gemini CLI, Claude Code, Codex — build Android apps; setup tokens fell over 70%, tasks ran 3x faster.
READ POST ↗I-DLM: A Diffusion LLM That Matches Same-Scale AR Quality
Diffusion LM meets AR self-checks: I-DLM-8B matches Qwen3-8B on 15 benchmarks, beats LLaDA-2.1-mini by 26 points on AIME-24, and runs on stock SGLang at 2.9-4.1x the throughput.
READ POST ↗NVD Gives Up on CVE Backlog as AI Inflow Accelerates
NIST now enriches only KEV, federal and critical-software CVEs; the pre-March 2026 backlog is 'Not Scheduled'. CVE submissions rose 263% from 2020 to 2025.
READ POST ↗A2A Turns One: 150+ Orgs, Five SDKs, Stable v1.0 Spec
A2A turned one on April 9, 2026: 150+ orgs, five production SDKs, a stable v1.0 spec with Signed Agent Cards, and deployments on Azure AI Foundry and Amazon Bedrock AgentCore.
READ POST ↗Gemma 4 Ships Under Apache 2.0: Google's Open Model Reset
Google DeepMind released Gemma 4 on April 2, 2026 under Apache 2.0 — four sizes from 2B edge to 31B dense, 140+ languages, 256K context, Arena top-3. What it changes for builders.
READ POST ↗PrismML 1-bit Bonsai: An 8B LLM Squeezed Into 1.15 GB
Caltech spinoff PrismML emerged from stealth with 1-bit Bonsai, an end-to-end 1-bit LLM family: the 8B flagship is 1.15 GB, 14x smaller than 16-bit peers, Apache 2.0 licensed.
READ POST ↗Meta's TRIBE v2 Predicts Brain Responses Like a Digital Twin
Meta FAIR's TRIBE v2, trained on fMRI from 700+ volunteers, zero-shot predicts brain responses for new subjects and languages at 70x the resolution of similar models.
READ POST ↗Cohere Open-Sources Transcribe, Tops ASR Leaderboard
Cohere's Transcribe, open-sourced March 26, 2026, is a 2B-parameter ASR model under Apache 2.0 covering 14 languages, first on the Hugging Face Open ASR Leaderboard at 5.42% WER.
READ POST ↗Cursor Composer 2 Caught Building on Kimi K2.5 Weights
Cursor launched Composer 2 as its own model; a leaked API model ID exposed its Kimi K2.5 base within hours. Cursor's admission ignited an open-weights attribution debate.
READ POST ↗USCC: China's Open-Source AI Reinforces Industrial Power
A new USCC report argues China's open-source AI strategy is self-reinforcing: open models spread globally while factory deployment feeds data back. Qwen alone has 100,000+ derivatives.
READ POST ↗Rakuten AI 3.0: Japan's Largest Model Goes Open-Weight
Rakuten open-sources Rakuten AI 3.0 under Apache 2.0: about 671B total parameters, 40B active per token, GENIAC-backed and tuned for Japanese. Benchmarks and the reality of running it.
READ POST ↗Mistral Open-Sources Leanstral, a Lean 4 Proof Agent
Mistral released Leanstral-120B-A6B under Apache 2.0 — the first open-source agent purpose-built for Lean 4 proof engineering. One pass costs $18; pass@2 beats Claude Sonnet at roughly 1/15 the cost.
READ POST ↗LeCun's AMI Labs Raises $1.03B Seed to Bet on World Models
AMI Labs, Yann LeCun's post-Meta startup, raised a $1.03B seed at a $3.5B valuation — Europe's largest ever — to build JEPA world models, starting with healthcare partner Nabla.
READ POST ↗Sakana AI's Doc-to-LoRA: Documents Become LoRA in One Pass
Sakana AI open-sources Doc-to-LoRA and Text-to-LoRA, hypernetworks that generate LoRA adapters in a single sub-second forward pass — long documents in under 50 MB, task adapters from one sentence.
READ POST ↗Steerling-8B: Guide Labs' Inherently Interpretable Open LLM
Guide Labs open-sources Steerling-8B, an 8B base model with a built-in concept layer — 33K supervised plus 100K discovered concepts, full token traceability, competitive at fewer FLOPs.
READ POST ↗Alibaba Open-Sources Qwen3.5 for the Agentic AI Era
Alibaba open-sourced Qwen3.5 on Lunar New Year's Eve: a multimodal 397B-A17B MoE with 201 languages, 8.6x Qwen3-Max decoding throughput, and visual agents that operate phones and desktops.
READ POST ↗Cohere's Tiny Aya: A 3.35B Open Model for 70+ Languages
Cohere Labs released Tiny Aya: a 3.35B open-weight model covering 70+ languages, built to run on a laptop, debuted at the India AI Impact Summit. Architecture, deployment, and license, broken down.
READ POST ↗Latam-GPT: Latin America's First Homegrown Open-Source LLM
Chile launched Latam-GPT on Feb 10: a 70B open-source model from CENIA and 15 partner countries, trained on 8 TB of regional data to counter US-centric bias.
READ POST ↗Zhipu Open-Sources GLM-5: From Vibe Coding to Agentic Engineering
On February 11, 2026, Zhipu AI (Z.ai) released GLM-5, an open-source flagship built for stronger coding and long-horizon agent tasks — a launch titled 'From Vibe Coding to Agentic Engineering.'
READ POST ↗Mistral Voxtral Transcribe 2: Open Speech-to-Text, On-Device
Mistral's Voxtral Transcribe 2 ships batch and streaming transcription models: ~4% WER on FLEURS, Apache 2.0 open weights for the 4B realtime model, and $0.003/min API pricing.
READ POST ↗Clawdbot to OpenClaw: Open-Source Agent Hits Security Wall
Viral open-source agent Clawdbot became OpenClaw on January 30, 2026 after an Anthropic trademark complaint. Behind the rename chaos: exposed control panels, scams, and shadow enterprise use.
READ POST ↗AI Found All 12 OpenSSL Zero-Days in One Release
On January 27, 2026, OpenSSL patched 12 zero-day vulnerabilities, every one found by AISLE's autonomous AI analyzer — including a CVSS 9.8 overflow with code dating back to 1998 and the SSLeay era.
READ POST ↗Kimi K2.5 Goes Open Source: Native Vision and Agent Swarm Coordination
Moonshot AI released Kimi K2.5 on January 27, 2026: an MIT-licensed open-source MoE model with native vision and Agent Swarm coordination. What it means for self-hosted agent stacks.
READ POST ↗NVIDIA Earth-2: First Fully Open AI Weather Stack
NVIDIA open-sourced the Earth-2 weather AI stack — Atlas 15-day forecasts, StormScope storm nowcasting, HealDA GPU-fast data assimilation — with national weather agencies already running it.
READ POST ↗Overworld Open-Sources Waypoint-1, a Real-Time World Model
Overworld open-sourced Waypoint-1, a real-time interactive video diffusion world model trained on 10,000 hours of gameplay, running at 30 FPS on an RTX 5090 via its WorldEngine stack.
READ POST ↗EaPU Cuts AI Training Energy Nearly a Million-Fold vs GPUs
Zhejiang Lab and Fudan University researchers published EaPU: probabilistic weight updates cut memristor writes by over 99% and training energy by nearly six orders of magnitude versus GPUs.
READ POST ↗Ultralytics YOLO26: NMS-Free Vision AI Built for the Edge
Ultralytics shipped YOLO26 on January 14, 2026: NMS-free end-to-end detection, DFL removed, up to 43% faster nano CPU inference, five scales from edge to server, AGPL-3.0 dual licensing.
READ POST ↗
2025
2 ARTICLESHugging Face ships SmolVLA, a lean open robotics model
In early June 2025, Hugging Face open-sourced SmolVLA, a 450M-parameter vision-language-action model that runs on a MacBook and cuts task time about 30% — a base for low-cost open robotics.
READ POST ↗Google AI Edge Gallery runs open AI models offline on phones
In late May 2025 Google quietly released AI Edge Gallery, an experimental Android app running open models like Gemma fully offline on phones. Features, limits, and why it matters.
READ POST ↗