2026
40 篇文章Megakernel 不是炫技:把 decode 從「等 kernel」改成「等資料」的實作取捨
Cohere 用單一 CUDA 檔把 North Mini Code 的 decode 做成 persistent megakernel,batch size 1 吞吐從 vLLM 的 185 tok/s 拉到 292 tok/s。
閱讀文章 ↗在 SageMaker HyperPod 上部署 Qwen3.8-2.4T-A95B:單節點跑 2.4T 開源模型的實戰配置
AWS 示範如何在單一 p6-b300 節點上用 NVFP4 量化與 vLLM 部署 2.4T 參數的 Qwen3.8,省下多節點成本。
閱讀文章 ↗蘋果 M6 與 M5 Ultra 登場:2nm 與四晶粒推向端側 AI
蘋果於 8 月 25 日發表首款 2nm 晶片 M6 與首款四晶粒架構的 M5 Ultra,Mac mini 與 Mac Studio 同步更新,最高 512GB 統一記憶體讓數千億參數模型能在桌面本機執行。
閱讀文章 ↗Cerebras CS-4 登場:三片 Turbo 晶圓的機架級推理系統
Cerebras 於 8 月 19 日發表 CS-4 機架級系統與 WSE-3T 晶圓:不做新矽,把同一片晶圓灌兩倍功率,宣稱推理比 GPU 快 30 倍,並把提示詞處理交給 AWS 與 AMD 加速器分擔。
閱讀文章 ↗記憶體一年漲五倍:AI 資料中心把 DRAM 吃光了
PCPartPicker 數據顯示 DDR5 均價一年上漲約 500%,128GB 套裝從 329 美元漲到 3,399 美元;HBM 排擠標準 DRAM 產能、雲端巨頭用預付款圈走 2027 年貨源,新產能卻要 2028 年後才落地。
閱讀文章 ↗Nvidia 縮手:OpenAI 資料中心貸款擔保從 2,500 億美元砍至 1,200 億以下
《華爾街日報》報導 Nvidia 縮減 OpenAI 資料中心債務擔保,初期將低於 1,200 億美元,遠低於 7 月談及的 2,500 億美元;交易僅有備忘錄,縮水被視為對股東的訊號,循環融資疑慮浮上檯面。
閱讀文章 ↗RAMageddon 延燒:傳 DRAM 與 HBM 2027 年產能已被採購一空
DIGITIMES 報導經 TweakTown 披露:三星、SK 海力士與美光的 DRAM 與 HBM 產能,據傳已被五年期長約包到 2027 年;AI 是最大買家,消費端 SSD 與 RAM 價格持續攀升。
閱讀文章 ↗AMD 收購 Taalas:把模型權重直接蝕刻進晶片
2026 年 8 月 6 日 AMD 宣布收購多倫多新創 Taalas:以權重蝕刻加 SRAM 快取的專用推理晶片路線,HC1 已在 Llama 3.1 8B 跑出每秒近 1.7 萬 token。
閱讀文章 ↗NVIDIA 投資 SSI:封閉兩年的實驗室,換來一個量級的算力
NVIDIA 於 7 月 27 日宣布投資 Ilya Sutskever 的 Safe Superintelligence 並達成長期策略合作,SSI 取得 Vera Rubin 平台存取權、算力擴大一個數量級,代價是讓 NVIDIA 罕見地看見其內部研究。
閱讀文章 ↗Google 揭露94.1億美元SpaceX持股,賺了超過百倍
Alphabet 在 2026 年 7 月 23 日提交的 Q2 10-Q 中揭露:持有約 6% SpaceX 股權,價值 941 億美元,相對 2015 年投入的 9 億美元成長逾百倍。本文整理持倉細節、分階段解禁安排,以及 Google 同時身為 SpaceX 股東與算力客戶的雙重關係。
閱讀文章 ↗Google 新晶片 Frozen v2 曝光:為 Gemini 而生、能效目標十倍
The Information 報導 Google 正開發內部代號 Frozen v2 的伺服器晶片,為 Gemini 模型打造,目標 2028 年推出、能效較現有晶片提升六到十倍。
閱讀文章 ↗Anthropic 洽談三星代工自研晶片:多源算力的第四條線
The Information 報導 Anthropic 正與三星洽談代工自研 AI 晶片,用途與規格皆未定;在 Google TPU、Amazon Trainium 與 Nvidia GPU 之外,多源硬體策略可能再添一條自己的供應線。
閱讀文章 ↗Etched 走出隱身:10 億美元訂單與台積電量產的推理晶片
AI 晶片新創 Etched 宣布走出隱身:手上握有 10 億美元合約訂單,首款晶片已由台積電成功量產,Frontier 推理叢集進入客戶測試,成為挑戰 Nvidia 的最新獨立勢力。
閱讀文章 ↗OpenAI 首顆自研推理晶片 Jalapeño 登場:九個月設計到量產,宣稱勝過 GB300
2026 年 6 月 24 日,OpenAI 與 Broadcom 發表首顆自研 LLM 推理晶片 Jalapeño,約九個月完成設計到量產,兩家公司並宣稱在關鍵推理基準擊敗 NVIDIA GB300。本文解析自研晶片對推理成本與算力供應的意義。
閱讀文章 ↗Amazon 擬對外出售 Trainium 晶片,劍指 NVIDIA
Bloomberg 報導,Amazon AI 主管 Peter DeSantis 證實 AWS 正與潛在買家洽談對外出售 Trainium 晶片;執行長 Andy Jassy 估算晶片事業獨立運作的年營收可達約 500 億美元,直接挑戰 NVIDIA 在資料中心 AI 晶片市場的主導地位。
閱讀文章 ↗Google 每月付 SpaceX 9.2 億美元租算力:IPO 前的橋接合約
2026 年 6 月 5 日 SpaceX 在 SEC 文件披露:Google 將月付 9.2 億美元、租用約 11 萬顆 NVIDIA GPU,期間為 2026 年 10 月至 2029 年 6 月,作為 Gemini Enterprise 需求超預期的橋接產能,也是 SpaceX 上市前一週的最新算力大單。
閱讀文章 ↗XCENA 融資 1.35 億美元:賭 AI 瓶頸不在算力而在記憶體
韓國晶片新創 XCENA 以 5.7 億美元估值完成 1.35 億美元 B 輪。其 MX1 晶片把數千顆 RISC-V 核心放到 DRAM 旁,在記憶體模組內管理 KV cache,號稱十台伺服器的工作一台就能做完,2026 年底由三星 4 奈米製程量產。
閱讀文章 ↗黃仁勳宣稱 Vera CPU 打開 2,000 億美元新市場:agent 晶片戰開打
NVIDIA 財報會上黃仁勳宣稱為 agent AI 而生的 Vera CPU 打開 2,000 億美元全新市場,今年獨立銷售已達 200 億美元;在 AWS 自研晶片與 Meta 轉單之際,CPU 成為 AI 晶片戰的新前線。
閱讀文章 ↗AMD Q1 2026 財報:資料中心營收年增 57%,伺服器 CPU 市場預測翻倍
AMD 公布 2026 年第一季財報:資料中心營收 58 億美元、年增 57%,總營收 102.5 億美元。蘇姿豐將伺服器 CPU 市場年增率預測從 18% 上調至 35%、2030 年超過 1,200 億美元。本文拆解財報數字、MI450 與 Helios 時程,以及 agentic AI 帶動的 CPU 需求。
閱讀文章 ↗Cerebras IPO 定價出爐:35 億美元募資、瞄準 266 億美元市值
2026 年 5 月 4 日,Cerebras 啟動 IPO 定價:釋出 2,800 萬股、每股 115 至 125 美元,高點估募 35 億美元、市值達 266 億美元,可望成為 2026 年迄今最大科技 IPO;Bloomberg 報導承銷訂單已湧入 100 億美元。
閱讀文章 ↗Cerebras 重新送件 IPO:5.1 億美元營收與 OpenAI 百億大單
AI 晶片公司 Cerebras 於 2026 年 4 月向 SEC 送件申請 IPO,計畫五月中掛牌。文件顯示 2025 年營收 5.1 億美元,加上 OpenAI 超過百億美元的運算訂單與 AWS 合作,這是它兩年內第二次挑戰公開市場。
閱讀文章 ↗Amazon 自研晶片年營收破 200 億美元:Jassy 股東信的算力藍圖
2026 年 4 月 9 日 Jassy 年度股東信揭露:Graviton、Trainium 與 Nitro 年營收突破 200 億美元、年增三位數百分比,獨立計價約 500 億美元,並首度鬆口未來可能把機櫃賣給第三方。
閱讀文章 ↗GPU 變成放貸標的:Forum Markets 切入 AI 晶片過橋融資
2026 年 4 月 8 日,Forum Markets 宣布以 60 至 120 天過橋貸款資助 Neocloud 業者採購 NVIDIA GPU,首筆承諾 2,500 萬至 5,000 萬美元、目標年化報酬中雙位數,並計畫把貸款代幣化。
閱讀文章 ↗韓國國家基金投資 Rebellions 1.66 億美元,啟動「K-Nvidia」晶片計畫
2026 年 3 月 26 日,韓國金融委員會批准透過國家成長基金直接投資 AI 晶片新創 Rebellions 約 1.66 億美元,是「K-Nvidia」計畫首筆直接投資,用於 NPU 量產與次世代 AI 晶片開發。本文解析主權 AI 晶片競賽。
閱讀文章 ↗Meta 與 Nebius 簽五年 270 億美元協議:認購 Vera Rubin 世代產能
2026 年 3 月 16 日,Nebius 宣布與 Meta 簽署五年協議:120 億美元專屬 AI 容量,最高再以 150 億美元認購未來叢集剩餘產能,總值約 270 億美元,建立於 NVIDIA Vera Rubin 平台,2027 年初開始交付,是 Neocloud 迄今最大合約之一。
閱讀文章 ↗GTC 2026 開幕:Vera Rubin 七款晶片全量投產,訂單總額約一兆美元
NVIDIA 執行長黃仁勳在 GTC 2026 主題演講發表 Vera Rubin 平台,七款新晶片已全面量產;CNBC 報導 Blackwell 加 Vera Rubin 到 2027 年的合計訂單約 1 兆美元。本文解析這場演講對算力供應鏈與採購規劃的意義。
閱讀文章 ↗Tesla Terafab 倒數七天:Musk 押上 200 億美元的自建 AI 晶圓廠
2026 年 3 月 14 日,Musk 宣布 Terafab Project 七天後啟動。這座造價約 200 億美元、瞄準 2nm 製程的 AI 晶圓廠,目標月產十萬片晶圓,為 FSD 與 Optimus 供應 AI5 晶片。本文解析一家車廠跳進半導體最前沿的垂直整合豪賭。
閱讀文章 ↗Meta 四款自研 MTIA 晶片登場:兩年四代的矽節奏
2026 年 3 月 11 日,Meta 發表 MTIA 300、400、450、500 四款自研 AI 晶片,數十萬顆已在生產環境運轉,並宣稱約每六個月推出新一代。本文解析逐代規格跳幅、chiplet 模組化設計,與 NVIDIA、Broadcom 之間的分層算力佈局。
閱讀文章 ↗Ayar Labs 募得 5 億美元 Series E:用共封裝光學打通 AI 叢集的銅線天花板
2026 年 3 月 3 日,光互連公司 Ayar Labs 宣布 5 億美元 Series E,投後估值 37.5 億美元,NVIDIA 與 AMD 續投。執行長 Mark Wade 直指 AI 基礎設施正撞上互連效率造成的電力高牆,本文解析 CPO 共封裝光學如何讓數千顆 GPU 如同一台機器運作。
閱讀文章 ↗Meta 向 Google 租 TPU:自建雲巨頭回頭當對手的客戶
2026 年 2 月 26 日,路透社報導 Meta 簽下數十億美元多年期協議,租用 Google Cloud TPU 訓練下一代模型,並洽談購買數百萬顆 TPU 自行部署。本文解析交易動機與對 Nvidia 的衝擊。
閱讀文章 ↗NVIDIA FY2026 Q4 財報:單季營收 681 億美元創紀錄,全年 2,159 億美元
2026 年 2 月 25 日,NVIDIA 公布 FY2026 第四季財報:單季營收 681 億美元創新高、季增 20%,淨利 430 億美元;全年營收 2,159 億美元、年增 65%。本文拆解數字背後的訊號,以及對採購與產品團隊的意義。
閱讀文章 ↗Meta 向 NVIDIA 鎖定數百萬顆 Blackwell:供給先被大單吃掉
Meta 與 NVIDIA 簽長期協議、鎖定數百萬顆 Blackwell GPU:訓練算力天花板大幅拉高,百萬顆級的排程與容錯成為新工程關卡,供給被大單提前佔用後,其他買家取得算力更貴也更慢。
閱讀文章 ↗Cerebras 完成 10 億美元 Series H 融資:晶圓級推論的 230 億美元賭注
2026 年 2 月,Cerebras 完成 10 億美元 Series H 輪融資,投後估值約 230 億美元,Tiger Global 領投、AMD 與 Benchmark 等跟投。本文解析晶圓級引擎的技術賭注、OpenAI 大單後的資本驗證,以及推論市場的多供應商格局。
閱讀文章 ↗Microsoft Maia 200 登場:FP4 破 10 petaFLOPS 的自研推理晶片
2026 年 1 月 26 日,Microsoft 發表第二代自研 AI 加速器 Maia 200:FP4 算力超過 10 petaFLOPS、1,400 億顆以上電晶體、750W 封裝,宣稱 FP4 效能為 Amazon Trainium3 的三倍。本文解析規格對比與自研矽的成本帳。
閱讀文章 ↗台積電 Q4 獲利創新高,2026 資本支出上看 560 億美元
台積電 1 月 15 日公布 2025 年第四季財報:單季營收首破 1 兆新台幣、淨利年增 35% 創新高,2 奈米開始量產,2026 年資本支出指引 520–560 億美元。本文拆解數字背後的 AI 需求訊號、記憶體短缺與美國擴產的拉扯。
閱讀文章 ↗OpenAI 與 Cerebras 簽逾 100 億美元協議:750 MW 晶圓級算力直攻低延遲推論
2026 年 1 月 14 日,OpenAI 與 Cerebras 簽署多年協議:到 2028 年部署最多 750 MW 晶圓級算力,CNBC 報導交易價值超過 100 億美元,全部用於低延遲推論。本文解析條款細節、Cerebras 對 G42 的營收依賴,以及 OpenAI 的多供應商算力佈局。
閱讀文章 ↗Etched 募得 5 億美元:用 transformer 專用晶片挑戰 NVIDIA
2026 年 1 月 13 日,Bloomberg 報導 AI 晶片新創 Etched 募得約 5 億美元,由 Stripes 領投、估值約 50 億美元。其 Sohu 晶片採 TSMC 4nm 製程,是專為 transformer 模型設計的 ASIC,直接挑戰 NVIDIA 的通用 GPU。
閱讀文章 ↗OpenAI 與 SoftBank 合投 SB Energy 10 億美元,簽下 1.2 GW Stargate 租約
2026 年 1 月 9 日前後,OpenAI 與 SoftBank 各投資 5 億美元於 SB Energy,合計 10 億美元,並伴隨 1.2 GW 的 Stargate 資料中心租約。本文解析算力與電力深度綁定的最新階段。
閱讀文章 ↗Intel CES 2026 發表 Core Ultra Series 3:18A 製程上機的 AI PC 首役
Intel 在 CES 2026 發表 Core Ultra Series 3(Panther Lake):首款運算晶片採用 18A 製程的 AI PC 平台,搭配 Arc B390 內顯與 200 款以上筆電設計,1 月 27 日全球開賣。本文解析規格、時程與對邊緣 AI 佈局的意義。
閱讀文章 ↗AMD CES 2026 開局:Helios 機架、MI455X 與 AI Everywhere 全線推進
2026 年 1 月 5 日,AMD 執行長蘇姿豐在 CES 發表「AI Everywhere」主題演講:每機架 3 AI exaflops 的 Helios 平台、MI455X 與 MI440X GPU,加上 60 TOPS 的 Ryzen AI 400 系列。本文解析 AMD 從資料中心到終端的完整算力佈局。
閱讀文章 ↗
2025
5 篇文章HBM 行情發酵:美光、SK 海力士股價齊揚
2025年6月26日記憶體行情延燒:美光財報後股價走揚,SK 海力士市值創下約1,570億美元新高;AI 伺服器帶動 HBM 需求,兩大記憶體巨頭同步改寫市場對記憶體股的評價。
閱讀文章 ↗美光財報:HBM 年內售罄,單季營收創新高
2025年6月25日美光公布會計年度第三季財報,單季營收93億美元創新高,HBM 收入季增近五成;執行長梅羅特拉表示 HBM 2025 年產能已全數售罄,第四季營收指引107億美元也高於市場預期。
閱讀文章 ↗黃仁勳:輝達財測將排除中國市場,H20季度損失約80億美元
2025年6月中,輝達執行長黃仁勳接受CNN專訪表示,因H20出口管制短期難解,公司財測將不再計入中國市場,第二季營收損失約80億美元;若政策反轉,對輝達而言只是額外紅利。
閱讀文章 ↗AMD 發表 MI350 系列:4 倍算力、35 倍推論
2025 年 6 月 12 日,AMD 在 Advancing AI 活動發表 Instinct MI350 系列:288GB HBM3E、原生 FP4/FP6,宣稱對 MI300X 算力最高 4 倍、推論最高 35 倍;MI350X 與 MI355X 已量產出貨,並預告 2026 年 Helios 機架與 MI400。
閱讀文章 ↗輝達 GTC 巴黎:歐洲 AI 算力兩年要拚十倍
2025 年 6 月 11 日,黃仁勳在巴黎 VivaTech 的 GTC 主題演講宣布歐洲 AI 算力兩年內將成長十倍:Mistral 首階段部署 18,000 套 Grace Blackwell 系統、德國打造全球首座工業 AI 雲端、英國首階段 14,000 顆 Blackwell GPU,各國主權 AI 部署合計超過 3,000 EFLOPS。
閱讀文章 ↗
2026
41 ARTICLESA Single Persistent Kernel Changes How You Serve Code Models
Cohere's North Mini Code megakernel serving engine hits 62% of H100 memory bandwidth, 1.58× faster than vLLM at batch size 1.
READ POST ↗What Deploying Qwen3.8-2.4T-A95B on HyperPod Changes for Self-Hosting Frontier Models
A practical walkthrough for serving a 2.4T open-weights MoE on a single 8-GPU node with vLLM and SageMaker HyperPod.
READ POST ↗Apple M6 and M5 Ultra: 2nm and Quad-Die for On-Device AI
Apple's first 2nm chip (M6) and first quad-die design (M5 Ultra) arrive in Mac mini and Mac Studio, with up to 512GB unified memory for hundred-billion-parameter local LLMs.
READ POST ↗Cerebras CS-4: Three WSE-3T Wafers, 30x Faster Inference
Cerebras launched its CS-4 rack-scale system: the same wafer pushed twice as hard, up to 30x faster inference than GPUs, and prefill offloaded to AWS and AMD accelerators.
READ POST ↗Memory Prices Up 500% in a Year as AI Data Centers Eat DRAM
PCPartPicker data shows DDR5 prices up roughly 500% in 12 months — a 128GB kit went from $329 to $3,399. HBM crowds out standard DRAM, and new capacity lands only after 2028.
READ POST ↗Nvidia Cuts OpenAI Data-Center Guarantee to Under $120B
Nvidia expects to initially guarantee under $120B of OpenAI data-center debt, down from $250B discussed in July. Nothing was ever signed; the cut reads as a signal.
READ POST ↗RAMageddon: 2027 DRAM and HBM Capacity Reportedly Sold Out
A DIGITIMES report says Samsung, SK Hynix, and Micron have sold DRAM and HBM capacity through 2027 via five-year deals, mostly to AI buyers — and consumers are paying.
READ POST ↗AMD Buys Taalas: Etching AI Models Straight Into Silicon
AMD agreed to buy Toronto startup Taalas, whose chips etch model weights into silicon: Llama 3.1 8B at nearly 17,000 tokens per second on a 6nm test chip.
READ POST ↗NVIDIA Invests in SSI and Opens Up Vera Rubin Compute
NVIDIA announced a long-term partnership and equity investment in Sutskever's Safe Superintelligence, giving the secretive lab an order of magnitude more compute on Vera Rubin.
READ POST ↗Google Discloses $94.1 Billion SpaceX Stake, Up Over 100x
Alphabet's Q2 10-Q discloses $94.1 billion in SpaceX stock, about 6% and over 100x its 2015 investment — plus Google's dual role as shareholder and compute customer.
READ POST ↗Google's Frozen v2: A Chip Built for Gemini Efficiency
The Information reports Google is building a server chip codenamed Frozen v2 for Gemini, targeting a 2028 launch and six to ten times better efficiency than today's chips.
READ POST ↗Anthropic in Talks With Samsung for a Custom AI Chip
Anthropic is reportedly in talks with Samsung to build a custom AI chip, with purpose and specs undecided — a possible fourth supply line beyond TPUs, Trainium, and Nvidia GPUs.
READ POST ↗Etched Emerges With $1B in Orders and Made-at-TSMC Silicon
AI chip startup Etched emerges with $1 billion in booked orders, TSMC-manufactured silicon, and Frontier inference clusters now in customer testing.
READ POST ↗OpenAI's First Custom Inference Chip: Broadcom-Built Jalapeño, Nine Months to Production
OpenAI and Broadcom unveiled Jalapeño on June 24, 2026: OpenAI's first custom LLM-inference chip, designed to production in about nine months and claimed to beat NVIDIA's GB300 on key benchmarks.
READ POST ↗Amazon Weighs Selling Trainium Chips Beyond AWS
Amazon is in talks to sell Trainium chips beyond AWS; Jassy sizes the standalone chips business at ~$50 billion a year, taking direct aim at Nvidia's AI data center dominance.
READ POST ↗Google to Pay SpaceX $920M a Month for GPU Compute
SpaceX's June 5 filing shows Google paying $920M monthly through June 2029 for ~110,000 NVIDIA GPUs — bridge capacity for Gemini Enterprise, a week before the IPO.
READ POST ↗Ollama 0.30: GGUF Support, Up to 20% Faster NVIDIA, Vulkan On by Default
Ollama 0.30 deepens its GGUF engine via llama.cpp: up to 20% faster NVIDIA throughput under a stated RTX 5090 condition, Vulkan on by default, and tool calling that carries to coding agents.
READ POST ↗XCENA Raises $135M: AI's Bottleneck Is Memory, Not Compute
Korean startup XCENA raised $135M at a $570M valuation. Its MX1 chip puts thousands of RISC-V cores beside DRAM to manage KV caches in-module; Samsung ramps production late 2026.
READ POST ↗Nvidia's Vera CPU: Jensen Huang's $200B Bet on Agent Silicon
Jensen Huang says the Vera CPU opens a $200B market Nvidia never addressed before, with $20B in standalone sales this year. CPUs are the new front in the AI chip war.
READ POST ↗AMD Q1 2026: Data Center Up 57%, Server CPU Outlook Doubled
AMD Q1 2026: data center revenue $5.8bn, up 57% YoY; AMD doubled its server CPU outlook to $120bn+ by 2030. Inside the numbers, MI450 and Helios timing, and agentic AI demand.
READ POST ↗Cerebras IPO: $115–$125 Range, $3.5B Raise, $26.6B Cap
Cerebras set its IPO at $115–$125 for 28M shares — a $3.5B raise, $26.6B market cap at the top, 2026's biggest tech IPO so far. Bloomberg counts $10B in orders.
READ POST ↗Cerebras Files for IPO Again With $510M Revenue
Cerebras refiled for a US IPO with a mid-May target. The filing shows $510M of 2025 revenue, a reported $10B+ OpenAI compute deal, and an AWS deployment pact.
READ POST ↗Amazon's Custom Chips Pass $20B Run Rate in Jassy Letter
Jassy's April 9, 2026 letter: Amazon's Graviton, Trainium and Nitro chips top a $20B annual run rate, growing triple digits, and might one day sell racks to third parties.
READ POST ↗GPUs as Collateral: Forum's Bridge Loans for Neoclouds
On April 8, 2026, Forum Markets said it will fund neocloud GPU purchases with 60-120 day bridge loans — a $25M-$50M first deal, mid-teens returns — and tokenize the debt.
READ POST ↗Korea Puts $166M Into Rebellions in 'K-Nvidia' Chip Push
South Korea's National Growth Fund is putting 250 billion won ($166M) into NPU maker Rebellions — the first direct deal under its K-Nvidia plan for a homegrown AI chip champion.
READ POST ↗Meta and Nebius Sign Up to $27 Billion AI Compute Pact
Nebius and Meta signed a five-year deal worth up to $27 billion: $12 billion dedicated capacity plus up to $15 billion for residual cluster capacity, on NVIDIA's Vera Rubin platform from early 2027.
READ POST ↗GTC 2026 Opens: Vera Rubin's Seven Chips in Full Production, About $1T in Orders
At GTC 2026, Jensen Huang unveiled the Vera Rubin platform with seven new chips already in full production; CNBC tallies about $1T in combined Blackwell and Vera Rubin orders through 2027.
READ POST ↗Tesla Terafab: Musk's $20B AI Chip Fab Starts Its Countdown
On March 14, 2026, Musk posted that the Terafab Project launches in seven days. The reported $20B, 2nm-class fab targets 100,000 wafer starts a month to feed FSD and Optimus with AI5 silicon.
READ POST ↗Meta Unveils Four MTIA Chips: A Six-Month Silicon Cadence
Meta announced four MTIA chips on March 11, 2026 — hundreds of thousands already in production, 25x FLOPS gains across generations, and a chiplet platform built for a roughly six-month cadence.
READ POST ↗Ayar Labs Raises $500M Series E to Scale Co-Packaged Optics
On March 3, 2026, Ayar Labs closed a $500M Series E at $3.75B, with NVIDIA and AMD re-investing. CEO Mark Wade: AI infrastructure is hitting a power wall driven by interconnect inefficiency.
READ POST ↗Meta Signs Multibillion-Dollar Deal to Rent Google TPUs
Meta signed a multibillion-dollar deal to rent Google Cloud TPUs for next-generation models, and is negotiating to buy millions more for its own data centers. The AI chip market just changed shape.
READ POST ↗NVIDIA's Q4 FY2026: A Record $68.1B Quarter Caps a $215.9B Year
NVIDIA's Q4 FY2026 report on Feb 25: record quarterly revenue of $68.1B, up 20% QoQ, with $43B net income; full-year revenue $215.9B, up 65%. What the numbers say about AI capex and buyers.
READ POST ↗Meta Locks In Millions of Blackwell GPUs in NVIDIA Deal
Meta's multiyear NVIDIA partnership commits to millions of Blackwell GPUs, claiming supply years ahead and raising acquisition costs for everyone else.
READ POST ↗Cerebras Raises $1B Series H at a $23B Valuation
Cerebras closed a $1 billion Series H at about a $23 billion valuation, led by Tiger Global with AMD and Benchmark joining. What the raise says about wafer-scale inference and its IPO path.
READ POST ↗Microsoft's Maia 200: A 10-PetaFLOPS Bet on AI Inference
Microsoft's Maia 200, announced January 26, 2026, delivers 10+ petaFLOPS at FP4 in a 750W package and claims 3x Trainium3 FP4 throughput. Inside the specs and the custom-silicon cost math.
READ POST ↗TSMC's Record Q4 and $56B Capex Plan Show AI Demand Is Real
TSMC closed 2025 with a record quarter — revenue above NT$1 trillion for the first time, net income up 35% — and set 2026 capex guidance at $52–56 billion. The numbers behind the AI chip boom.
READ POST ↗OpenAI's $10B Cerebras Deal Adds 750MW of Inference Compute
OpenAI signed a Cerebras deal worth over $10 billion: 750MW of wafer-scale compute through 2028 for low-latency ChatGPT inference. What the terms, the G42 risk, and the supplier mix mean.
READ POST ↗Etched Raises $500M to Take On Nvidia With Transformer Chips
Etched raised about $500 million led by Stripes at a reported $5 billion valuation to take on Nvidia. Its TSMC 4nm Sohu ASIC runs transformer models from OpenAI, Google, Microsoft and Anthropic.
READ POST ↗OpenAI and SoftBank Put $1B Into SB Energy Alongside a 1.2 GW Stargate Lease
Around January 9, 2026, OpenAI and SoftBank each invested $500M in SB Energy — $1B total — alongside a 1.2 GW Stargate data-center lease. Compute and power, bound into one deal.
READ POST ↗Intel's Core Ultra Series 3 Brings 18A Chips to AI PCs
At CES 2026 Intel launched Core Ultra Series 3 (Panther Lake), its first AI PC platform on 18A, with Arc B390 graphics, 200+ laptop designs, and January 27 availability. What it changes.
READ POST ↗AMD at CES 2026: Helios Racks and the Yotta-Scale AI Play
At CES 2026 Lisa Su laid out AMD's AI Everywhere vision: Helios racks at up to 3 AI exaflops each, new MI455X and MI440X GPUs, a 2027 MI500 preview, and 60-TOPS Ryzen AI 400 chips.
READ POST ↗
2025
5 ARTICLESAI memory trade heats up: Micron and SK hynix stocks
On June 26, 2025, Micron shares rose on AI memory demand bets while SK hynix hit a record $157 billion market cap — memory chips moved to the center of the AI hardware trade.
READ POST ↗Micron fiscal Q3: HBM sold out for calendar 2025
Micron's record fiscal Q3: $9.3 billion revenue on June 25, 2025, with HBM up nearly 50% sequentially. CEO Mehrotra said HBM is sold out for calendar 2025; Q4 guidance topped estimates.
READ POST ↗Huang: Nvidia drops China from forecasts after H20 curbs
June 2025: Jensen Huang told CNN Nvidia's forecasts now exclude China after H20 export controls cost about US$8 billion in Q2 revenue; any policy reversal would just be a bonus.
READ POST ↗AMD's Instinct MI350 launch: 4x compute, 35x inference
At Advancing AI on June 12, 2025, AMD launched Instinct MI350X and MI355X GPUs with 288GB HBM3E, claiming up to 4x compute and 35x inference over MI300X, and previewed the 2026 Helios MI400 rack.
READ POST ↗Nvidia GTC Paris: Europe's AI compute set to grow tenfold
GTC Paris, June 11, 2025: Jensen Huang said Europe's AI compute will grow tenfold in two years, backed by 18,000 Grace Blackwell systems for Mistral and 3,000-plus exaflops of sovereign compute.
READ POST ↗