Google Designs a Gemini-Specific AI Chip Called Frozen v2
Google is designing Frozen v2, a chip that hardwires parts of its Gemini model into silicon and could serve six to ten times more tokens per watt, targeting deployment by 2028.
XLA, short for Accelerated Linear Algebra, is a compiler designed to optimize tensor computations used by machine learning frameworks. It can combine operations, remove redundant work, plan memory use, and generate efficient code for CPUs, GPUs, and specialized accelerators. These transformations may reduce training time, inference latency, and memory overhead without changing the model’s intended mathematics. Benefits vary by workload, shapes, framework integration, and hardware, and compilation itself adds time that may not help short-lived jobs. Numerical behavior can also differ slightly because operations are reordered or fused. Teams benchmark complete applications, verify accuracy, and retain fallback paths when an optimization is unsupported or unstable.
Google is designing Frozen v2, a chip that hardwires parts of its Gemini model into silicon and could serve six to ten times more tokens per watt, targeting deployment by 2028.
OpenAI launched Codex Micro, a $230 mechanical keypad built with Work Louder that gives developers physical controls for its Codex coding agent, its first hardware to reach market.
SK Hynix raised $26.5 billion in its Nasdaq debut, the largest US listing ever by a foreign company, as investors rushed to buy into the AI memory chip boom.
Meta plans to begin manufacturing its in-house Iris AI chip in September as it aims to double computing capacity to 14 gigawatts next year and cut its reliance on Nvidia.
Apple expanded its partnership with Broadcom in a multi-year deal worth more than $30 billion, its largest US manufacturing commitment, to produce over 15 billion American-made chips.
Nvidia’s Kyber rack for its 2027 Rubin Ultra chips has slipped more than a year to 2028 over a hard-to-manufacture circuit board, SemiAnalysis says, with no proven fallback.
OpenAI teased its first branded hardware, Codex Micro, a programmable macro pad built with Work Louder to give developers physical shortcuts for its Codex coding agent, launching July 15.
A record European heatwave is highlighting a growing threat to the AI boom: extreme weather that stresses data centers and the power grids they depend on at once.
SK Hynix filed to raise about $29 billion through a Nasdaq ADR listing set for July 10, roughly double earlier estimates, to fund its AI memory expansion.
Qualcomm agreed to acquire Modular, the AI software startup led by Chris Lattner, to build a hardware-agnostic software layer and challenge Nvidia in data center AI.