Marvell Issues Google 2.2B Stock Warrant in Custom AI Silicon Deal

Marvell Technology has granted Alphabet's Google a warrant to purchase up to 58.9 million shares of common stock at an exercise price of $206.58 per share, establishing an equity arrangement valued at up to $12.18 billion. The agreement expands the companies' partnership to co-develop custom artificial intelligence silicon, specialized networking, and next-generation datacenter infrastructure. Following the announcement, Marvell shares rose more than 11% in premarket trading, while primary cust

1 min
Marvell Issues Google 2.2B Stock Warrant in Custom AI Silicon Deal

Marvell Technology has granted Alphabet's Google a warrant to purchase up to 58.9 million shares of common stock at an exercise price of $206.58 per share, establishing an equity arrangement valued at up to $12.18 billion. The agreement expands the companies' partnership to co-develop custom artificial intelligence silicon, specialized networking, and next-generation datacenter infrastructure.

Following the announcement, Marvell shares rose more than 11% in premarket trading, while primary custom silicon competitor Broadcom traded down approximately 3%.

Expanding Custom AI Silicon and Near-Memory Architectures

The commercial agreement deepens Google's custom Application-Specific Integrated Circuit (ASIC) development pipeline beyond its existing Tensor Processing Unit (TPU) programs. Under the expanded framework, Marvell will develop:

  • Dedicated AI inference accelerators optimized for large-scale model serving
  • Advanced storage controllers and high-bandwidth interconnect solutions
  • Memory interface controllers and near-memory computing architectures
  • Optical interconnects and scale-up datacenter networking switching infrastructure
Marvell and Google Custom AI Silicon Architecture

Hyperscaler Silicon Diversification

The arrangement underscores hyperscaler strategies to diversify hardware suppliers and control datacenter economics. While Google maintains a long-term agreement with Broadcom extending through 2031 to co-develop TPU architectures and next-generation compute racks, the Marvell partnership broadens Google's architectural alternatives for inference workloads, datacenter interconnects, and specialized memory subsystems.

Custom silicon designs allow cloud providers to avoid general-purpose GPU premiums, optimize power envelopes per token, and tailor hardware specifically to transformer attention patterns and mixture-of-experts routing. Structuring commercial commitments around multi-billion-dollar equity warrants aligns long-term manufacturing allocation and design roadmaps across multi-year hardware generations.

Sources

Written by

More to read

  • Cache-Aware Load Balancing in Production LLM Serving: Architecture, Prefix Affinity, and Multi-Replica Routing Trade-Offs

    Cache-Aware Load Balancing in Production LLM Serving: Architecture, Prefix Affinity, and Multi-Replica Routing Trade-Offs When scaling large language model inference across multiple GPU worker nodes, standard Layer-4 and Layer-7 load balancing algorithms create an unseen performance cliff. Round-robin, least-connections, and random routing distribute HTTP/gRPC requests uniformly across compute replicas. However, modern LLM inference engines rely on prompt caching mechanisms, such as vLLM Automa

    1 min
  • SwiGLU and Gated Linear Units: How Bilinear Gating Replaced Standard FFNs in Modern LLMs

    Every modern open-weight and frontier large language model, from Meta's LLaMA 3 and Mistral to Alibaba's Qwen 2.5 and DeepSeek-V3, has abandoned the standard two-layer Feed-Forward Network (FFN) originally introduced in the 2017 Transformer architecture. In its place, model architectures have converged almost universally on Gated Linear Units (GLU), specifically the Swish-Gated Linear Unit (SwiGLU). While the original Transformer relied on standard non-linear activations like ReLU or Gaussian E

    1 min
  • Nvidia Acts as Matchmaker for Nordic Datacenter Capacity to Ease AI Compute Bottlenecks

    Nvidia is directly brokering compute infrastructure deals by connecting enterprise customers holding graphics processing units with datacenter operators in the Nordic region that possess available power, cooling, and floor capacity, according to reporting by CNBC. The matchmaking initiative reflects Nvidia's efforts to mitigate severe power grid bottlenecks in North America and Western Europe that threaten to stall AI cluster deployments. By pairing hardware buyers directly with site operators

    1 min