Newsletter FIT
Dashboard Search Watchlist Alerts Billing
SEARCH // ⌘K
Sign in Get started

Billing Watchlist

Article // Intelligence

SemiAnalysis · Sep 7, 2026 · Dylan Patel

TPU Inference Externalization Full Steam Ahead - InferenceX

InferenceX, Up to 50% Better Performance per Dollar, Rapid Externalization of TPU stack, Growing Customer Base, Ironwood, TPUv8i, Reducing CUDA Moat

Open original article View publication

What's interesting

  • Mentions Anthropic
  • Mentions Google
  • Mentions DeepMind
  • Discusses FP4
  • Discusses FP8
  • Discusses GPU

Engagement

Likes: 124

Comments: 1

Restacks: 10

Entities

Topics

FP4 FP8 GPU TPU TPU externalization Performance per dollar / inference economics TPUv7 Ironwood vs NVIDIA B200/B300 TorchTPU software stack Open-weight model serving vLLM/SGLang TPU support TPU total cost of ownership Anthropic TPU adoption

Companies

Anthropic SGLang vLLM TorchTPU Gemini B300 TPUv7 Ironwood Google DeepMind NVIDIA AMD SemiAnalysis Inferact Red Hat RadixArk B200

Sponsorship

No sponsorship detected

Related articles

What is So Hard About Behind-The-Meter Power For Datacenters? Part 1

SemiAnalysis · Google · Same publication

🍻 Tretas da bolha tech da semana

@manodeyvin · Google

The AI Trade is a Psyop

Notes from the Circus · Google

The Tribute

Notes from the Circus · Google

When Biden Dies

Political Currents by Ross Barkan · Google

Meta Settled a Massive Lawsuit Around Harming Kids. Then It Threw Itself a Parade.

Techish · Google

Your creator is on competitors radar

The Playbook · Google

AI Doesn’t account for past wins

The Playbook · Google

SEARCH // ESC

Type to search