TSMC
Where most AI chips are manufactured.
TSMC is a chip foundry company based in Hsinchu, Taiwan. The world's largest semiconductor foundry.
The AI supply chain in 7 terms · foundry, memory, chip, system, training, inference, pricing.
Where most AI chips are manufactured.
TSMC is a chip foundry company based in Hsinchu, Taiwan. The world's largest semiconductor foundry.
High bandwidth memory, the throughput backbone.
HBM3e (High Bandwidth Memory 3e (Enhanced)) is a JEDEC HBM memory generation from 2024, with 1,180 GB/s of bandwidth per stack.
Where the silicon becomes an accelerator.
A GPU is the accelerator that trains and serves most large AI models.
“A GPU is the accelerator that trains and serves most large AI models.”
Read full chapterMany chips wired into one rack-scale system.
DGX GB200 NVL72 is a NVIDIA AI system with 72 NVIDIA B200 and 36 NVIDIA Grace.
What these systems are used for first.
The phase where a model learns from internet-scale text via next-token prediction · typically 15-30T tokens and millions of GPU hours.
“Pretraining is the moat.”
Read full chapterWhat they are used for every day after that.
The process of running a trained model to generate predictions · every API call is inference.
“Inference optimization is where the next 10× cost reduction lives. Every frontier lab is racing to ship the best serving stack.”
Read full chapterHow model compute finally becomes a price.
The tokens in your prompt · billed per million, typically 3-5× cheaper than output tokens.
“Input token discipline separates teams that can scale from teams that can't. Cache aggressively.”
Read full chapterBy the end you understand the full stack from silicon to inference API · and why each layer sets pricing for the one above it.
Seven terms that decode whether AI is overpriced, fairly priced, or criminally underpriced. Read in order.