Rung 05 / How chips and sites talk

Networking & Interconnect

How chips and sites talk to each other.

900 GB/s

Fourth-generation NVLink bandwidth per H100 GPU — what one accelerator can say to its neighbours

4
Sources
1
Concepts
1
Entities
Paper2ReportAnalysisNews2
A brushed steel fibre coupler with copper collars, on paper.

In scope: optical interconnect and co-packaged optics, scale-up versus scale-out fabrics, NVLink, InfiniBand and Ethernet economics, bandwidth per dollar, topology, cross-datacenter links. Out: memory bandwidth inside a package (hardware), photonic compute (out of the knowledge base entirely).

A cable bundle of copper conductor and glass fibre terminating in one cartridge.
Optical interconnectCPOFabricsTopology

This rung has no compiled overview yet. Until it does, the base is best read across the stack rather than down one rung.

Start here

  1. 01
    LLM Inference Prices Have Fallen Rapidly but Unequally Across Tasks

    Start at the number the whole base chases: what a token costs, and how unevenly that price has actually fallen.

    inference-economics
  2. 02
    Memory-Centric Computing: A Paradigm Shift for Sustainable and Efficient Systems

    The reason the price falls the way it does — the bottleneck is moving data, not doing arithmetic on it.

    hardware
  3. 03
    AI Hyperscaler Capex 2026: Why Microsoft, Google, Meta and Amazon Are All Spending at Once

    Where the money physically lands once the bottleneck is priced: buildings, land and power, committed years ahead.

    datacenters
  4. 04
    Speculative Decoding Meets Quantization: Compatibility Evaluation and Hierarchical Framework Design

    How the cost is actually cut in software, and what the two cheapest tricks do to each other when you stack them.

    serving
  5. 05
    Agentic Reasoning for Large Language Models

    The demand side: the systems wrapped around the model are what decide how many tokens the price applies to.

    harnesses

Companies on this rung

All 23