Quick Answer
NVIDIA's post-Rubin roadmap announced at GTC 2026 includes Kyber (Vera Rubin Ultra with vertical 144-GPU racks, shipping H2 2027), Feynman (next-next-gen platform, expected 2028), and CPO switches entering volume production for Rubin Ultra (NVIDIA newsroom, 2026).
Data last verified September 9, 2026 from NVIDIA newsroom and the GTC 2026 keynote.
Why a post-Rubin roadmap matters
Vera Rubin is shipping in Q3 2026. But the AI compute demand curve is steep — inference workloads are doubling every 9-12 months, and physical AI (robotics, autonomous vehicles, digital twins) adds a new growth vector. To stay ahead, NVIDIA needs to announce the next two generations publicly so customers can plan multi-year infrastructure investments (NVIDIA, 2026).
Kyber — Vera Rubin Ultra (2027)
Kyber is NVIDIA's next-rack architecture. Key changes vs Vera Rubin:
- 144 GPUs per rack (vs 72 in Vera Rubin NVL72) — doubles compute density.
- Vertical compute trays — shorter cable length reduces latency by ~40%.
- CPO switches — co-packaged optics integrated into the switch ASIC, replacing pluggable transceivers.
- Liquid-cooled fully integrated — direct-die cooling on every GPU.
Expected shipping: H2 2027. Pricing TBA (NVIDIA, 2026).
Feynman — Post-Rubin platform (2028)
Feynman is named after the Nobel-winning physicist Richard Feynman — NVIDIA's tradition of naming architectures after scientists (Pascal, Volta, Turing, Hopper, Ampere, Ada Lovelace, Grace Hopper, Blackwell, Rubin). Details are limited, but the key directions:
- Next-generation HBM — likely HBM5 (Samsung/SK Hynix roadmap, 2028).
- Optical interconnect — silicon photonics for chip-to-chip communication.
- Sub-2nm process — TSMC A16 or N2 process node.
CPO switches — the network inflection
Co-Packaged Optics is one of the most consequential technology shifts in AI infrastructure. Traditional pluggable transceivers (QSFP-DD, OSFP) consume ~3-5 pJ/bit. CPO reduces this to <1 pJ/bit, saving 30-50% of network power in a 100k-GPU cluster (NVIDIA, 2026).
For Vera Rubin Ultra, NVIDIA is integrating CPO directly into the NVLink Switch 6 ASIC — first-of-its-kind in volume production. Competitors (Broadcom Tomahawk, Cisco Silicon One) are following with their own CPO roadmaps for 2027.
Competitor roadmap comparison
| Vendor | 2026 | 2027 | 2028 |
|---|---|---|---|
| NVIDIA | Vera Rubin (72-GPU rack) | Kyber / Vera Rubin Ultra (144-GPU) | Feynman |
| AMD | MI355X (rack-scale) | MI400 series | MI500 series |
| Google TPU | v6 Trillium | v7 (CPO integration) | v8 (silicon photonics) |
| AWS Trainium | Trainium 3 (NV networking) | Trainium 4 | — |
Source: vendor announcements + NVIDIA GTC 2026 (2026).
Why the rack-scale shift matters
The shift from per-GPU to per-rack performance is the defining trend of 2026-2028 AI infrastructure. Customers now buy "an NVL72" or "a Kyber rack" as a unit, not "2,000 GPUs." This changes procurement, deployment, networking, and cooling — every data center built in 2026-2027 will be sized for rack-scale AI. NVIDIA's roadmap positions them to lead that shift through 2028.
For the broader NVIDIA growth thesis, see our analysis of Jensen Huang's $1T order outlook.
What Kyber means for data center design
Vertical compute trays change the rack-level design. Today's racks are wide (24-inch wide, 42U tall) with horizontal compute trays. Kyber's vertical trays require new rack dimensions:
- Width - same (24-inch wide for data center compatibility).
- Depth - 48-inch deep (vs 36-inch today) for vertical tray clearance.
- Height - 52U tall (vs 42U) to accommodate the vertical orientation.
- Weight - about 3,500 lbs per rack (vs about 2,500 today) due to dense GPU packaging.
Existing data centers can accommodate Kyber with retrofits; new builds (2027+) will spec for the larger form factor.
Why CPO matters
Co-Packaged Optics (CPO) is the most consequential networking change in 10 years:
| Metric | Pluggable (QSFP-DD) | CPO |
|---|---|---|
| Power per bit | 3-5 pJ | less than 1 pJ |
| Latency | about 5 ns (transceiver) | less than 1 ns (on-die) |
| Bandwidth density | 1x baseline | 4-8x baseline |
| Cost per port | $500-1000 | $200-400 |
Source: NVIDIA + Broadcom CPO roadmap (2026).
Competitor CPO timelines
- NVIDIA NVLink Switch 6 (Vera Rubin Ultra, 2027) - first volume CPO.
- Broadcom Tomahawk 6 (2026) - CPO for hyperscaler Ethernet.
- Intel Silicon Photonics (2027) - PCIe over optics.
- AMD Ultra Ethernet (2027) - optical fabric for MI400 racks.
Why this matters for the broader AI ecosystem
This announcement fits into a larger pattern of the 2026 AI industry consolidation wave. Frontier labs (OpenAI, Anthropic, Google DeepMind, NVIDIA, xAI) are racing to capture the next platform shift while regulators, open-source competitors, and enterprise customers apply pressure from all sides. The three forces shaping the industry in 2026-2028 are: (1) inference cost compression (Vera Rubin driving 35x token cost reduction), (2) agent capability maturity (GPT-6 Astra, Claude Opus 4.5, Gemini 3.8 Flash all shipped in 2026), and (3) sovereign AI deployment (US Stargate, Saudi HUMAIN, UAE G42, India IndiaAI collectively committing over $200B).
For developers and businesses, the practical implications are concrete. Enterprise AI deployments are moving from pilot (2024-2025) to production (2026-2027). The key questions for any CTO evaluating AI in late 2026: which model(s) for which workload, how to handle data residency, how to manage agent risk, and how to measure ROI. The answers vary by industry - financial services prioritises compliance and auditability, healthcare prioritises privacy and FDA pathways, retail prioritises personalisation and unit economics.
What to watch next
Three upcoming events will validate or revise this analysis:
- NVIDIA GTC Berlin (Oct 20-22, 2026) - European AI sovereignty + Vera Rubin EU rollout.
- Made by Google October 2026 - Pixel 11, Gemini Spark 2, Android XR 2 launch.
- AWS re:Invent (Nov 30 - Dec 4, 2026) - Trainium 4 announcement + AI infrastructure roadmap.
Cross-references
For related TutorsBot coverage, see our guides on Jensen Huang's $1T order outlook, Vera Rubin shipping Q3 2026, and the 2026 AI chip war landscape. For the broader market context, our analysis of NVDA's $4.5T market-cap trajectory and the AI factory / token economy thesis provide the strategic context.






