Supercomputers/Systems/Together AI / Hypertec Cloud GB200 Cluster
USA · operational · ai training
Together AI / Hypertec Cloud GB200 Cluster
Operated by Together AI at Together AI / Hypertec (undisclosed site) .
Phase 1 seed dataset, compiled by hand. These rows were built from public operator, laboratory and vendor sources. A mechanical second-reader pass has since fetched every cited source: 1092 of 1153 systems have a readable citation that names them, and 40 are genuinely weakly sourced. Every claim carries its source and a confidence tier. Treat anything below verified as a lead, not a citation.
Measured performance
- Rmax
- -
- Rpeak
- -
- Rmax ÷ Rpeak
- -
- Cores
- -
- Power
- -
- Per watt
- - GF/W
Figures are as last publicly reported for the configuration described below, not a live measurement. Where a system was upgraded in place, the post-upgrade configuration is shown and the earlier one appears in the timeline.
What this machine is made of
One row per supplier relationship. “Supplier at build” is the company that shipped the part at the time; where that company has since been acquired, the parent it rolls up to today is shown beside it. That distinction is what makes ticker-level aggregation possible across a thirty-year dataset.
| Role | Part | Supplier at build | Quantity | Confidence | Source |
|---|---|---|---|---|---|
| Integrator | - Cluster co-built by Hypertec Cloud with Together AI | Hypertec Cloud | - | Reported | together.aiprnewswire.com |
| Accelerator | NVIDIA GB200 NVL72 Accelerators Arm CPU + CUDA GPU · TSMC 4NP 36,000+ NVIDIA GB200 NVL72 GPUs starting Q1 2025 | NVIDIA | 36,000 reported | Reported | together.aiprnewswire.com |
| Interconnect | NVIDIA InfiniBand NDR (Quantum-2) Interconnect Fat tree or dragonfly+ · 400 Gb/s per port · Under 0.6 microseconds NVIDIA NVLink and InfiniBand, up to 1.8 TB/s GPU-to-GPU bandwidth per the GB200 NVL72 rack | NVIDIA | - | Reported | together.ai |
AI datacenter metrics
Figures beyond an accelerator count, each with who stated it. What can and cannot be compared across AI datacenters.
| Metric | Value | As of | Basis | Note | Source |
|---|---|---|---|---|---|
| Accelerators | 36,000 accelerators | 2024-11 | Design target | Together AI and Hypertec Cloud announced a 36,000 GB200 NVL72 GPU cluster for availability from Q1 2025. Announcement; delivery not confirmed on a fetched page. | together.ai |
The read
A cluster of more than 36,000 NVIDIA GB200 NVL72 GPUs that Together AI and Hypertec Cloud announced on 18 November 2024 they would co-build starting Q1 2025, complementing existing H100 and H200 capacity the two companies already operate across North America and Europe. Together AI's own blog and the joint press release (syndicated by PR Newswire) are the same underlying announcement rather than independent confirmation, so the tier stays reported. Neither names an exact city, so the site is recorded as an undisclosed North American location and no map coordinates are set. No FLOPS or HPL figure has been published for it.
Site & facility
Change history
- 2026-10-04 Added AI datacenter metrics for together-hypertec-gb200
- 2026-09-24 Added the Together AI / Hypertec Cloud 36,000-GPU GB200 NVL72 cluster
Source check
We fetched this system's own citations and recorded whether each page actually mentions it. This is a corroboration signal, not a fact check, and it is published so you can see how well the sourcing holds up rather than take it on trust.
| Cited source | Result | Found on the page |
|---|---|---|
| prnewswire.com | names system, part and count | NVIDIA GB200 NVL72 · counts: 36000 |
| together.ai | names system, part and count | NVIDIA GB200 NVL72, NVIDIA InfiniBand NDR (Quantum-2) · counts: 36000 |
Checked 2026-10-09 by pnpm verify. Re-run it and the table changes with the web.
Further reading
Inside a node
Why accelerators exist, what the memory hierarchy costs, and how the CPU and accelerator merged onto one package.
The interconnect
Topologies, why latency and tail behaviour matter more than bandwidth, and how the fabric market consolidated.
Operating a supercomputer
Acceptance testing, why "installed" and "in production" are six months apart, and the machine lifecycle.