NVIDIA GB300 Blackwell Ultra Launches: 288 GB HBM3e, NVLink 5 at 1.8 TB/s
In one sentence NVIDIA begins shipping the GB300 Blackwell Ultra GPU featuring 288 GB HBM3e per chip, NVLink 5 at 1.8 TB/s, and double the FP8 throughput of the B200, dramatically lowering inference costs for frontier AI models.
Picture the most powerful engine ever built for artificial intelligence — that is essentially what NVIDIA has released with the GB300 Blackwell Ultra GPU. This is not a graphics card for gaming; it is a specialized processor designed to run the large AI models that power chatbots, translation tools, image generators, and more.
The headline number is 288 gigabytes of ultra-fast memory sitting directly on the chip. That is roughly 24 times the RAM in a typical smartphone, all accessible at blistering speeds thanks to a memory technology called HBM3e. More memory means bigger AI models can be loaded and run without splitting them across dozens of chips, which saves time and money.
In datacenters, these GPUs are never used alone. They work in racks of eight or more, connected at high speed. NVIDIA upgraded that connection too: NVLink 5 moves data between chips at 1.8 terabytes per second, fast enough to transfer a full 4K movie library in under a second.
Why does this matter beyond the datacenter walls? Because running a frontier AI model is expensive. Every question you ask a chatbot, every image you generate, costs fractions of a cent in compute time. With more powerful hardware, those fractions get smaller. NVIDIA claims the GB300 cuts inference cost per token by 40-50% compared to the previous B200 generation, meaning cheaper and faster AI services for end users.
The first units have already shipped to major hyperscalers like Google, Microsoft, and Amazon. AMD is competing with its MI350 chip, but NVIDIA retains a clear lead in both raw specs and software ecosystem maturity.
Companies
NVIDIA, AMD
Tools
GB300 Blackwell Ultra, NIM
Tags
Sources