VRAM
432 GB
Memory
HBM4
Bandwidth
19600 GB/s
TDP
-
Large Language Models
Training and inference for models like GPT-4, Llama 70B+
Deep Learning Training
High-performance training for neural networks
High-Throughput Inference
Optimized for batched inference workloads
Enterprise Deployment
Designed for 24/7 datacenter operations
Launched 2026-07-23 at Advancing AI. 40 PF FP4 / 20 PF FP8 dense per AMD (FP16 derived at half FP8). Helios rack platform; shipments ramp from Q4 2026 into H1 2027.
Estimates based on INT8 quantization. Actual fit depends on framework and batch size.
Added Jul 30, 2026
Last updated: Jul 30, 2026
Every month new models become cheaper, faster, and more capable. Inferbase ensures your application automatically benefits without changing a single API call.