NVIDIA

NVIDIA H100

80GB HBM3 Hopper Architecture — the enterprise benchmark for large-scale training and inference with NVLink scale-out

Reserved and on-demand, deployed in Thailand

$2.85/hr per GPU

All prices in USD, ex-tax.

Compare GPU models

Deploys high-density NVIDIA H100 SXM5 8-GPU HGX nodes with NVLink 4 interconnect and full root access in Tier III regional infrastructure — optimized for enterprise foundation model training, distributed fine-tuning, and high-throughput inference.

Memory
80GB HBM3
3.35 TB/s bandwidth
Architecture
Hopper
Family · NVIDIA Hopper
Compute (FP16)
7.9 PFLOPS (8-GPU, FP16 tensor)
BF16: 7.9 PFLOPS (8-GPU, BF16 tensor) · FP8: 15.8 PFLOPS (8-GPU, FP8 tensor)
Interconnect
NVLink 4 (900 GB/s) / NDR InfiniBand
TDP · 700W

Highlights

  • Enterprise HGX H100 SXM5 architecture with native NVSwitch all-to-all GPU mesh
  • NVLink 900GB/s in-node connectivity for multi-GPU training
  • Production-ready CUDA, NCCL, and TensorRT-LLM stack for distributed LLM workloads

Software stack

CUDA 12.x
PyTorch 2.x
Triton
vLLM
TensorRT-LLM
NCCL

Best-fit scenarios

  • Large-scale LLM training and fine-tuning
  • High-throughput inference (TensorRT-LLM / vLLM)
  • Mixed-precision scientific computing

Ready to deploy NVIDIA H100?

Spin up a dedicated server in days, or talk to our sales engineers about custom configurations.

Contact Us