
NVIDIA H100 80GB PCIe Gen5
NVIDIA H100 Tensor Core GPU in PCIe form factor with 80GB HBM3 memory. Ideal for deploying AI inference and training in standard servers without NVLink clustering requirements. Contact our sales team for volume pricing and current lead time.
Price Available Upon Request
Custom configuration and pricing
What it's for
The NVIDIA H100 Tensor Core GPU delivers high performance, scalability, and security for every workload. Built on the NVIDIA Hopper architecture, H100 features Transformer Engine technology that accelerates large language model training and inference. Available in PCIe form factor, it brings datacenter-class AI to standard enterprise servers. Request a quote for configurations and lead time.
Key Features
- ✓80GB HBM3 memory with 3 TB/s bandwidth
- ✓Fourth-generation Tensor Cores with FP8 precision
- ✓1,513 TFLOPS FP8 performance with Transformer Engine
- ✓989 TFLOPS FP16 Tensor Core performance
- ✓PCIe Gen5 x16 interface for broad server compatibility
- ✓Secure Boot and confidential computing ready
- ✓350W TDP with passive cooling support
- ✓Full NVIDIA AI Enterprise software stack compatibility
Use Cases
- →Large Language Model (LLM) inference at scale
- →Generative AI model deployment
- →Multi-tenant AI infrastructure
- →Recommendation systems
- →Computer vision and video analytics
- →Scientific computing and simulation
Technical Specifications
| Architecture | Hopper |
| GPU Memory | 80 GB HBM3 |
| Memory Bandwidth | 3 TB/s |
| FP64 Performance | 60 TFLOPS |
| FP32 Performance | 120 TFLOPS (TF32) |
| FP16 Performance | ~989 TFLOPS |
| FP8 Performance | ~1,513 TFLOPS |
| INT8 Performance | ~3,026 TOPS |
| CUDA Cores | 16,896 |
| Tensor Cores | 528 (4th Gen) |
| Max TDP | 350W |
| Thermal Solution | Passive (blower available) |
| Form Factor | Dual-Slot PCIe |
| PCIe Interface | PCIe Gen5 x16 |
Related Products

NVIDIA H100 80GB SXM5
NVIDIA H100 Tensor Core GPU in SXM5 form factor with NVLink for multi-GPU scaling. Designed for HGX server platforms and large-scale AI training clusters. Enterprise volume discounts available - contact sales for custom configurations.
From
$32,999

NVIDIA H200 141GB HBM3e SXM5
Hopper GPU with 141GB HBM3e and 4.8TB/s bandwidth — the memory-bandwidth pick for LLM inference. Contact us for current lead time.
From
$39,999

NVIDIA B200 192GB Blackwell
Blackwell architecture with 192GB HBM3e and FP4 precision for next-gen AI. Contact us for availability and pricing.
Request pricing

AMD Instinct MI300X 192GB HBM3 OAM
AMD's flagship AI accelerator with 192GB HBM3 — the largest single-GPU memory in the class — and 5.3TB/s bandwidth. Strong performance per dollar for generative AI and large language models. Quoted per project with ROCm support; lead time confirmed at quote.
From
$29,999
Recommended Reading
The adjacent decisions buyers usually miss
Relevant guides tied to this product's real deployment questions: VRAM, power, cooling, financing, and market timing.

H100 vs MI300X: Complete Buyer's Guide (2026 Update)
NVIDIA H100 or AMD MI300X? Compare performance, pricing, TCO, and real-world benchmarks. Includes LLM training data, software ecosystem analysis, MI350X preview, and buying recommendations.

AMD MI350X vs NVIDIA B200: Which Next-Gen AI GPU Should You Buy?
Side-by-side comparison of AMD MI350X (288GB, CDNA 4) and NVIDIA B200 (192GB, Blackwell): specs, inference benchmarks (MI350X 20-30% faster on large models), training performance, pricing (~50% cheaper), software ecosystem, and cooling requirements.

H100 PCIe vs SXM: When Does Form Factor Matter?
Deep dive into H100 PCIe vs SXM differences: NVLink bandwidth, memory throughput, power consumption, and multi-GPU scaling. Decision framework for infrastructure engineers.