Brand: NVIDIA | Category: GPUs
SKU: NVH100NVLTCGPU-KIT | Part #: NVH100NVLTCGPU-KIT | MPN: NVH100NVLTCGPU-KIT
Contact for Pricing — Request a Quote
The PNY NVIDIA H100 NVL 94GB HBM3 PCIe Accelerator (MPN: NVH100NVLTCGPU-KIT) is built on NVIDIA's Hopper GPU architecture and is specifically engineered for large language model (LLM) inference and training at scale. The H100 NVL variant ships in a dual-GPU NVLink configuration delivering a combined 188GB of HBM3 memory and 7.8 TB/s aggregate memory bandwidth, enabling the acceleration of models that cannot fit within the memory envelope of a single GPU. The PCIe Gen5 host interface makes this solution broadly compatible with modern enterprise server platforms without requiring proprietary interconnect fabrics at the host level.
At the silicon level, the H100 NVL leverages the GH100 die with fourth-generation Tensor Cores supporting FP8, FP16, BF16, TF32, and FP64 precision, alongside the Transformer Engine—an NVIDIA capability that dynamically selects precision per layer to maximize throughput without sacrificing model accuracy. Compared to prior-generation A100 PCIe solutions, the H100 NVL delivers substantially higher throughput for transformer-based inference workloads, making it a practical choice for production AI serving infrastructure. NVIDIA's Confidential Computing capability, also present on Hopper, allows tenant isolation and data-in-use protection in multi-tenant cloud and enterprise deployments.
As a PCIe form-factor card distributed by PNY, the NVH100NVLTCGPU-KIT targets enterprise data centers seeking Hopper-generation performance within standard OCP and proprietary rack environments. Launched in Q1 2024, the H100 NVL has experienced persistent supply constraints driven by acute global demand for LLM training and inference infrastructure. Organizations deploying this accelerator commonly pair it with NVIDIA's software stack including CUDA 12.x, cuDNN, TensorRT-LLM, and NEMO frameworks to realize full hardware capability across generative AI, scientific computing, and high-performance data analytics pipelines.
| Manufacturer | NVIDIA |
| Brand | PNY (NVIDIA H100 NVL) |
| Manufacturer Part Number | NVH100NVLTCGPU-KIT |
| GPU Architecture | NVIDIA Hopper (GH100) |
| GPU Memory per Card | 94 GB HBM3 |
| GPU Memory Configuration | Dual-GPU NVLink pair — 188 GB HBM3 combined |
| Memory Bandwidth per Card | 3.9 TB/s |
| Aggregate NVLink Pair Memory Bandwidth | 7.8 TB/s |
| NVLink Interconnect Bandwidth | 600 GB/s bidirectional (NVLink 4.0) |
| Host Interface | PCIe Gen5 x16 |
| Tensor Core Generation | 4th Generation (FP8, FP16, BF16, TF32, FP64) |
| Transformer Engine | Yes (2nd Generation) |
| Supported Precisions | FP8, FP16, BF16, TF32, FP32, FP64, INT8 |
| Confidential Computing | Yes (NVIDIA Hopper Confidential Computing) |
| Form Factor | PCIe Dual-slot (per GPU card in NVLink pair) |
| Thermal Design | Passive cooling (requires adequate chassis airflow) |
| Max Thermal Design Power (per GPU) | 400 W |
| ECC Memory Support | Yes |
| CUDA Compute Capability | 9.0 |
| Launch Quarter | Q1 2024 |
Available from Omnixon Global. Submit an RFQ and our team will confirm configuration and availability for your order.
| Brand | NVIDIA |
| Category | GPUs |
| SKU | NVH100NVLTCGPU-KIT |
| Part Number | NVH100NVLTCGPU-KIT |
| Condition | New |
| Manufacturer Part Number | NVH100NVLTCGPU-KIT |
| GPU Architecture | NVIDIA Hopper (GH100) |
| GPU Memory per Card | 94 GB HBM3 |
| GPU Memory Configuration | Dual-GPU NVLink pair — 188 GB HBM3 combined |
| Memory Bandwidth per Card | 3.9 TB/s |
| Aggregate NVLink Pair Memory Bandwidth | 7.8 TB/s |
| NVLink Interconnect Bandwidth | 600 GB/s bidirectional (NVLink 4.0) |
| Host Interface | PCIe Gen5 x16 |
| Tensor Core Generation | 4th Generation (FP8, FP16, BF16, TF32, FP64) |
| Transformer Engine | Yes (2nd Generation) |
| Supported Precisions | FP8, FP16, BF16, TF32, FP32, FP64, INT8 |
| Confidential Computing | Yes (NVIDIA Hopper Confidential Computing) |
| Form Factor | PCIe Dual-slot (per GPU card in NVLink pair) |
| Thermal Design | Passive cooling (requires adequate chassis airflow) |
| Max Thermal Design Power (per GPU) | 400 W |
| ECC Memory Support | Yes |
| CUDA Compute Capability | 9.0 |
| Launch Quarter | Q1 2024 |
The PNY NVIDIA H100 NVL 94GB HBM3 PCIe Accelerator accelerates AI/ML training, inference, scientific HPC, and virtualization (vGPU) workloads. Typical deployments include LLM training clusters, computer-vision pipelines, financial risk modeling, and rendering farms.
The PNY NVIDIA H100 NVL 94GB HBM3 PCIe Accelerator accelerates AI/ML training, inference, scientific HPC, and virtualization (vGPU) workloads. Typical deployments include LLM training clusters, computer-vision pipelines, financial risk modeling, and rendering farms.
Key specifications for the PNY NVIDIA H100 NVL 94GB HBM3 PCIe Accelerator: new condition; manufacturer NVIDIA; brand PNY (NVIDIA H100 NVL); manufacturer part number NVH100NVLTCGPU-KIT; gpu architecture NVIDIA Hopper (GH100); gpu memory per card 94 GB HBM3; gpu memory configuration Dual-GPU NVLink pair — 188 GB HBM3 combined. Manufacturer part number NVH100NVLTCGPU-KIT. For the full datasheet with electrical, environmental, and compliance details, contact our pre-sales engineering team.
Key specifications for the PNY NVIDIA H100 NVL 94GB HBM3 PCIe Accelerator: new condition; manufacturer NVIDIA; brand PNY (NVIDIA H100 NVL); manufacturer part number NVH100NVLTCGPU-KIT; gpu architecture NVIDIA Hopper (GH100); gpu memory per card 94 GB HBM3; gpu memory configuration Dual-GPU NVLink pair — 188 GB HBM3 combined. Manufacturer part number NVH100NVLTCGPU-KIT. For the full datasheet with electrical, environmental, and compliance details, contact our pre-sales engineering team.
The PNY NVIDIA H100 NVL 94GB HBM3 PCIe Accelerator requires a PCIe Gen4 or Gen5 x16 slot, server power adequate for the card's TDP, and CUDA/ROCm driver support in your hypervisor or bare-metal OS. Sales engineering will confirm chassis fit (1U/2U/4U), PCIe lane count, and PSU headroom before quoting.