Brand: Gigabyte | Category: GPUs
SKU: GV-NH200NVL-141G | Part #: GV-NH200NVL-141G | MPN: GV-NH200NVL-141G
Contact for Pricing — Request a Quote
PCIe Gen5 x16 interface with server-grade 10-pin PCIe power connector determines chassis and PSU compatibility. The Gigabyte NVIDIA H200 NVL 141GB HBM3e PCIe GPU Card (part number GV-NH200NVL-141G) brings enterprise-class AI acceleration to demanding datacenter workloads. Built on NVIDIA's Hopper architecture (GH200), this full-height, full-length dual-slot accelerator delivers 141 GB of HBM3e memory paired with a 4480-bit memory bus and 4.8 TB/s bandwidth—enabling rapid processing of massive datasets. The card houses 16,896 CUDA cores across 132 streaming multiprocessors, supported by 4th generation Tensor Cores with multi-precision compute capabilities: FP8 reaches 3958 TOPS, FP16 and BF16 each deliver 1979 TFLOPS, TF32 provides 989 TFLOPS, and FP64 achieves 67 TFLOPS. Support for FP64, FP32, TF32, BF16, FP16, FP8, INT8, and INT4 precisions ensures flexibility across machine learning, scientific computing, and high-performance analytics applications.
AI infrastructure teams, machine learning engineers, and GPU-accelerated research institutions purchase the Gigabyte H200 NVL to overcome memory and bandwidth bottlenecks in transformer model training, large language model inference, and enterprise AI pipelines. The 2nd generation Transformer Engine accelerates sparsity and mixed-precision workloads critical to modern deep learning. Passive thermal design requires forced-air datacenter cooling infrastructure, positioning this card for purpose-built server environments rather than edge deployments.
Organizations evaluating this accelerator for production AI clusters should contact Omnixon Global to request pricing, availability, and configuration guidance for the GV-NH200NVL-141G. Our enterprise solutions team stands ready to support your RFQ and technical validation process.
| Brand | Gigabyte |
| Category | GPUs |
| SKU | GV-NH200NVL-141G |
| Part Number | GV-NH200NVL-141G |
| Condition | New |
| Manufacturer Part Number | GV-NH200NVL-141G |
| GPU Architecture | NVIDIA Hopper (GH200) |
| GPU Memory | 141 GB HBM3e |
| Memory Bus Width | 4480-bit |
| Memory Bandwidth | 4.8 TB/s |
| Interface | PCIe Gen5 x16 |
| Tensor Core Generation | 4th Generation |
| CUDA Cores | 16896 |
| Streaming Multiprocessors | 132 |
| FP8 Tensor Core Throughput | 3958 TOPS |
| FP16 Tensor Core Throughput | 1979 TFLOPS |
| BF16 Tensor Core Throughput | 1979 TFLOPS |
| TF32 Tensor Core Throughput | 989 TFLOPS |
| FP64 Tensor Core Throughput | 67 TFLOPS |
| Supported Precisions | FP64, FP32, TF32, BF16, FP16, FP8, INT8, INT4 |
| Transformer Engine | 2nd Generation |
| Thermal Design | Passive (forced-air datacenter cooling required) |
| Form Factor | Full-height, full-length (FHFL) dual-slot |
| Power Connector | 10-pin PCIe power (server-grade) |
| ECC Memory Support | Yes |
| NVIDIA Software Ecosystem | CUDA, cuDNN, TensorRT, NCCL, Triton Inference Server |
Reference servers include Dell PowerEdge XE9680 / XE9712, HPE Cray XD670, Lenovo ThinkSystem SR685a / SR675 V3, Supermicro AS-A21GE / SYS-821GE, Gigabyte G593 / G894, ASUS ESC. Share your target platform in the RFQ and we will confirm chassis-to-GPU compatibility and recommended NIC pairing.
Highly model-dependent. L40S / RTX-class: typically 3-6 weeks. H100/H200/B200 in SXM form factor: 12-16 weeks for whole-platform allocations. We quote genuine-channel ETAs only — no grey-market promises.
Yes — Omnixon stocks the full NVIDIA networking lineup (Quantum-2 / Quantum-X InfiniBand, Spectrum-X Ethernet, ConnectX NICs, BlueField DPUs) so we can quote a complete training-cluster BOM, not just the GPUs.
Yes. We hold genuine-channels for NVIDIA AI Enterprise software subscriptions. Add it to your RFQ and we quote node-aligned licensing along with the hardware.