NVIDIA H200 NVL 141GB HBM3e PCIe 5.0
€27,273.61
Product information "NVIDIA H200 NVL 141GB HBM3e PCIe 5.0"
The NVIDIA H200 NVL 141GB is a high-end computing accelerator card designed for demanding AI, HPC, and data center workloads. Based on the NVIDIA Hopper architecture and equipped with 141 GB of HBM3e ECC memory, it offers enormous memory bandwidth and computing power for large language models (LLMs), generative AI, machine learning, scientific simulations, and data-intensive applications.
The exceptionally large HBM3e memory enables the processing of extensive AI models and datasets directly on the GPU. Multiple H200 NVL GPUs can be interconnected via NVLink for appropriately scalable workloads. The passive cooling is designed for use in dedicated server and workstation systems with sufficient system airflow.
The H200 NVL has no display outputs and is designed as a compute accelerator for GPU computing. It is particularly well-suited for professional systems where maximum AI computing power, large GPU memory, and high memory bandwidth are paramount.
Package Contents:
• NVIDIA H200 NVL 141GB PCIe
| AI (Artificial Intelligence): | Ja |
|---|---|
| Bandwith: | 1280 GB/s |
| Bus Interface: | PCI-Express 5.0 x16 |
| CUDA: | 11.6 |
| Cooling: | Passiv (Kühlkörper) |
| Data Center: | Ja |
| DirectX: | 12 |
| ECC-Support: | Ja |
| FP64 (Double Precision Performance): | 12.04 TFLOPS |
| GPU Architecture: | Hopper |
| GPU Name: | GH100 |
| Memory Interface Connection: | 5120 bit |
| NVLink (Scalable Link Interface): | Ja |
| Number of Shading Units: | 7296 |
| OpenCL: | 3.0 |
| OpenGL: | 4.6 |
| Power Consumption: | 350 Watt |
| RAM Size: | 80 GB |
| RAM Type: | HBM2e |
| Single Precision Performance (FP32): | 24.08 TFLOPS |
| Slot Width: | Full Size ATX / Dual-Slot |
| Vulkan: | 1.3 |
Cross-Selling
The NVIDIA L40S is a high-performance data center GPU designed for professional visualization, AI, rendering, and GPU-accelerated applications. Based on the NVIDIA Ada Lovelace architecture, it combines advanced CUDA, RT, and Tensor Cores with 48 GB of graphics memory.The L40S is suitable for demanding workloads such as AI inference, virtual workstations, 3D rendering, and digital twins. Thanks to its combination of high computing power and large graphics memory, it is specifically designed for high-performance server and data center environments.Package Contents:• NVIDIA L40S 48GB
The NVIDIA L4 24GB is a compact and energy-efficient data center GPU designed for AI inference, generative AI, video processing, and professional graphics workloads. It is suitable for applications such as recommendation systems, visual search, AI assistants, video analysis, and other GPU-accelerated workloads.With 24 GB of graphics memory, the NVIDIA L4 offers sufficient capacity for demanding AI and inference applications. Its compact single-slot design and high energy efficiency enable its use in servers, data centers, and edge systems with limited space and power requirements.Package Contents:• NVIDIA L4 24GB PCIe
Der NVIDIA L40 liefert eine noch nie dagewesene visuelle Rechenleistung für das Rechenzentrum und bietet Grafik-, Rechen- und KI-Funktionen der nächsten Generation für GPU-beschleunigte Anwendungen. Basierend auf der Ada Lovelace GPU-Architektur verfügt der L40 über RT Cores der dritten Generation, die die Echtzeit-Raytracing-Fähigkeiten verbessern, und Tensor Cores der vierten Generation mit Unterstützung für das FP8-Datenformat, die eine Inferenzleistung von über einem Petaflop liefern. Diese neuen Funktionen werden mit CUDA Cores der neuesten Generation und 48 GB Grafikspeicher kombiniert, um Visual Computing-Workloads von hochleistungsfähigen virtuellen Workstation-Instanzen bis hin zu großen digitalen Zwillingen in NVIDIA Omniverse zu beschleunigen. Scope of delivery: • NVIDIA L40
The NVIDIA A10 Tensor Core GPU delivers a versatile platform for mainstream enterprise workloads, like AI inference, training, and HPC. With TF32 and FP64 Tensor Core support, as well as an end-to-end software and hardware solution stack, A10 ensures that mainstream AI training and HPC applications can be rapidly addressed. Multi-instance GPU (MIG) ensures quality of service (QoS) with secure, hardware-partitioned, right-sized GPUs across all of these workloads for diverse users, optimally utilizing GPU compute resources. Scope of delivery: • NVIDIA A10
The NVIDIA A100 Tensor Core GPU delivers unprecedented acceleration at every scale for AI, data analytics, and HPC to tackle the world’s toughest computing challenges. As the engine of the NVIDIA data center platform, A100 can efficiently scale up to thousands of GPUs or, using new Multi-Instance GPU (MIG) technology, can be partitioned into seven isolated GPU instances to accelerate workloads of all sizes. A100’s third-generation Tensor Core technology now accelerates more levels of precision for diverse workloads, speeding time to insight as well as time to market. Scope of delivery: • NVIDIA A100 • 8-pin male (graphics card) to 2x 8-pin female (power supply)
NVIDIA® A40 delivers the data center-based solution designers, engineers, artists, and scientists need to meet today’s challenges. Built on the NVIDIA Ampere architecture, the A40 combines the latest generation RT Cores, Tensor Cores, and CUDA® Cores with 48GB of graphics memory for unprecedented graphics, rendering, compute, and AI performance. From powerful virtual workstations accessible from anywhere, to dedicated render nodes, the A40 is built to tackle the most demanding visual computing workloads from the data center. Scope of delivery: • NVIDIA A40 • 8-pin male (graphics card) to 2x 8-pin female (power supply)
The NVIDIA A2 Tensor Core GPU provides entry-level inference with low power, a small footprint, and high performance for NVIDIA AI at the edge. Featuring a low-profile PCIe Gen4 card and a low 40-60 watt (W) configurable thermal design power (TDP) capability, the A2 brings adaptable inference acceleration to any server. Scope of delivery: • NVIDIA A2 • Low-Profile (SFF) Bracket mounted • Additional Full-Height (ATX) Bracket
GPU name: GH100 Architecture: Hopper Memory size: 80 GB Memory type: HBM2e Memory bus: 5120 bit Memory bandwidth: 1280 GB/s Supported APIs: OpenCL / DirectCompute / OpenACC / CUDA Shader units: 7,296 Tensor cores: 456 FP32 (Float) Power: 24.08 TFLOPS FP64 (double) Power: 12.04 TFLOPS Power consumption: 350 Watt Thermal solution: Passive
The AMD Radeon AI PRO R9700 32GB is a high-performance professional workstation GPU designed for on-premises AI inference, AI development, and demanding visualization workflows. With 32 GB of GDDR6 graphics memory and advanced AMD RDNA 4 architecture, it provides a powerful platform for large language models (LLMs), generative AI, machine learning, and GPU-accelerated applications.Thanks to its support for AMD ROCm™ and PyTorch®, the Radeon AI PRO R9700 is particularly well-suited for modern AI and development environments. At the same time, it delivers high performance for professional applications in areas such as CAD, 3D visualization, rendering, and content creation. Its large graphics memory and multi-GPU support also make it an attractive option for powerful AI and high-performance workstations.Package Contents:• AMD Radeon AI PRO R9700 32GB PCIe 5.0 x16
The AMD Radeon AI PRO R9700 32GB is a professional workstation graphics card designed for on-premises AI inference, AI development, visualisation and memory-intensive workloads. With 32 GB of GDDR6 graphics memory and cutting-edge AMD RDNA 4 architecture, it offers a powerful platform for developers, creators and businesses.The large amount of graphics memory is particularly well-suited to large language models (LLMs), generative AI, text-to-image applications, machine learning and other GPU-accelerated workloads. Thanks to support for AMD ROCm™ and PyTorch®, the R9700 can be integrated into modern AI and development environments.In addition to AI applications, the Radeon AI PRO R9700 is suitable for professional applications such as CAD, 3D visualisation, rendering and content creation. The combination of 32 GB VRAM, high computing power and multi-GPU support makes it a versatile GPU for high-performance workstations.Package contents:AMD Radeon AI PRO R9700 32GB PCIe 5.0 x16
More AI accelerators
The AMD Radeon AI PRO R9700 32GB is a high-performance professional workstation GPU designed for on-premises AI inference, AI development, and demanding visualization workflows. With 32 GB of GDDR6 graphics memory and advanced AMD RDNA 4 architecture, it provides a powerful platform for large language models (LLMs), generative AI, machine learning, and GPU-accelerated applications.Thanks to its support for AMD ROCm™ and PyTorch®, the Radeon AI PRO R9700 is particularly well-suited for modern AI and development environments. At the same time, it delivers high performance for professional applications in areas such as CAD, 3D visualization, rendering, and content creation. Its large graphics memory and multi-GPU support also make it an attractive option for powerful AI and high-performance workstations.Package Contents:• AMD Radeon AI PRO R9700 32GB PCIe 5.0 x16
The AMD Radeon AI PRO R9700 32GB is a professional workstation graphics card designed for on-premises AI inference, AI development, visualisation and memory-intensive workloads. With 32 GB of GDDR6 graphics memory and cutting-edge AMD RDNA 4 architecture, it offers a powerful platform for developers, creators and businesses.The large amount of graphics memory is particularly well-suited to large language models (LLMs), generative AI, text-to-image applications, machine learning and other GPU-accelerated workloads. Thanks to support for AMD ROCm™ and PyTorch®, the R9700 can be integrated into modern AI and development environments.In addition to AI applications, the Radeon AI PRO R9700 is suitable for professional applications such as CAD, 3D visualisation, rendering and content creation. The combination of 32 GB VRAM, high computing power and multi-GPU support makes it a versatile GPU for high-performance workstations.Package contents:AMD Radeon AI PRO R9700 32GB PCIe 5.0 x16
The NVIDIA A10 Tensor Core GPU delivers a versatile platform for mainstream enterprise workloads, like AI inference, training, and HPC. With TF32 and FP64 Tensor Core support, as well as an end-to-end software and hardware solution stack, A10 ensures that mainstream AI training and HPC applications can be rapidly addressed. Multi-instance GPU (MIG) ensures quality of service (QoS) with secure, hardware-partitioned, right-sized GPUs across all of these workloads for diverse users, optimally utilizing GPU compute resources. Scope of delivery: • NVIDIA A10
The NVIDIA A100 Tensor Core GPU delivers unprecedented acceleration at every scale for AI, data analytics, and HPC to tackle the world’s toughest computing challenges. As the engine of the NVIDIA data center platform, A100 can efficiently scale up to thousands of GPUs or, using new Multi-Instance GPU (MIG) technology, can be partitioned into seven isolated GPU instances to accelerate workloads of all sizes. A100’s third-generation Tensor Core technology now accelerates more levels of precision for diverse workloads, speeding time to insight as well as time to market. Scope of delivery: • NVIDIA A100 • 8-pin male (graphics card) to 2x 8-pin female (power supply)
The NVIDIA A2 Tensor Core GPU provides entry-level inference with low power, a small footprint, and high performance for NVIDIA AI at the edge. Featuring a low-profile PCIe Gen4 card and a low 40-60 watt (W) configurable thermal design power (TDP) capability, the A2 brings adaptable inference acceleration to any server. Scope of delivery: • NVIDIA A2 • Low-Profile (SFF) Bracket mounted • Additional Full-Height (ATX) Bracket
NVIDIA® A40 delivers the data center-based solution designers, engineers, artists, and scientists need to meet today’s challenges. Built on the NVIDIA Ampere architecture, the A40 combines the latest generation RT Cores, Tensor Cores, and CUDA® Cores with 48GB of graphics memory for unprecedented graphics, rendering, compute, and AI performance. From powerful virtual workstations accessible from anywhere, to dedicated render nodes, the A40 is built to tackle the most demanding visual computing workloads from the data center. Scope of delivery: • NVIDIA A40 • 8-pin male (graphics card) to 2x 8-pin female (power supply)
GPU name: GH100 Architecture: Hopper Memory size: 80 GB Memory type: HBM2e Memory bus: 5120 bit Memory bandwidth: 1280 GB/s Supported APIs: OpenCL / DirectCompute / OpenACC / CUDA Shader units: 7,296 Tensor cores: 456 FP32 (Float) Power: 24.08 TFLOPS FP64 (double) Power: 12.04 TFLOPS Power consumption: 350 Watt Thermal solution: Passive
The NVIDIA H100 NVL is a high-performance data center GPU based on the NVIDIA Hopper architecture and was specifically designed for demanding AI, machine learning, and high-performance computing workloads. Its large HBM3 memory with high memory bandwidth enables the efficient processing of large AI models and massive datasets. The H100 NVL is particularly well-suited for large language models (LLMs), generative AI, deep learning, scientific computing, and other GPU-accelerated applications. Compatible GPUs can be interconnected via NVIDIA NVLink for particularly memory- and compute-intensive workloads. The passive cooling is designed for appropriately configured server and data center systems.Package Contents:• NVIDIA H100 NVL 94GB PCIe
The NVIDIA L4 24GB is a compact and energy-efficient data center GPU designed for AI inference, generative AI, video processing, and professional graphics workloads. It is suitable for applications such as recommendation systems, visual search, AI assistants, video analysis, and other GPU-accelerated workloads.With 24 GB of graphics memory, the NVIDIA L4 offers sufficient capacity for demanding AI and inference applications. Its compact single-slot design and high energy efficiency enable its use in servers, data centers, and edge systems with limited space and power requirements.Package Contents:• NVIDIA L4 24GB PCIe
Der NVIDIA L40 liefert eine noch nie dagewesene visuelle Rechenleistung für das Rechenzentrum und bietet Grafik-, Rechen- und KI-Funktionen der nächsten Generation für GPU-beschleunigte Anwendungen. Basierend auf der Ada Lovelace GPU-Architektur verfügt der L40 über RT Cores der dritten Generation, die die Echtzeit-Raytracing-Fähigkeiten verbessern, und Tensor Cores der vierten Generation mit Unterstützung für das FP8-Datenformat, die eine Inferenzleistung von über einem Petaflop liefern. Diese neuen Funktionen werden mit CUDA Cores der neuesten Generation und 48 GB Grafikspeicher kombiniert, um Visual Computing-Workloads von hochleistungsfähigen virtuellen Workstation-Instanzen bis hin zu großen digitalen Zwillingen in NVIDIA Omniverse zu beschleunigen. Scope of delivery: • NVIDIA L40
The NVIDIA L40S is a high-performance data center GPU designed for professional visualization, AI, rendering, and GPU-accelerated applications. Based on the NVIDIA Ada Lovelace architecture, it combines advanced CUDA, RT, and Tensor Cores with 48 GB of graphics memory.The L40S is suitable for demanding workloads such as AI inference, virtual workstations, 3D rendering, and digital twins. Thanks to its combination of high computing power and large graphics memory, it is specifically designed for high-performance server and data center environments.Package Contents:• NVIDIA L40S 48GB