|
|
|
List of Xi Servers with NVIDIA H100 NVL Tensor Core GPU
|
|
|
|
 |
Unprecedented performance, scalability, and security for every data cente
The NVIDIA H100 NVL Tensor Core GPU is the most optimized platform for LLM
Inferences with its high compute density, high memory bandwidth, high energy
efficiency, and unique NVLink architecture. It also delivers unprecedented acceleration
to power the world’s highest-performing elastic data centers for AI, data analytics, and
high-performance computing (HPC) applications. NVIDIA H100 NVL Tensor Core
technology supports a broad range of math precisions, providing a single accelerator for
every compute workload.
|
|
|
 |
Transformational AI Training
NVIDIA H100 GPUs feature fourth-generation Tensor Cores and the Transformer Engine with FP8 precision that provides up to 9X faster training over the prior generation for mixture-of-experts (MoE) models. The combination of fourth-generation NVlink, which offers 900 gigabytes per second (GB/s) of GPU-to-GPU interconnect; NVLINK Switch System, which accelerates communication by every GPU across nodes; PCIe Gen5; and NVIDIA Magnum IO™ software delivers efficient scalability from small enterprises to massive, unified GPU clusters.
. |
 |
Real-Time Deep Learning Inference
H100 further extends NVIDIA’s market-leading inference leadership with several advancements that accelerate inference by up to 30X and deliver the lowest latency. Fourth-generation Tensor Cores speed up all precisions, including FP64, TF32, FP32, FP16, and INT8, and the Transformer Engine utilizes FP8 and FP16 together to reduce memory usage and increase performance while still maintaining accuracy for large language models.
|
 |
Exascale High-Performance Computing
H100 also features DPX instructions that deliver 7X higher performance over NVIDIA A100 Tensor Core GPUs and 40X speedups over traditional dual-socket CPU-only servers on dynamic programming algorithms, such as Smith-Waterman for DNA sequence alignment. |
To learn more about the NVIDIA H100 Tensor Core GPU, visit www.nvidia.com/h100
© 2022 NVIDIA Corporation. All rights reserved. NVIDIA, the NVIDIA logo, CUDA, DGX, HGX, HGX H100, NVLink, NVSwitch, OpenACC, TensorRT,
and Volta are trademarks and/or registered trademarks of NVIDIA Corporation in the U.S. and other countries. OpenCL is a trademark of Apple
Inc. used under license to the Khronos Group Inc. All other trademarks and copyrights are the property of their respective owners. Jun20
|
|
|
|
DATA CENTER GPUs QUICK SPECS: |
|
Model Name |
NVIDIA H100 NVL for PCIe |
NVIDIA H100 for SXM |
|
FP32 CUDA Cores |
2 x 16896 (paired) |
16,896 |
|
SM Tensor Cores |
2 x 528 (paired) |
528 |
|
FP64 |
38 teraFLOPS |
34 teraFLOPS |
|
FP64 Tensor
Core |
134 teraFLOPS |
67 teraFLOPS |
|
FP32 |
134 teraFLOPS |
67 teraFLOPS |
|
TF32 Tensor
Core |
1,979 teraFLOPS* |
939 teraFLOPS* |
|
BFLOAT16 Tensor
Core |
3,958 teraFLOPS* |
1,979 teraFLOPS* |
|
FP16 Tensor
Core |
3,958 teraFLOPS* |
1,970 teraFLOPS* |
|
FP8 Tensor Core
|
7.916 teraFLOPS* |
3,958 teraFLOPS* |
|
INT8 Tensor
Core |
7.916 TOPS* |
3,958 TOPS* |
|
GPU Memory |
2 x 94 GB (188GB) HBM3 |
80 GB HBM2e |
|
GPU Memory Bandwidth |
7.8TB/s**** |
3.35TB/s |
|
Interconnect |
NVIDIA NVLink: 600GB/s PCIe Gen5: 128GB/s** |
NVIDIA NVLink 900 GB/s PCIe Gen5 128 GB/s |
|
Multi-instance GPUs |
Various instance
Up to 14 MIGS @ 12GB each |
Various instance
Up to 7 MIGS @ 10GB each |
|
Form Factor |
2x PCIe dual-slot |
SXM |
|
Max TDP Power |
2x 350-400W |
700W |
|
Networking |
- |
- |
|
Server options |
Partner and
NVIDIA-Certified Systems
with 2-4 pairs (minimun 2) |
NVIDIA HGX H100 Partner and NVIDIA-Certified Systems with 4 or 8 GPUs NVIDIA DGX H100 with 8 GPUs |
|
NVIDIA AI Enterprise |
Supported by VMWare |
Add-on |
|
* With sparsityIe GPUs via NVLink Bridge for up to 2-GPUs
*** Up to 400 Gb/s (NDR or 400GbE), dual-port QSFP112 (with aggregated bandwidth of 400 GB/s), Ethernet or InfiniBand
**** Aggregate HBM bandwidth
Hopper family
NVIDIA H100 NVL for PCIe -
NVIDIA H100 for SXM |
|
Best Optimized Data Center Servers |
Xi® NetRAIDer™ 64LT
Xi® NetRAIDer™ 64XLT
Xi® NetRAIDer™ 64XE
Xi® WebRAIDer™ 64X-1U |
|
|
*Published Prices. Prices shown (provided by way of a Quotation or a Pre-Configured List) are subject to change without prior notice to the prices in effect at the time of Quotation or placing an Order. Seller reserves the right to make any corrections to prices quoted due to availability, clerical errors or errors of omission. In the event of any specific requirements (including without limitation any design, specification, ordered quantity, or shipment changes) representing a price increase.
|
|
|
|
Follow us on @Xi Social Media:
|
|
|
|
|
.:Total views:. 8636
|
Typographic errors are subject to correction. Merchandise enlarged to show detail and may not always be exactly as pictured.
All trademarks and brands mentioned on this website may be legally registered in
the U.S. and/or other countries. They are subject without restriction to the terms of applicable registered trademark rights and the ownership rights of the respective registered owners. The mention of a trademark should not be taken to indicate that such a trademark is not subject to third-party rights.
If you wish to view this webpage properly, please use Firefox, Google Chrome or Microsoft Edge.
|
|
Home
| Products
|
Support
|
Company
|
Contact |
Terms & Privacy
© 1996- @Xi® Computer Corporation | All Rights reserved. |
|
|