High-Performance Computing & Research
Accelerated computing for science and engineering: the HPC SDK compilers and libraries, CUDA-X, CUDA-Q for quantum research, PhysicsNeMo for physics AI, and Grace CPUs with HGX and DGX systems.
Technology profiles for this category are in research.
Overview
Research computing centers run long simulations in climate, chemistry, materials and engineering, often in Fortran and C++ code built up over decades. Porting and scaling that code is the main task in this category.
The HPC SDK provides the NVC, NVC++ and NVFORTRAN compilers with OpenACC, CUDA Fortran and standard C++ parallel algorithms, plus cuBLAS, cuFFT and other math libraries, a CUDA-aware Open MPI, NCCL, NVSHMEM and the Nsight profilers. CUDA-Q is an open-source platform for hybrid quantum and classical programs that can simulate circuits on GPUs or run on quantum processors. For hardware, NVIDIA lists HGX and DGX platforms and Grace CPUs for HPC; the Grace CPU Superchip joins two CPUs over NVLink-C2C.
Codes dominated by serial logic gain little from GPUs. Profile before porting, and check the GPU Apps Catalog in case a GPU-ready version of your application already exists.1234
Problems it addresses
Porting legacy Fortran1
NVFORTRAN supports OpenACC directives and CUDA Fortran, so code can move to GPUs step by step.
Scaling across nodes1
A CUDA-aware Open MPI, NCCL and NVSHMEM handle communication between GPUs and nodes.
Quantum research without hardware2
CUDA-Q GPU simulators let researchers develop and benchmark circuits without a quantum processor.
Costly repeated simulations5
PhysicsNeMo trains surrogate models that approximate expensive solvers.
A typical workflow
Profile the CPU code6
Find the loops that dominate run time before porting anything.
Port step by step1
Add OpenACC directives or standard parallel algorithms and compile with NVFORTRAN or NVC++.
Call tuned libraries1
Replace hand-written math with cuBLAS, cuFFT, cuSOLVER or cuSPARSE.
Scale out1
Use CUDA-aware MPI and NCCL across nodes.
Profile again1
Check GPU use with Nsight Systems and Nsight Compute, both included in the HPC SDK.
AI Factory Efficiency Lab
Model token and infrastructure costs for your own numbers.
Next steps
Check the GPU Apps Catalog for an existing GPU version of your application.
Profile one representative run to find the loops worth porting.
Estimate GPU count and power for a planned simulation campaign in the AI Factory Efficiency Lab.
Sources
Thank you. Your correction was sent.
The editors check it against the sources. If you left an email address, they may reply about it.