IC Resources is looking to speak with specialist GPU Infrastructure Architects and Engineers with experience designing and bringing up large-scale NVIDIA compute environments.
The requirement sits at the infrastructure layer of large-scale AI and HPC platforms, covering the integration of GPU compute, high-speed networking, storage and bare-metal infrastructure.
We’re particularly interested in engineers who have been directly involved in the architecture, deployment or commissioning of multi-node NVIDIA GPU clusters rather than solely operating workloads once the platform is live.
Key experience:
-
Large-scale NVIDIA GPU infrastructure
-
NVIDIA HGX and/or DGX platforms
-
Hopper and/or Blackwell architectures
-
Multi-node GPU cluster design and deployment
-
InfiniBand / RDMA
-
High-speed Ethernet / AI fabrics
-
NVIDIA networking and interconnect technologies
-
Bare-metal compute infrastructure
-
GPU cluster topology
-
Compute, network and storage integration
-
Cluster commissioning and validation
-
Performance testing and benchmarking
-
Troubleshooting across hardware, networking and infrastructure layers
Additional experience with technologies such as NVLink, NVSwitch, ConnectX, BlueField, Spectrum-X, NCCL or GPUDirect would be particularly relevant.
This is a fully remote opportunity and we’re open to speaking with specialist engineers internationally.
If this opportunity could be of interest, please contact Bradley Wilson at IC Resources for more information.