Overview
The Network Architect for AI GPU Data Centres will play a pivotal role in transforming an existing data center into a next-generation AI-ready facility. Collaborating with specialists across various disciplines, this position focuses on leading the architectural design and deployment strategies for large-scale NVIDIA GPU infrastructure, ensuring alignment with AI compute requirements and data center capabilities.
Responsibilities
- Define and validate architecture for large-scale NVIDIA GPU deployments.
- Develop AI cluster design standards supporting training and inference workloads.
- Work with data center design teams to align compute requirements with facility capabilities.
- Define rack density, power, and cooling strategies for high-density AI environments.
- Provide technical leadership around NVIDIA DGX, HGX, Blackwell and future GPU platforms.
- Design GPU networking and fabric requirements including NVLink, NVSwitch and InfiniBand.
- Support capacity planning and scaling strategies for a 20-30MW AI compute estate.
- Collaborate with infrastructure, cloud, networking and operations teams.
Requirements
- Proven experience designing large-scale GPU infrastructure environments.
- Experience with NVIDIA AI platforms including DGX and/or HGX systems.
- Strong understanding of AI training and inference architectures.
- Deep knowledge of GPU networking technologies such as NVLink, NVSwitch, InfiniBand, and high-speed Ethernet.
- Strong understanding of data center power and cooling constraints.
- Experience supporting high-density rack deployments.
- Ability to communicate with both technical and executive stakeholders.
- Experience with liquid cooling deployments and cloud-scale infrastructure.