Launch your career with Hamilton Barnes’ Graduate Hub
Take me there

Senior Solution Engineer (GPU) - Systems Integrator

1731616
  • Up to £150,000 Per Annum
  • United Kingdom
  • Permanent
  • 150000
  • Artificial Intelligence


Ready to take the next step in your career?

Join a high-performance neocloud provider purpose-built for AI, HPC, and cloud-native infrastructure. The organisation delivers bare-metal and Kubernetes-based NVIDIA GPU solutions for large-scale AI and HPC workloads, offering an alternative to traditional hyperscalers with a focus on performance and predictable pricing.

This company is looking for a Senior Solution Engineer to act as the primary technical architect for major AI and HPC customer engagements. The role involves designing NVIDIA Blackwell GPU clusters for foundation model training and inference, owning solutions from initial design through to Bill of Materials, and working directly with leading AI labs. The position is fully remote, with a four-day working week, uncapped holiday, and flexible hours focused on output.

Ready to take the next step in your career?


Responsibilities:

  • Author comprehensive High-Level and Low-Level Design documentation for enterprise-scale GPU supercomputing clusters
  • Produce detailed Bills of Materials covering compute nodes, NVLink switches, network fabrics, cabling, cooling, power distribution and high-performance storage
  • Architect scale-up and scale-out topologies (NVLink/NVSwitch, Fat-Tree, Rail-Optimised) for NVIDIA Blackwell B300 and GB300NVL platforms
  • Design high-throughput, low-latency fabrics using InfiniBand and RoCE/RoCEv2, including lossless Ethernet mechanisms such as PFC, ECN and Adaptive Routing
  • Act as technical lead alongside sales and commercial teams on high-value AI infrastructure opportunities, engaging directly with customer CTOs, Chief AI Officers and ML engineering leads
  • Architect and oversee proof-of-concept deployments, benchmarking with tools such as NCCL tests, GPUDirect RDMA and MLPerf to validate real-world workload performance


Skills/Must Have:

  • Experience: 5+ years in Solution Architecture, Systems Engineering or Technical Pre-Sales, focused on high-performance cloud, HPC or AI infrastructure
  • Core Tech/Domain: Deep hands-on knowledge of NVIDIA HGX/DGX platforms, NVLink/NVSwitch fabrics and Blackwell architectures (B300, GB300NVL, GB200 NVL72/NVL36)
  • Methodology/Protocols: Expert-level InfiniBand (Quantum-2/Quantum-X800, Adaptive Routing) and RoCE/RoCEv2 (Spectrum-X/Spectrum-4), plus Kubernetes orchestration (NVIDIA GPU Operator, MPI Operator) and bare-metal tooling (Slurm, Ansible, Terraform)
  • Soft Skills: Strong technical leadership and presentation skills, able to translate complex hardware and network trade-offs for executive stakeholders; strong spoken and written English is essential


Desirable Skills:

  • NVIDIA Certified Professional: AI Infrastructure (NCP-AII)
  • NVIDIA Certified Professional: AI Networking (NCP-AIN)
  • NVIDIA Certified Professional: InfiniBand (NCP-IB)
  • NVIDIA Certified Associate/Professional: AI Workload Deployment & Cloud Native
  • Familiarity with high-density data centre environments, liquid cooling and power delivery for 100kW+ racks


Benefits:

  • Four-day working week as standard (five-day only when attending events)
  • Uncapped holiday
  • Fully remote, flexible working with no fixed hours
  • Direct exposure to some of the largest GPU cluster deployments for major global AI labs
  • Collaborative, inclusive culture with genuine autonomy


Salary:

  • Up to £150,000 Per Annum
Jamie Maher Head of IP & AI Infrastructure

Apply for this role