Live opening · Posted 6 days ago
At a glance
The key details from the original listing.
Your early-applicant advantage
Live timing from JobBeeper.
About the role
Description supplied by the original job listing.
You will be the most senior technology leader at 1Legion, reporting directly to the CEO, with end-to-end ownership of how we design, build and operate GPU infrastructure across five countries. You will turn signed contracts into delivered, SLA-backed capacity on time and on budget, and you will be the person our customers, lenders and investors look to when they ask whether the platform is engineered by someone who has done this before.
What you will do
Own the reference architecture for every deployment: rack, pod and cluster design for NVIDIA HGX B300 and rack-scale NVL72-class systems, including liquid cooling at 100 kW and above per rack
Lead delivery of new capacity from site to acceptance: cluster build, fabric, storage, provisioning, burn-in and customer sign-off against defined benchmarks
Run the bare metal platform: provisioning and re-imaging, tenant isolation without a hypervisor, InfiniBand and Ethernet fabrics, parallel storage, observability and SLA reporting
Own 24/7 operations across all sites, including facilities, NOC and incident management
Act as design authority on our data centre builds and conversions, from M&E design through commissioning and handover
Represent the platform in technical diligence with customers, lenders and investors
Hire and lead engineering and site operations teams in each country we operate in
What we are looking for
10+ years in infrastructure engineering, with at least 3 years in senior leadership at a GPU cloud, a hyperscaler GPU business, or a data centre operator hosting GPU workloads
You have built or converted at least one data centre and operated live facilities with paying tenants, not only deployed servers into someone else's building
You have delivered multi-thousand-GPU clusters into production for external customers
Deep bare metal fundamentals: HGX and NVL systems, BMC and firmware lifecycle at fleet scale, InfiniBand and RoCE fabric design, parallel file systems, Kubernetes and Slurm on bare metal
You have run infrastructure in more than one country and hired across borders
You are comfortable committing to delivery dates that carry contractual consequences, and building the buffers to hit them
Nice to have
Liquid-cooled Blackwell deployments specifically
Experience on the issuer side of a lender or investor technical diligence process
Work arrangement
Yes
More openings worth a look
Recently tracked roles with full details and direct application links.