Open role
Principal Deployment Engineer
nscaleoperationsukltd
Houston; New York; San Francisco; Seattle; USPosted Jul 21, 2026 · 7h ago$175k – $225k
Full-time$175k – $225kPrincipalHybridAI / ML
About this role
Lead the architecture and deployment of large-scale GPU superclusters for AI workloads. This technical leadership role involves defining standards for deployment, validation, and scaling of high-performance computing infrastructure. You will own the full lifecycle from rack design to production readiness, ensuring peak performance and reliability.
What we are looking for
6- Architect and lead large-scale GPU cluster bringup
- Define technical standards for node, rack, and cluster deployment
- Architect high-performance network fabrics (IB, RoCE, Ethernet)
- Establish cluster acceptance criteria and validation frameworks
- Debug performance issues across compute and network layers
- Drive automation for provisioning and cluster validation
