jobs
HPC/GPU Systems Engineer
About Nscale Nscale is the vertically integrated AI cloud engineered for AI. We own and operate the full stack — energy, data centres, GPU superclusters, orchestration, and AI services — delivering high-performance infrastructure to AI-nati
- Company
- Nscale
- Location
- Houston; San Francisco; Seattle
- Status
- Open
- Posted
- 2026-06-04T15:34:15+00:00
Infrastructure Support Engineers maintain Nscale's GPU fleets across GPU nodes, high-performance networks, Linux, and data centers. The role handles tickets, alerts, hardware faults, and customer issues, with GPU node triage, nvidia-smi/DCGM log reading, component swaps, vendor RMA evidence, fabric diagnostics, storage investigations, and DCIM/NetBox records. It requires 3–4+ years in infrastructure support, service desk experience, Linux CLI skills, and hands-on server hardware troubleshooting. It is L2/L3 support with on-call, travel, and a stated path toward Senior.
Original job posting