jobs
Senior AI Infrastructure Systems Engineer (R&D / GPU / AI)
About Nebius: Nebius is leading a new era in cloud infrastructure for the global AI economy. We are building a full-stack AI cloud platform that supports developers and enterprises from data and model training through to production deployme
In inglese
- Company
- Nebius
- Location
- United States
- Status
- Open
- Posted
- 2026-01-28T21:39:47+00:00
A senior systems engineering role at Nebius, focused on finding the root cause of hardware and platform failures in large GPU-based AI infrastructure. The engineer designs tests, reads system logs and low-level hardware data, and separates GPU, host, firmware, interconnect, thermal and power faults — often where no troubleshooting procedure exists. The posting asks for five or more years across Linux, server hardware, firmware, PCIe and NVIDIA GPU platforms, kernel-level debugging, and scripting in Python or Go. Work is onsite in the United States.
Original job posting