jobs
Site Reliability Engineer in Network Infrastructure
About Nebius: Nebius is leading a new era in cloud infrastructure for the global AI economy. We are building a full-stack AI cloud platform that supports developers and enterprises from data and model training through to production deployme
In inglese
- Company
- Nebius
- Location
- Amsterdam, Netherlands; Remote - Europe
- Status
- Open
- Posted
- 2026-06-17T16:40:56+00:00
Nebius is hiring a site reliability engineer for the network layer of its AI cloud platform. The person will set reliability targets (SLIs, SLOs, error budgets) for network services and inter-site connectivity, run incident response and postmortems, build observability and alerting, and automate change workflows with CI/CD, canarying and rollbacks. The posting asks for Linux and networking fundamentals, experience operating high-availability systems, and Go or Python automation plus infrastructure-as-code and container tooling. Datapath work such as load balancing, NAT64, eBPF/XDP or DPDK is listed as a bonus.
Original job posting