Skip to content
AI.info

jobs

Site Reliability Engineer in Network Infrastructure

About Nebius: Nebius is leading a new era in cloud infrastructure for the global AI economy. We are building a full-stack AI cloud platform that supports developers and enterprises from data and model training through to production deployme

Company
Nebius
Location
Amsterdam, Netherlands; Remote - Europe
Status
Open
Posted
2026-06-17T16:40:56+00:00

Nebius is hiring a site reliability engineer for the network layer of its AI cloud platform. The person will set reliability targets (SLIs, SLOs, error budgets) for network services and inter-site connectivity, run incident response and postmortems, build observability and alerting, and automate change workflows with CI/CD, canarying and rollbacks. The posting asks for Linux and networking fundamentals, experience operating high-availability systems, and Go or Python automation plus infrastructure-as-code and container tooling. Datapath work such as load balancing, NAT64, eBPF/XDP or DPDK is listed as a bonus.

Original job posting