Nebius is building a full-stack AI cloud platform for developers and enterprises, supporting data and model training through production deployment. The Site Reliability Engineer will support the Hardware Infrastructure team by ensuring reliable, scalable, and uninterrupted services, solving infrastructure problems, and improving CI/CD processes.