← Back to list
Job · Senior

Senior Site Reliability Engineer

DevOps / SRE • Senior • Remote • Full-time • Czechia Czechia

Senior SRE at Nebius, ensuring fault-tolerance, scale and uninterrupted operation of cloud infrastructure services. You will solve infrastructure problems with cutting-edge cloud technology and improve CI/CD processes.

Responsibilities

  • ▹Ensure fault-tolerance, scale and uninterrupted operations for the service
  • ▹Use cutting-edge cloud technology to solve a variety of infrastructure problems
  • ▹Implement and improve CI/CD processes

Requirements

  • ▹Solid experience with programming languages (like Go, Python or C++)
  • ▹Solid understanding of classic algorithms and data structures
  • ▹Commercial experience with and deep understanding of Unix systems and network technology
  • ▹Experience with systems for containerization and configuration management (Ansible, Salt, Terraform, Docker, K8s, Helm)

Nice to have

  • ▹Desire to be involved in backend development
  • ▹Experience designing, developing and running high-load distributed systems
  • ▹Commercial experience with a variety of cloud platforms

What we offer

  • ▹Competitive compensation
  • ▹Career growth and learning opportunities
  • ▹Flexibility and ownership
  • ▹Collaborative and innovative culture
  • ▹Opportunity to work on impactful AI projects
  • ▹International environment and talented teams

About the company

Nebius is building a full-stack AI cloud platform from data and model training through to production deployment. Listed on Nasdaq and headquartered in Amsterdam, it has R&D hubs across Europe, the UK, North America and Israel and a team of 1,500+.

Similar jobs