← Back to list
Job · Senior

Senior Site Reliability Engineer

DevOps / SRE • Senior • Remote • Full-time • European Union EU/EMEA

The Reliability team at Latitude.sh is responsible for the health and resilience of the infrastructure behind its global bare metal cloud. The Senior SRE builds reliable, observable, self-healing systems at scale, at the intersection of software engineering and infrastructure.

Responsibilities

  • ▹Continuously improve the platform's reliability and performance
  • ▹Design, build and maintain tools to automate operational tasks and incident response
  • ▹Implement and improve observability solutions, including monitoring, alerting and tracing
  • ▹Collaborate with engineering and platform teams to design scalable and resilient systems
  • ▹Participate in on-call rotations and lead post-incident reviews with a focus on learning
  • ▹Develop and document processes and runbooks that ensure operational excellence
  • ▹Contribute to SLO/SLI definition and reliability metrics adoption across teams

Requirements

  • ▹Strong verbal and written English communication skills
  • ▹Advanced knowledge of Linux/Unix systems in production environments
  • ▹Experience with Kubernetes and container orchestration
  • ▹Proficiency with infrastructure automation tools (e.g. Terraform, Ansible)
  • ▹Experience with observability stacks (e.g. Prometheus, Grafana, Loki, ELK)
  • ▹Familiarity with scripting and programming languages such as Bash, Python, Go or Ruby
  • ▹Working knowledge of Git and CI/CD pipelines
  • ▹Solid understanding of incident management and root cause analysis
  • ▹Knowledge of cloud-native reliability and security best practices

Soft skills

Learning-focused mindsetCommunicationTeamwork

What we offer

  • ▹Contractor (PJ) agreement
  • ▹Paid time off
  • ▹Competitive compensation
  • ▹Wellhub (formerly Gympass)
  • ▹Annual bonus based on company and team performance
  • ▹Flexible work hours
  • ▹Opportunities for professional growth and development

About the company

Latitude.sh launched its global computing platform in 2019, letting businesses programmatically deploy single-tenant bare metal instances around the world. The team is building the fastest, easiest-to-use, developer-centric single-tenant cloud infrastructure.

Languages: Angol: erős szóbeli és írásbeli

Similar jobs