← Zurück zur Liste
Stelle · Mid-level

Site Reliability Engineer

DevOps / SRE • Mid-level • Remote • Vollzeit Spanien Spanien

Tinybird's Platform team is looking for an experienced Site Reliability Engineer to keep large-scale distributed systems reliable as they grow.

Responsibilities

  • Design, build and operate distributed cloud architectures and large-scale production systems
  • Operate production-grade Kubernetes clusters, write custom controllers or operators
  • Tune autoscaling mechanisms (KEDA, Karpenter)
  • Debug incidents and improve observability and service reliability

Requirements

  • Strong experience designing and running distributed cloud architectures and large-scale web-based production systems
  • Deep knowledge of Kubernetes
  • Skilled in AWS and GCP
  • Coding skills, primarily Python and some C++
  • SQL experience, comfortable working with real-time analytical systems

Nice to have

  • Experience with ClickHouse or launching database systems at scale
  • Familiarity with Traefik, Varnish, Redis, Terraform or Ansible

Soft skills

Systems thinking, attention to failure modesBias toward action and fast iterationOwnership, willingness to fix broken things

About the company

Tinybird helps developers and data teams unlock real-time data, publishing low-latency, high-concurrency APIs.

Ähnliche Stellen