← Back to list

Job
Site Reliability Engineer
DevOps / SRE
• On-site
• Full-time
•
Budapest, Hungary
Site Reliability Engineer role operating a Kubernetes-based platform, supporting SLOs and reliability, and handling on-call incident response within a modern cloud native technology stack.
Responsibilities
- ▹Operate a Kubernetes-based platform
- ▹Day-to-day management of clusters and nodes (upgrades, patching, node cordon/drain, scaling)
- ▹Operate persistent storage (e.g. CSI/Longhorn-type solutions), basic capacity planning
- ▹Participate in defining and monitoring SLIs/SLOs
- ▹Track error budgets, feed back incidents and trends to the team
- ▹Use and perform basic configuration of monitoring and log-collection systems (dashboards, alerts)
- ▹Participate in the on-call rotation: first-line handling of alerts, resolving incidents based on runbooks
- ▹Support automation of deployment and configuration (CI/CD, Git-based workflows)
- ▹Create and maintain runbooks and operational documentation
- ▹Collaborate daily with development, operations and business stakeholders
- ▹Propose improvements to processes and platform reliability
Requirements
- ▹At least 2–3 years of experience operating Linux-based systems
- ▹Hands-on experience with containerized environments (Docker) and Kubernetes, ideally in production
- ▹Experience with a monitoring/logging stack (e.g. Prometheus/Grafana, ELK, Zabbix, etc.)
- ▹Basic experience with CI/CD systems and Git-based workflows (Gitea, ArgoCD)
- ▹Operational-level knowledge of a Cloud Native stack (Linux, Ubuntu, containerd, Docker, Kubernetes, Prometheus/Grafana, ELK, Zabbix, RabbitMQ, MinIO, PostgreSQL)
- ▹Willingness to participate in on-call rotation with structured troubleshooting thinking
- ▹Good communication skills, collaboration across multiple teams and organizational units
Soft skills
Structured troubleshooting mindsetGood communication skillsCross-team and cross-department collaboration
What we offer
- ▹Varied projects based on modern technologies
- ▹Innovative company with a stable background
- ▹5 extra paid days off per year (for non-contractor status)
- ▹Participation in professional events, workshops and hackathons
- ▹Flexible working hours and a friendly atmosphere
- ▹Team-building programs and shared leisure activities
- ▹Real impact on products and customers
Similar jobs

Job
Sovereign Engineering Platform SRE - T Cloud Public (REF5740Q)
Deutsche Telekom IT Solutions
Argocd
+8
💰 Salary: not specified
🏢 On-site
Budapest

Job
DevOps Engineer (REF5727Q)
Deutsche Telekom IT Solutions
ArgocdArtifactory
+14
💰 Salary: not specified
🏢 On-site
Budapest

Job
DevOps Engineer
Zenitech
Azure Devops
+8
💰 Salary: not specified
🏢 On-site
Budapest

Job
DevOps Engineer (German speaking)
Deutsche Telekom IT Solutions
+5
💰 Salary: not specified
🏢 On-site
Budapest

Job
Devops Engineer (REF5687R)
Deutsche Telekom IT Solutions
ArgocdArtifactory
+14
💰 Salary: not specified
🏢 On-site
Budapest

Job
Site Reliability Engineer
Betsson
Couchbase
+15
💰 Salary: not specified
🏢 On-site
Budapest