← Back to list
Job · Senior

Site Reliability Engineer (SRE)

DevOps / SRE • Senior • Remote • Full-time European Union EU/EMEA

Site Reliability Engineer role at ArangoDB, responsible for the reliability, scalability, and performance of distributed database systems running on Kubernetes and cloud environments (AWS, Google Cloud).

Responsibilities

  • Design, implement, and maintain cloud infrastructure on AWS and Google Cloud
  • Ensure scalability, performance, and reliability of Kubernetes-based distributed database systems
  • Collaborate with developers to write efficient, production-grade Golang code to automate infrastructure management
  • Optimize and automate CI/CD pipelines, deployment processes, and monitoring systems
  • Develop strategies for disaster recovery, high availability, and fault tolerance

Requirements

  • Experience operating Kubernetes-based systems
  • Cloud platform knowledge: AWS and Google Cloud
  • Golang programming skills, or willingness to learn
  • Experience with CI/CD pipelines
  • Ability to troubleshoot complex system issues

Soft skills

Problem-solving mindsetCollaboration with development teamsProactive approach to system reliability

About the company

ArangoDB builds a unified, multimodel contextual data platform that powers AI agents, assistants, and applications. Customers include NVIDIA, HPE, and the London Stock Exchange. ArangoDB is a member of the NVIDIA Inception Program and AWS ISV Accelerate Program.

Similar jobs