← Zurück zur Liste

Stelle
· Principal
Staff SRE Engineer
DevOps / SRE
• Principal
• Remote
• Vollzeit
•
Ungarn
Stellar Cyber's Staff SRE Engineer drives reliability, scalability, and efficiency across production systems, bringing deep expertise in cloud infrastructure, Kubernetes, and observability.
Stack
Responsibilities
- ▹Administer and maintain container orchestration platforms and containerized workloads
- ▹Monitor and troubleshoot production systems, participating in on-call rotations
- ▹Drive observability improvements across monitoring, logging, and alerting
- ▹Administer and optimize cloud-based environments across multiple providers
- ▹Manage and support distributed data platforms and real-time processing systems
- ▹Develop and maintain CI/CD pipelines for reliable deployments
- ▹Own and implement Infrastructure as Code (IaC) practices
- ▹Automate and orchestrate infrastructure using programming and scripting languages
- ▹Perform system administration and networking tasks for internal and external environments
- ▹Collaborate with engineers and stakeholders across time zones
Requirements
- ▹5+ years of experience in Site Reliability Engineering, DevOps, or Platform Engineering
- ▹Proven success leading large-scale production systems in cloud environments (AWS, GCP, Azure, or OCI)
- ▹Demonstrated leadership in incident response, on-call best practices, and reliability-focused culture
- ▹Strong experience with production on-call operations and incident management
- ▹Advanced proficiency in Kubernetes administration and troubleshooting
- ▹Hands-on experience with observability tools: Prometheus, Grafana, Loki, and Alertmanager
- ▹Familiarity with chat-based operations interfaces and/or AI-agent-based auto-remediation controllers
- ▹Understanding of AI agents for auto-triaging alerts and correlating signals
- ▹Expertise operating data platforms (Elasticsearch, MongoDB, Spark, Kafka, Redis)
- ▹Proficiency with public cloud services (AWS, Azure, GCP, or OCI)
About the company
Stellar Cyber is a fast-growing global cybersecurity company trusted by nearly 30% of the world's top MSSPs, protecting organizations against sophisticated cyber threats using AI and automation technologies.
Ähnliche Stellen

Stelle
· Principal
Senior / Staff Site Reliability Engineer
RADAR
+6
172 914–259 370 €/Jahr
brutto
🏢 Vor Ort
New York
🗣️ EN

Stelle
· Principal
Principal Site Reliability Developer (SRE/SRD)
Oracle
+4
💰 Gehalt: keine Angabe
🏢 Vor Ort
🗣️ EN

Stelle
· Principal
Staff DevOps Engineer (Platform)
Phantom
Github Actions
+7
💰 Gehalt: keine Angabe
🌍 Remote
🗣️ EN

Stelle
· Principal
Staff DevOps Engineer
Anduril
Datadog
+5
165 133–218 736 €/Jahr
brutto
🏢 Vor Ort
Costa Mesa
🗣️ EN

Stelle
· Principal
Staff Site Reliability Engineer
Skydio
ArgocdDatadogGithub Actions
+8
💰 Gehalt: keine Angabe
🔀 Hybrid
San Mateo
🗣️ EN

Stelle
· Principal
Principal DevOps Engineer (Poland Remote)
Turnitin, LLC
Dynamodb
+4
71 377 €/Jahr
brutto
🌍 Remote
Krakow
🗣️ EN