← Back to list

Job
· Principal
Staff SRE Engineer
DevOps / SRE
• Principal
• Remote
• Full-time
•
Hungary
Stellar Cyber's Staff SRE Engineer drives reliability, scalability, and efficiency across production systems, bringing deep expertise in cloud infrastructure, Kubernetes, and observability.
Stack
Responsibilities
- ▹Administer and maintain container orchestration platforms and containerized workloads
- ▹Monitor and troubleshoot production systems, participating in on-call rotations
- ▹Drive observability improvements across monitoring, logging, and alerting
- ▹Administer and optimize cloud-based environments across multiple providers
- ▹Manage and support distributed data platforms and real-time processing systems
- ▹Develop and maintain CI/CD pipelines for reliable deployments
- ▹Own and implement Infrastructure as Code (IaC) practices
- ▹Automate and orchestrate infrastructure using programming and scripting languages
- ▹Perform system administration and networking tasks for internal and external environments
- ▹Collaborate with engineers and stakeholders across time zones
Requirements
- ▹5+ years of experience in Site Reliability Engineering, DevOps, or Platform Engineering
- ▹Proven success leading large-scale production systems in cloud environments (AWS, GCP, Azure, or OCI)
- ▹Demonstrated leadership in incident response, on-call best practices, and reliability-focused culture
- ▹Strong experience with production on-call operations and incident management
- ▹Advanced proficiency in Kubernetes administration and troubleshooting
- ▹Hands-on experience with observability tools: Prometheus, Grafana, Loki, and Alertmanager
- ▹Familiarity with chat-based operations interfaces and/or AI-agent-based auto-remediation controllers
- ▹Understanding of AI agents for auto-triaging alerts and correlating signals
- ▹Expertise operating data platforms (Elasticsearch, MongoDB, Spark, Kafka, Redis)
- ▹Proficiency with public cloud services (AWS, Azure, GCP, or OCI)
About the company
Stellar Cyber is a fast-growing global cybersecurity company trusted by nearly 30% of the world's top MSSPs, protecting organizations against sophisticated cyber threats using AI and automation technologies.
Similar jobs

Job
· Principal
Senior / Staff Site Reliability Engineer
RADAR
+6
$200,000–$300,000/yr
gross
🏢 On-site
New York
🗣️ EN

Job
· Principal
Principal Site Reliability Developer (SRE/SRD)
Oracle
+4
💰 Salary: not specified
🏢 On-site
🗣️ EN

Job
· Principal
Staff DevOps Engineer (Platform)
Phantom
Github Actions
+7
💰 Salary: not specified
🌍 Remote
🗣️ EN

Job
· Principal
Staff DevOps Engineer
Anduril
Datadog
+5
$191,000–$253,000/yr
gross
🏢 On-site
Costa Mesa
🗣️ EN

Job
· Principal
Staff Site Reliability Engineer
Skydio
ArgocdDatadogGithub Actions
+8
💰 Salary: not specified
🔀 Hybrid
San Mateo
🗣️ EN

Job
· Principal
Principal DevOps Engineer (Poland Remote)
Turnitin, LLC
Dynamodb
+4
308,025 zł/yr
gross
🌍 Remote
Krakow
🗣️ EN