← Back to list
Job · Senior

Senior Site Reliability Engineer (SRE) - Kubernetes

DevOps / SRE • Senior • Remote • Full-time • Poland Poland

Owns production reliability for the platform powering ServiceNow's AI-first user interfaces (a Node.js SSR runtime and a Java/Glide layer) on Kubernetes. The role covers observability, incident response, and troubleshooting both the Node.js and JVM sides of the system.

Responsibilities

  • ▹Operate and ensure reliability of Kubernetes-based production services
  • ▹Monitor service health and investigate incidents
  • ▹Participate in on-call, root cause analysis, postmortems
  • ▹Troubleshoot networking and service-to-service issues
  • ▹Support CI/CD and GitOps-based deployments

Requirements

  • ▹5+ years of experience in SRE, DevOps, or Platform Engineering
  • ▹3+ years hands-on production Kubernetes experience
  • ▹Splunk experience for log aggregation and troubleshooting
  • ▹Prometheus and Grafana: building alert rules and dashboards
  • ▹CI/CD and infrastructure-as-code (Helm, ArgoCD/Flux)
  • ▹Strong Linux and networking fundamentals (DNS, HTTP/2)
  • ▹Production troubleshooting for Node.js and JVM/Java
  • ▹Service-to-service authentication (mTLS, certificate rotation, JWT)
  • ▹Good spoken and written English

Nice to have

  • ▹Web Components / Lit experience
  • ▹Server-side rendering experience
  • ▹Canary rollout and multi-version production operations
  • ▹Distributed tracing
  • ▹KEDA or event-driven autoscaling

What we offer

  • ▹Flexible employment and remote work
  • ▹International projects with leading global clients
  • ▹International business trips
  • ▹Language classes, internal and external training
  • ▹Private healthcare and insurance
  • ▹Multisport card

About the company

Software Mind is an international software engineering company building technology solutions for global clients; this project serves ServiceNow's AI user-interface platform.

Similar jobs