← Back to list
Job · Mid-level

AI Infrastructure & Platform Operations Engineer (remote in the EU)

Other • Mid-level • Remote • Full-time • Germany Germany

Mirantis is building a European AI Infrastructure & Platform Operations team that operates environments based on NVIDIA GPUs, high-performance networking and Kubernetes across multiple datacenters. The role is remote in the EU, with a gross annual salary range of USD 60,000-67,000.

Responsibilities

  • ▹Monitor, operate and support production AI infrastructure platforms
  • ▹Investigate and resolve infrastructure, networking, hardware and platform-related incidents
  • ▹Support NVIDIA GPU infrastructure and associated platform services
  • ▹Monitor and troubleshoot Kubernetes-based environments
  • ▹Investigate performance, availability and reliability issues across infrastructure and platform components
  • ▹Collaborate with engineering teams, hardware vendors, datacenter personnel and service delivery teams
  • ▹Participate in incident response, root cause analysis and operational improvement
  • ▹Contribute to improvements in monitoring, observability, automation and operational processes
  • ▹Maintain operational documentation, runbooks and knowledge articles

Requirements

  • ▹3+ years of experience in infrastructure operations, platform operations, network operations, SRE, cloud operations, datacenter operations or related technical roles
  • ▹Strong Linux administration and troubleshooting skills
  • ▹Good understanding of networking concepts and experience diagnosing infrastructure-related issues
  • ▹Working knowledge of Kubernetes in production environments
  • ▹Experience supporting production infrastructure and services
  • ▹Experience working within structured operational and incident management processes
  • ▹Ability to work within a shift-based operational environment

Nice to have

  • ▹NVIDIA GPU infrastructure and accelerated computing platforms
  • ▹InfiniBand networking and NVIDIA UFM
  • ▹Kubernetes platform operations
  • ▹AI infrastructure or HPC environments
  • ▹Site Reliability Engineering (SRE) or Platform Engineering
  • ▹Observability platforms such as Grafana, Prometheus, ELK or OpenTelemetry
  • ▹Infrastructure automation technologies and Infrastructure-as-Code practices
  • ▹Large-scale distributed systems and production platforms

Soft skills

Strong analytical and problem-solving skillsExcellent communication and collaboration skills

What we offer

  • ▹Work with some of the most advanced AI infrastructure environments in production today
  • ▹Exposure to NVIDIA GPU technologies, Kubernetes platforms and high-performance networking
  • ▹Help define how next-generation AI infrastructure is operated
  • ▹Shape AI-powered operations through k0rdent AI
  • ▹Gross annual salary range: USD 60,000-67,000

About the company

Mirantis is the Kubernetes-native AI infrastructure company, helping organizations build and operate scalable, secure and sovereign infrastructure for AI, machine learning and data-intensive applications. Its customers include Adobe, PayPal and Volkswagen.

Similar jobs