← Zurück zur Liste

Stelle
· Senior
Senior AI Infrastructure & Platform Operations Engineer
Sonstige
• Senior
• Remote
• Vollzeit
•
Polen
Role Overview As a Senior AI Infrastructure & Platform Operations Engineer, you will serve as a technical leader within the operations organization, providing deep expertise across infrastructure, networking, platform operations, and service reliability. You will be responsible for driving operational excellence across complex production environments while acting as a key escalation point for critical incidents and challenging technical issues. What You Will Do Lead the investigation and resolution of complex infrastructure, networking, and platform-related incidents. Support large-scale NVIDIA GPU infrastructure and high-performance networking environments. Troubleshoot complex Linux, Kubernetes, networking, storage, and hardware-related issues. Why It Might Be a Fit We offer: Operate some of the most advanced AI infrastructure environments in production today. Work with the latest NVIDIA GPU technologies, Kubernetes platforms, and high-performance networking environments. Help define operational standards and reliability practices for next-generation AI infrastructure services. Requirements 7+ years of experience in infrastructure operations, platform operations, site reliability engineering, network operations, cloud operations, datacenter operations, or related technical roles. Expert-level Linux administration and troubleshooting skills. Strong networking expertise, including experience diagnosing complex performance, connectivity, and reliability issues. Strong experience operating Kubernetes in production environments. Experience supporting large-scale production infrastructure and distributed systems. Proven experience leading technical investigations and managing complex incidents. Experience performing root cause analysis and driving long-term operational improvements. Strong understanding of observability, monitoring, and service reliability practices. Excellent troubleshooting and analytical skills across multiple infrastructure domains. Strong communication, collaboration, and stakeholder management skills. Benefits Operate some of the most advanced AI infrastructure environments in production today. Work with the latest NVIDIA GPU technologies, Kubernetes platforms, and high-performance networking environments. Help define operational standards and reliability practices for next-generation AI infrastructure services. Influence the adoption of AI-powered operational capabilities through k0rdent AI. Work alongside highly skilled engineers solving complex infrastructure and platform challenges at scale. Join a growing organisation investing heavily in AI infrastructure, platform services, and operational innovation. Originally posted on Himalayas
Ähnliche Stellen

Stelle
· Senior
Senior DevSecOps Engineer (Kubernetes, CI/CD)
Capgemini
💰 Gehalt: keine Angabe
🏢 Vor Ort
Wrocław
🗣️ EN

Stelle
· Senior
L3 Linux/Kubernetes support/specialist
Infotree Global Solutions
Openshift
+2
💰 Gehalt: keine Angabe
🌍 Remote
🗣️ EN
Himalayas

Stelle
· Senior
Systems Engineer - Senior
SOFTSWISS
Argocd
+15
💰 Gehalt: keine Angabe
🌍 Remote
🗣️ EN
Himalayas

Stelle
· Senior
SENIOR IT OPERATIONS SPECIALIST
JYSK
Github Actions
+5
💰 Gehalt: keine Angabe
🏢 Vor Ort
Gdańsk

Stelle
· Senior
Senior API Engineer
Equinix, Inc
+2
7 454 €/Mon.
brutto
🔀 Hybrid
Warsaw
🗣️ EN

Stelle
· Senior
Senior Systems Engineer – Performance & Reliability (Analysis)
Graphcore
CppGithub Actions
+6
6 766–9 153 €/Mon.
brutto
🏢 Vor Ort
Gdańsk
🗣️ EN