← Zurück zur Liste
Stelle · Senior

Senior Systems HPC Engineer

Sonstige • Senior • Remote • Vollzeit • Niederlande Niederlande

Nebius is looking for a Senior Systems HPC Engineer to optimize performance across large-scale GPU clusters, working at the intersection of hardware, system software, networking (InfiniBand/RoCE) and virtualization (KVM/QEMU).

Responsibilities

  • ▹Investigate and troubleshoot GPU cluster performance issues
  • ▹Evaluate and integrate new hardware and tuning approaches
  • ▹Support complex performance escalations from internal teams and customers
  • ▹Collaborate with infrastructure, software and hardware vendor teams (NVIDIA, Mellanox, Intel)
  • ▹Contribute to hardware and cluster qualification/acceptance

Requirements

  • ▹5+ years in system-level software development with a performance focus
  • ▹3+ years hands-on Linux administration, troubleshooting and tuning
  • ▹Deep knowledge of server architecture (PCIe, NICs, Linux kernel, HPC systems)
  • ▹Strong proficiency in C/C++, Go or Python
  • ▹Coding interview as part of the hiring process

Soft skills

Systems thinking across multiple layersCollaboration with hardware vendor partnersProblem-solving and troubleshooting mindset

What we offer

  • ▹Competitive compensation
  • ▹Career growth and learning opportunities
  • ▹Flexibility and ownership
  • ▹International environment, talented teams
  • ▹Opportunity to work on impactful AI projects

About the company

Nebius, listed on Nasdaq (NBIS) and headquartered in Amsterdam, builds a full-stack AI cloud platform with a 1,500+ strong team and global R&D hubs.

Ähnliche Stellen