NVIDIA is looking for a Senior Performance Engineer to profile, benchmark, and optimize AI and HPC workloads across large-scale GPU and CPU clusters, working closely with hardware, networking, and software teams.
Responsibilities
- ▹Profile, benchmark, and analyze AI and HPC workloads on GPU and CPU clusters
- ▹Explore performance characteristics of high-performance networking and collective communications (NCCL, RDMA, MPI, RoCE)
- ▹Identify performance bottlenecks across networking, compute, memory, and system architecture
- ▹Develop and enhance performance analysis, benchmarking, and diagnostic tools
- ▹Define performance test plans and establish expectations for new technologies and platforms
- ▹Collaborate across hardware, firmware, networking, systems, and software teams to provide actionable performance insights
- ▹Support telemetry collection and data refinement efforts to enable accurate performance analysis
- ▹Maintain high standards for data quality, reproducibility, and traceability of performance results
Requirements
- ▹B.Sc. or M.Sc. in Computer Science, Computer Engineering, Software Engineering, or equivalent experience
- ▹5+ years of experience in performance analysis, systems engineering, or HPC/AI infrastructure
- ▹Demonstrated expertise in performance analysis methodologies
- ▹Hands-on experience with high-performance networking (RDMA, MPI, NCCL, congestion control)
- ▹Strong understanding of system performance metrics (latency, throughput, resource utilization)
- ▹Exposure to hardware, firmware, or embedded telemetry environments
Nice to have
- ▹Knowledge of CUDA, NCCL internals, and congestion control algorithms
- ▹Deep system-level understanding of CPU architectures, GPUs, HCAs, memory, and PCIe
- ▹Experience with NVIDIA GPUs, CUDA, and deep learning frameworks such as PyTorch or TensorFlow
- ▹Experience with cloud platforms
- ▹Proficiency in Python, experience with Bash and C/C++ is a plus, strong experience working in Linux environments
Soft skills
Strong analytical and problem-solving skillsStrong communication skillsAbility to work effectively in cross-functional, fast-paced R&D teams
What we offer
- ▹Highly competitive salaries and a comprehensive benefits package
About the company
NVIDIA is passionate about supercomputing and ground-breaking technologies, with products spanning high-performance computing, machine learning, cloud services, and storage. The company is widely considered one of the technology world's most desirable employers and is committed to fostering a diverse work environment.
Education: B.Sc. vagy M.Sc. diploma számítástechnika, számítógép-mérnöki vagy szoftvermérnöki területen, vagy ezzel egyenértékű tapasztalat
Similar jobs

Job
· Senior
Senior Machine Learning Scientist
Axon
AI/MLCppLlm
+2
💰 Salary: not specified
🏢 On-site
Scottsdale
🗣️ EN

Job
· Senior
Senior Machine Learning Scientist
Axon
AI/MLCppLlm
+2
$159,750–$255,600/yr
gross
🏢 On-site
Seattle
🗣️ EN

Job
· Senior
Senior Machine Learning Scientist
Axon
AI/MLCppLlm
+2
$159,750–$255,600/yr
gross
🏢 On-site
Boston
🗣️ EN

Job
· Senior
Senior HPC AI Cluster Engineer
NVIDIA
AI/ML
+6
💰 Salary: not specified
🌍 Remote
🗣️ EN
Himalayas

Job
· Senior
Senior Android BSP Engineer
Jabil
Cpp
+3
💰 Salary: not specified
🌍 Remote
🗣️ EN
Himalayas

Job
· Senior
Senior Systems HPC Engineer
Nebius
AI/MLCpp
💰 Salary: not specified
🌍 Remote
🗣️ EN
Himalayas