Skip to content
ITA Jobs
All vacancies
Nebius logo

Senior Systems HPC Engineer

Nebius

Remote · EuropeSalary not disclosedfull-timeVerified recentlyOver a month oldNebius Careers

We are looking for a Senior Systems HPC Engineer to play a key role in building our hyperscaler platform, working across its core components while analyzing and optimizing the performance of large-scale GPU clusters at the intersection of hardware and software. You will operate across the full stack—from hardware and system software to networking (InfiniBand/RoCE), virtualization (KVM/QEMU), and distributed communication layers (e.g., MPI, NCCL).

Responsibilities

  • Focus on understanding system behavior across multiple layers, identifying performance bottlenecks, and driving improvements that shape how our clusters are built, operated, tuned, and validated.
  • Investigate and troubleshoot performance issues of GPU cluster under real workloads (training and inference)
  • Evaluate and integrate new hardware, system configurations and tuning approaches through software stack
  • Support complex performance-related escalations from internal teams and customers
  • Work closely with infrastructure, software engineering and hardware vendor teams (e.g. NVIDIA, Mellanox, Intel)
  • Contribute to hardware and cluster qualification (acceptance), ensuring systems meet performance expectations

Languages

    Work format
    Remote
    Seniority
    Senior
    Posted
    24 Apr 2026 (5mo ago)
    Last verified
    4 Oct 2026

    Keep looking