К содержимому

Senior Infrastructure Engineer (GPU Platform)

Грейд
Senior
Категория
DevOps
Источник
IT Outstaff Projects

Контакт HR — бесплатно после входа

Войдите через Telegram — покажем, кому писать напрямую. Без резюме и анкет.

Описание

We are looking for a Senior Infrastructure Engineer (GPU Platform)

Requirements:
• Production experience with multi-node GPU training infrastructure
• Strong Linux, containers, CUDA, and NVIDIA GPU stack knowledge
• Hands-on experience with NCCL and InfiniBand or RoCE/RDMA troubleshooting
• Deep experience with Kubernetes or Slurm
• Experience with infrastructure automation and observability
• Experience diagnosing issues across training workloads, networking, storage, and GPU hosts
• Strong incident leadership and provider-facing communication skills
• English – Upper-Intermediate or higher

Would be a plus:
• Experience in an AI lab, HPC environment, or specialist GPU cloud
• PyTorch, Megatron, DeepSpeed, or other distributed-training frameworks
• Experience with parallel storage and checkpoint optimization
• Experience working with multi-provider GPU platforms

📩 Send your CV:
cv@talentstoday.com
  • linux
  • containers
  • cuda
  • nvidia
  • gpu
  • nccl
  • infiniband
  • roce
  • rdma
  • kubernetes
  • slurm
  • pytorch
  • megatron
  • deepspeed

Оценка вакансии

38/100 · минимум информации

  • Описание полное
  • Зарплата не указана
  • Компания не названа
  • Контакт без валидации
  • Стек описан подробно
  • Формат не указан
Как считается

Похожие вакансии