About this role
Senior Manager, AI Infrastructure Operations at Vultr. Location: United States. Role: leading engineers, operating clusters, managing incidents Requirements: 6+ years in infrastructure/HPC; hands-on Linux, Kubernetes or Slurm, Terraform, Ansible; experience with GPU/bare-metal fleet ops, high-performance networking (InfiniBand/RDMA); leadership and incident management skills. Category: Engineering Seniority: Senior Level Tools: Linux, Kubernetes, Slurm, Terraform, Ansible, InfiniBand, RDMA Commitment: Full Time Workplace: Remote Languages: English