About this role
Member of Technical Staff, AI Supercomputing at Radical Numerics. Location: San Francisco or Tokyo. Role: operating clusters, building interfaces, maximizing throughput Requirements: Proven experience operating large-scale GPU clusters and orchestration (Kubernetes/Slurm); backend development (Python or Rust); strong Linux, networking, and systems skills; familiarity with deep learning frameworks and performance tuning. Category: Engineering Seniority: Mid Level Tools: Kubernetes, Slurm, Python, Rust, PyTorch, Triton, CUDA, C++, NCCL, Torchtitan, Megatron-LM Commitment: Full Time Workplace: Onsite Languages: English