About this role
Job Description Summary The HPC Infrastructure Services (Senior) Engineer executes and ensures service deployments and management to enable modern delivery flows of High-Performance Computing solutions. Job Description Job Description Major accountabilities:
• High Performance Computing (HPC) Specialist for Novartis Core Infrastructure Services team performs the operational health deployment and lifecycle of HPC technology • Designs innovative solutions utilizing projects and qualification frameworks to run and optimize our complex infrastructures. • Act as technical and organizational escalation point during major and critical incidents. • Responsible to track suppliers and partner for effective and efficient delivery of projects • Contributes to service / platform strategy development
Commitment to Diversity & Inclusion: We are committed to building an outstanding, inclusive work environment and diverse teams, representative of the patients and communities we serve.
What you’ll bring to the role
• 6+ years’ IT experience, of which at least 3 in an HPC/scientific computing environment • Associate's degree in computer science or related field, or equivalent industry experience • Experience designing and building infrastructure within a public cloud platform (e.g. AWS, GCP or Azure) as they relate to HPC workloads • Hands-on experience with system administrative task in Linux environment, and fluency in scripting (shell scripting in bash required, Python knowledge would be advantageous) • Can build software from sources including knowledge of build systems (e.g. Make, CMake...) • Extensive knowledge of one or more HPC scheduling mechanisms (e.g. Grid Engine, Slurm, LSF... etc.) • Extensive knowledge of one or more HPC cluster management software packages (e.g. Bright Cluster Manager / Base Command Manager, xCat, OpenHPC… etc.) • Solid knowledge of computer architectures and multi-threaded/parallel processing applications • Hands-on knowledge of network- and distributed filesystems (e.g. NFS, GPFS, Gluster, BeeGFS, Lustre or other parallel file systems, etc.) • Worked in environments with high speed / low latency network (e.g. InfiniBand) • Strong knowledge of Unix systems performance tuning • Experience with shared memory and/or GPU and/or distributed memory computation jobs • Working knowledge of DevOps (e.g. Ansible, Chef, CICD etc.) tools and practices • Desired: experience setting up, maintaining, and tuning infrastructure for AI/ML workloads • Desired: experience with HPC infrastructure supporting a scientific / research / pharmaceutical / bioinformatics / cheminformatics environment • Exceptional service orientation, excellent problem-solving abilities, keen attention to detail. • Fluent English speaker
Skills Desired Communication Skills (Inactive), Information Technology (IT) Infrastructure, Information Technology Operations, IT Service Management (ITSM), Problem Solving Skills (Inactive), System Integration, Vendor Management