About this role
Lead System Engineer
Job description: Job Title: Cloud Platform / Systems Engineer Role Summary We are seeking a highly skilled Cloud Platform / Systems Engineer with strong experience in Azure cloud infrastructure, Linux/Windows operating systems, automation, and platform engineering. The ideal candidate should possess hands-on production support and engineering experience in cloud-native environments, with a strong focus on automation, DevOps practices, Kubernetes, and Infrastructure as Code (IaC). This role will support the transition from traditional infrastructure operations to a modern platform engineering model with end-to-end ownership of infrastructure lifecycle management, automation, deployment, observability, and operational excellence. Key Responsibilities • Manage and support Azure cloud infrastructure and hybrid environments. • Perform day-to-day systems administration across Linux and Windows environments. • Design, implement, and maintain Infrastructure as Code (IaC) using Terraform, Bicep, or Ansible. • Support Kubernetes/containerized workloads and cloud-native applications. • Build and maintain CI/CD pipelines for automated deployments and infrastructure provisioning. • Automate operational tasks including patching, monitoring, deployments, and server configurations using Shell/Python scripting. • Support platform engineering initiatives with focus on scalability, automation, and operational efficiency. • Collaborate with Infrastructure, Security, Networking, and Application teams to support end-to-end platform lifecycle management. • Participate in cloud migrations, VM migrations, decommission activities, and modernization initiatives. • Monitor infrastructure health, troubleshoot production issues, and perform root cause analysis. • Ensure operational readiness, documentation, observability, and compliance standards are maintained. • Support infrastructure provisioning and deployment activities across lower and higher environments. • Contribute to improving operational maturity through automation and process optimization. Required Skills & Experience • Cloud & Infrastructure • Strong hands-on experience with Microsoft Azure. • Experience managing hybrid cloud and on-prem infrastructure environments. • Good understanding of virtualization technologies including VMware. • Experience with storage, networking, and server infrastructure concepts. Operating Systems • Strong Linux administration experience. • Good understanding of Windows Server administration. • Knowledge of OS-level troubleshooting, services, cron jobs, scheduled tasks, permissions, sudo privileges, system performance tuning, and patching.
Hands-on experience: • Terraform Bicep • Ansible • Shell scripting • Python scripting • Understanding implementing Infrastructure as Code (IaC). • Understanding with CI/CD tools and DevOps practices. • Containers & Kubernetes • Container orchestration • Understanding of cloud-native architectures and microservices. • Monitoring & Operations • Experience with observability, monitoring, and alerting tools. • Strong troubleshooting and incident management skills. • Ability to support production systems and perform operational automation. Preferred Qualifications • Azure Administrator / Azure DevOps / Kubernetes certifications preferred. • Experience in Platform Engineering or Site Reliability Engineering (SRE) environments. • Experience supporting AI/Cloud-native application platforms. • Familiarity with Agile and DevOps operating models. • Exposure to enterprise infrastructure transformation projects. Soft Skills • Strong communication and collaboration skills. • Ability to work independently in fast-paced environments. • Strong analytical and problem-solving abilities. • Ability to quickly learn and adapt to new technologies. • Ownership mindset with focus on operational excellence and automation. Key Expectations • Candidate should possess practical production experience, not just training/certification knowledge. • Strong preference for engineers with hands-on automation and cloud operations experience. • Ability to support future-state platform engineering initiatives and cloud transformation programs. • Focus on reducing manual operational effort through automation and standardization.