About this role
<div> <div> <p><span>Role Mission: </span><span> </span></p> </div> <div> <p><span>As a Lead CRM DevOps Engineer at StarHub, you will own the end-to-end DevOps strategy, platform reliability, and release governance for mission-critical CRM and Order Management (CRM/OM) platforms. You will lead the design and evolution of CI/CD pipelines, Kubernetes platform architecture, and cloud infrastructure, ensuring high availability, scalability, and secure delivery of CRM/OM microservices on AWS EKS.</span><span> </span></p> </div> <div> <p><span>This role goes beyond execution requiring you to drive DevOps best practices, establish SRE capabilities, and ensure end-to-end stability of customer journeys (order, provisioning, billing). You will act as a technical leader, guiding engineers and working across teams to enable faster, safer, and more reliable platform delivery.</span><span> </span></p> </div> <div> <p><span>Responsibilities: </span><span> </span></p> </div> <div> <ul style="list-style-type:disc"> <li> <p><span>Drive DevOps strategy and operating model for CRM/OM platforms, acting as the technical escalation point for complex production issues.</span><span> </span></p> </li> </ul> </div> <div> <ul style="list-style-type:disc"> <li> <p><span>Architect and govern end</span>‑<span>to</span>‑<span>end CI/CD pipelines using Jenkins (Pipeline</span>‑<span>as</span>‑<span>Code) and GitOps (Argo CD / Flux), enabling safe, repeatable releases.</span><span> </span></p> </li> </ul> </div> <div> <ul style="list-style-type:disc"> <li> <p><span>Lead trunk</span>‑<span>based development adoption, automated testing, quality gates, and release safety mechanisms across microservices.</span><span> </span></p> </li> </ul> </div> <div> <ul style="list-style-type:disc"> <li> <p><span>Design and operate multi</span>‑<span>cluster Kubernetes platforms (EKS/OCP), including networking, RBAC, scaling, and resilience.</span><span> </span></p> </li> </ul> </div> <div> <ul style="list-style-type:disc"> <li> <p><span>Govern AWS cloud infrastructure and Infrastructure</span>‑<span>as</span>‑<span>Code using Terraform and/or CloudFormation, ensuring security, scalability, and cost efficiency.</span><span> </span></p> </li> </ul> </div> <div> <ul style="list-style-type:disc"> <li> <p><span>Apply SRE practices by defining SLOs/SLIs, improving reliability and performance, and leading incident management and RCA.</span><span> </span></p> </li> </ul> </div> <div> <ul style="list-style-type:disc"> <li> <p><span>Ensure end</span>‑<span>to</span>‑<span>end reliability of CRM and API integrations across systems.</span><span> </span></p> </li> </ul> </div> <div> <ul style="list-style-type:disc"> <li> <p><span>Establish observability and operational excellence using Splunk, Prometheus, Grafana, CloudWatch, and on</span>‑<span>call tooling.</span><span> </span></p> </li> </ul> </div> <div> <ul style="list-style-type:disc"> <li> <p><span>Contribute to automation and code improvements, participate in architecture reviews, mentor engineers, and drive documentation and knowledge sharing.</span><span> </span></p> </li> </ul> </div> </div> <div> <div> <p><span> </span></p> </div> <div> <p><span>Qualifications:</span><span> </span></p> </div> <div> <ul style="list-style-type:disc"> <li> <p><span>Bachelor’s degree in Computer Science, Software Engineering, or a related field.</span><span> </span></p> </li> </ul> </div> <div> <ul style="list-style-type:disc"> <li> <p><span>6+ years of hands-on experience across DevOps, SRE, and platform or software engineering, with at least 2–3 years in a senior or technical leadership role.</span><span> </span></p> </li> </ul> </div> <div> <ul style="list-style-type:disc"> <li> <p><span>Strong proficiency in Java and Spring Boot, with hands-on experience supporting microservices-based and distributed architectures.</span><span> </span></p> </li> </ul> </div> <div> <ul style="list-style-type:disc"> <li> <p><span>Deep understanding of scalability, reliability engineering, and production system performance.</span><span> </span></p> </li> </ul> </div> <div> <ul style="list-style-type:disc"> <li> <p><span>Advanced hands-on experience with:</span><span> </span></p> </li> </ul> </div> <div> <ul style="list-style-type:circle"> <li> <p><span>Kubernetes (EKS preferred), Helm, and container orchestration</span><span> </span></p> </li> </ul> </div> <div> <ul style="list-style-type:circle"> <li> <p><span>CI/CD platforms (Jenkins, GitLab CI) and GitOps practices</span><span> </span></p> </li> </ul> </div> <div> <ul style="list-style-type:circle"> <li> <p><span>AWS cloud services and architecture</span><span> </span></p> </li> </ul> </div> <div> <ul style="list-style-type:circle"> <li> <p><span>Infrastructure as Code (Terraform and/or CloudFormation)</span><span> </span></p> </li> </ul> </div> <div> <ul style="list-style-type:disc"> <li> <p><span>Strong SQL skills with hands-on PostgreSQL administration and performance tuning experience.</span><span> </span></p> </li> </ul> </div> <div> <ul style="list-style-type:disc"> <li> <p><span>Proficiency in Python and shell scripting for automation and platform tooling.</span><span> </span></p> </li> </ul> </div> <div> <ul style="list-style-type:disc"> <li> <p><span>Experience implementing observability and monitoring using tools such as Splunk, Prometheus, Grafana, and CloudWatch.</span><span> </span></p> </li> </ul> </div> <div> <ul style="list-style-type:disc"> <li> <p><span>Proven experience working in Agile/Scrum environments, with strong stakeholder management, communication, and technical leadership skills.</span><span> </span></p> </li> </ul> </div> <div> <ul style="list-style-type:disc"> <li> <p><span>Ability to operate effectively in a fast-paced, production-critical environment.</span><span> </span></p> </li> </ul> </div> </div>