Now hiring

Senior DevOps Engineer - Healix (Toronto, ON, CA, M5H 0A9) @ Deloitte

Toronto, ON, CA, M5H 0A9OnsiteFull-time
Apply with ResuMinder

Opens on the employer's site

About this role

Job Type: Permanent Work Model: Hybrid Reference code: 135131 Primary Location: Toronto, ON All Available Locations: Toronto, ON; Burlington, ON; Calgary, AB; Edmonton, AB; Fredericton, NB; Halifax, NS; Kitchener, ON; Moncton, NB; Ottawa, ON; Regina, SK; Saint John, NB; Saskatoon, SK; St. John's, NL; Winnipeg, MB Our Purpose At Deloitte, our Purpose is to make an impact that matters. We exist to inspire and help our people, organizations, communities, and countries to thrive by building a better future. Our work underpins a prosperous society where people can find meaning and opportunity. It builds consumer and business confidence, empowers organizations to find imaginative ways of deploying capital, enables fair, trusted, and functioning social and economic institutions, and allows our friends, families, and communities to enjoy the quality of life that comes with a sustainable future. And as the largest 100% Canadian-owned and operated professional services firm in our country, we are proud to work alongside our clients to make a positive impact for all Canadians. By living our Purpose, we will make an impact that matters. Have many careers in one Firm. Enjoy flexible, proactive, and practical benefits that foster a culture of well-being and connectedness. Learn from deep subject matter experts through mentoring and on the job coaching Summary We are hiring a Senior DevOps Engineer to own and evolve our cloud platform on AWS, grounded in infrastructure as code, secure multi-account patterns, and reliable delivery. You will shape the DevOps roadmap (standards, tooling, automation, and operational excellence), support application releases, and provide production support for critical workloads. Amazon EKS is central to how we run workloads—we need someone with deep production-grade EKS expertise who has built and owned Kubernetes on AWS end-to-end, not only deployed apps to a cluster someone else runs. What will your typical day look like? You will lead how we adopt AI for infrastructure and platform work—not as a buzzword, but as a practical force multiplier: safe use of AI-assisted authoring and review for IaC and automation, clearer runbooks and incident workflows, and evaluation of tools and patterns that improve speed without weakening security, compliance, or change control. This role suits someone who combines deep AWS practice with leadership: you can define “how we build and run” while still being hands-on in pipelines, clusters, and incidents. Key Responsibilities: Roadmap & standards: Define and socialize DevOps priorities (security, reliability, cost, velocity). Align teams on AWS Well-Architected practices, tagging, guardrails, and repeatable patterns for networking, identity, secrets, and data. AI adoption for infra & platform: Drive a pragmatic AI strategy for the team—e.g. standards for AI-assisted IaC and pipeline changes (review gates, testing, drift detection), documentation and runbook quality, incident summarization and triage workflows where appropriate, and guardrails so AI tooling fits regulated or high-stakes environments. Stay current on vendor and open-source options; pilot, measure, and roll out what actually reduces toil. Infrastructure as code: Design, review, and implement changes using Terraform and Terragrunt, with clear module boundaries, environmentspecific config, and safe promotion across dev → non-prod → production. EKS (critical): Build, operate, and own the Kubernetes platform on AWS —cluster lifecycle (creation, upgrades, patching), node groups / capacity, networking (CNI, service mesh or ingress as used), security (RBAC, admission controls, pod security, secrets and IRSA), add-ons, and cost/ reliability tuning. Partner with app teams on standards for workloads, namespaces, and safe rollouts; be the escalation point for cluster-level incidents. Broader AWS platform: Operate and improve adjacent services—e.g. RDS/Aurora, DynamoDB, object storage and CDN, KMS, Secrets Manager, SNS (alerting), Lambda, EventBridge, and CI/CD (CodePipeline / CodeBuild, connections to source control)—plus IAM, VPC, and multi-tenant or multi-namespace patterns where applicable. Release engineering: Partner with development teams on release processes, deployment strategies, change management, rollbacks, and post-release verification in regulated or high-stakes environments (e.g. healthcare-adjacent workloads). Production support: Participate in on-call or escalation rotation as defined by the team; troubleshoot incidents, drive root-cause analysis, and implement preventive fixes (runbooks, dashboards, alarms, automation). Observability & operations: Improve monitoring, logging, tracing, and alerting; tune thresholds; reduce noise; document operational procedures. Collaboration: Work with security, architecture, and engineering leads to implement least-privilege access, encryption, backup/DR posture, and audit-friendly operations—including how AI-assisted workflows meet security and audit expectations. About the team Healix is an AI-native venture built inside Deloitte, creating a governed AI workflow layer that lifts the operational burden on hospital staff — without disrupting the systems they already rely on. Our platform sits across existing hospital infrastructure to deploy and scale AI-enabled workflows with standardized trust controls applied once, so hospitals can start with one use case and expand to many without restarting governance from scratch. The thesis is simple: do the work you already trust Deloitte with, just faster, smarter, and at scale. We're backed by Deloitte, the world's largest professional services firm, which means we have real enterprise distribution, real health system access, and real institutional credibility from day one. That's not a small thing in healthcare — trust is the product, and we come pre-loaded with it. We're also early enough that the person we hire here will define what Healix's product org looks like for the next decade. This isn't a point solution play. Healix is being built as the foundational AI operating layer for hospital operations — the platform every future AI-enabled workflow runs on top of, inheriting the same security, auditability, and human-oversight controls without multiplying IT burden or compliance risk. Enough about us, let’s talk about you You are someone who has these skills, experience & qualifications: 6+ years of experience in software engineering, systems engineering, DevOps, or Site Reliability Engineering (SRE), including at least 4 years working with AWS in production environments. Strong expertise in Infrastructure as Code (IaC) using Terraform, with experience designing modular, environment-driven infrastructure. Experience with Terragrunt or similar composition frameworks is considered an asset. Deep, hands-on expertise with Amazon EKS. This role requires direct experience building, operating, and owning Kubernetes platforms on AWS, not simply deploying applications to an existing cluster. You should be comfortable with: Cluster architecture, lifecycle management, upgrades, and patching Networking, including VPCs, CNI, DNS, and ingress Security and identity management, including RBAC, IRSA, secrets management, and governance controls Observability, monitoring, and troubleshooting Capacity planning, performance optimization, and production support Experience limited to basic Kubernetes usage or application deployment is not sufficient. Strong understanding of CI/CD pipelines, artifact promotion strategies, secrets management, and safe deployment practices across multiple environments. Experience managing production incidents, including triage, stakeholder communication, root cause analysis (RCA), and implementing long-term corrective actions. Demonstrated interest in applying AI to platform engineering and DevOps workflows, such as AI-assisted infrastructure development, code reviews, operational tooling, or documentation. Candidates should understand the limitations, validation requirements, and risks associated with using AI in production environments. Proven ability to influence technical direction without direct authority through standards, architecture reviews, and roadmap recommendations that are successfully adopted by engineering teams. Excellent communication and collaboration skills, with the ability to work effectively across distributed teams and with stakeholders beyond engineering. It would be great for you to have some of these nice to haves as well: AWS certifications such as AWS Solutions Architect Professional or AWS DevOps Engineer Professional, or equivalent practical experience demonstrating similar depth of expertise. Kubernetes certifications such as Certified Kubernetes Administrator (CKA) or Certified Kubernetes Security Specialist (CKS), or equivalent demonstrated experience managing Kubernetes and EKS platforms at scale. Experience with Helm, Helmfile, policy-as-code frameworks, or Kubernetes cluster baseline tooling. Familiarity with PostgreSQL, Amazon RDS, multi-tenant architectures, or technology environments subject to regulatory requirements. Experience defining and managing Service Level Objectives (SLOs), error budgets, and platform performance metrics. Exposure to cloud cost optimization and FinOps practices, including rightsizing resources, scheduling non-production workloads, and implementing storage lifecycle strategies. Hands-on experience evaluating or using AI coding assistants, internal LLM or RAG solutions for operational knowledge management, or AI tools that support platform engineering teams. Total Rewards The salary range for this position is $105,000 - $175,000, and individuals may be eligible to participate in our bonus program. Deloitte is fair and competitive when it comes to the salaries of our people. We regularly benchmark across a variety of positions, industries, sectors, targets, and levels. Our approach is grounded on recognizing people's unique strengths and contributions and rewarding the value that they deliver. Our Total Rewards Package extends well beyond traditional compensation and benefit programs and is designed to recognize employee contributions, encourage personal wellness, and support firm growth. Along with a competitive base salary and variable pay opportunities, we offer a wide array of initiatives that differentiate us as a people-first organization. On top of our regular paid vacation days, some examples include: $4,000 per year for mental health support benefits, a $1,300 flexible benefit spending account, firm-wide closures known as "Deloitte Days", dedicated days of for learning (known as Development and Innovation Days), flexible work arrangements and a hybrid work structure. Our promise to our people: Deloitte is where potential comes to life. Be yourself, and more. We are a group of talented people who want to learn, gain experience, and develop skills. Wherever you are in your career, we want you to advance. You shape how we make impact. Diverse perspectives and life experiences make us better. Whoever you are and wherever you’re from, we want you to feel like you belong here. We provide flexible working options to support you and how you can contribute. Be the leader you want to be Some guide teams, some change culture, some build essential expertise. We offer opportunities and experiences that support your continuing growth as a leader. Have as many careers as you want. We are uniquely able to offer you new challenges and roles – and prepare you for them. We bring together people with unique experiences and talents, and we are the place to develop a lasting network of friends, peers, and mentors. The next step is yours At Deloitte, we are all about doing business inclusively – that starts with having diverse colleagues of all abilities. Deloitte encourages applications from all qualified candidates who represent the full diversity of communities across Canada. This includes, but is not limited to, people with disabilities, candidates from Indigenous communities, and candidates from the Black community in support of living our values, creating a culture of Diversity Equity and Inclusion and our commitment to our AccessAbility Action Plan , Reconciliation Action Plan and the BlackNorth Initiative . We encourage you to connect with us at [email protected] if you require an accommodation for the recruitment process (including alternate formats of materials, accessible meeting rooms or other accommodations) or [email protected] for any questions relating to careers for Indigenous peoples at Deloitte (First Nations, Inuit, Métis). When you apply, we will review your application using Deloitte's Global Talent Standards to ensure a consistent recruitment experience. Our recruitment advisors and hiring teams will utilize human screening combined with AI technology to help identify the skills and qualities that matter most to our business, while safeguarding your privacy and using AI responsibly. Deloitte Canada has 20 offices with representation across most of the country. We acknowledge that Deloitte offices stand on traditional, treaty, and unceded territories in what is now known as Canada. We recognize that Indigenous Peoples have been the caretakers of this land since time immemorial, nurturing its resources and preserving its natural beauty. We acknowledge this land is still home to many First Nations, Inuit, and Métis Peoples, who continue to maintain their deep connection to the land and its sacred teachings. We humbly acknowledge that we are all Treaty people, and we commit to fostering a relationship of respect, collaboration, and stewardship with Indigenous communities in our shared goal of reconciliation and environmental sustainability.

Ready to apply?

Install the ResuMinder extension and we'll auto-fill the application in seconds — no rewriting.

See how your CV scores