About this role
Salary: £65,000 - 105,000 per year
Requirements: Deep, hands-on experience in SRE, DevOps, or Platform Engineering across both AWS and AzureStrong track record operating Docker and Kubernetes in production environmentsPractical, enterprise-level experience supporting IIS/.NET applications and Java Spring Boot servicesExperience administering both Linux and WindowsExperience building and maintaining CI/CD pipelines, primarily with GitHub Actions; Bamboo experience is a plusAbility to script confidently in Python, PowerShell, or BashExperience with cloud security best practices, including IAM, secrets management, and container/image scanningUnderstanding of core infrastructure fundamentals, including networking, storage, and DNSExperience with APM and telemetry tooling Responsibilities: Embed with Fitch Ratings development squads to provide Service Reliability Engineering expertiseLead the delivery of reliable, scalable, mission-critical servicesGuide squads on Kubernetes and modern deployment patternsMentor associate engineers and establish best practicesPartner closely with Fitch Ratings Development Squads and Operations to design and advance service builds, DevOps tooling, and operational excellencePartner with Core Engineering to architect and govern GitHub Actions CI/CD, including quality gates, canary/blue-green strategies, and AI-assisted redeploy checksOwn observability in Datadog, including defining SLIs/SLOs, dashboards, alerting, and MS Teams integrationsReduce incidents through telemetry-driven automation and blameless postmortemsChampion AI-enabled operations using AWS Bedrock, SageMaker, and Model Context Protocol for log analysis, anomaly detection, incident triage, and workflow orchestrationEstablish guardrails for AI-enabled operations adoptionDefine and enforce cloud guardrails and security controls in partnership with Security and RiskInfluence cross-functional roadmaps, lead complex release planning, and drive strategic platform initiatives across CI&PEServe as an escalation point and participate in the L3 on-call rotation Technologies: AIAWSArchitectAzureBambooBashCI/CDCloudDatadogDevOpsDockerGitHubIAMSupportJavaKubernetesLinuxMS TeamsPowerShellPythonSecuritySpringSpring BootWindowsASP.NETAgentic AIDevSecOpsEmbeddedMCPModel ServingNodeJS More:
We are Fitch Group, a leading global financial information services provider delivering credit and risk insights, data, and tools that support more efficient and transparent financial markets. Our technology and data organization is a dynamic team focused on innovation, cloud, AI, and modern engineering practices, and we are recognized by Built In as a Best Place to Work in Technology for three years in a row. This Service Reliability Engineer role is based in our Manchester office, embeds with Fitch Ratings development squads, and is part of Cloud Infrastructure & Platform Engineering. We offer a hybrid work environment with 2 to 3 days a week in the office, along with learning and mobility opportunities, retirement planning, tuition reimbursement, healthcare, parental leave, inclusive employee resource groups, and volunteer and donation matching programs.
last updated 36 week of 2026