About this role
Why Join DAZN? DAZN is one of the world’s leading sports streaming platforms, delivering live and on-demand content to millions of fans across the globe. In an always-on environment, service continuity is critical. As a Technical Incident Manager, you’ll be at the forefront of restoring services when issues arise, owning the process of critical incident response across our global tech infrastructure. This is your opportunity to directly influence the availability, performance, and integrity of the product used by millions daily. If you're passionate about rapid troubleshooting, operational excellence, and leading under pressure, this is your arena. The Role: As a Technical Incident Manager n our Global Live Operations team, you’ll be responsible for managing all critical technical incidents (S1–S3) across DAZN’s platforms and services. This is a hands-on, high-impact role where you’ll lead service restoration efforts, coordinate technical teams, and drive rapid problem resolution while ensuring quality communication with the business. Be part of a 1 in 4 on-call rotation, supporting major live events. You’ll also be instrumental in operationalising new projects, defining monitoring requirements, and ensuring support readiness. You will act as a central figure for all things incident response—partnering across engineering, operations, support, and vendor teams. Technically lead the resolution of critical incidents, managing SMEs, technical sub-channels, and service restoration efforts. Coordinate with DAZN teams and third-party vendors during incidents via war rooms, Teams channels, and structured communication points. Own the integrity of the Major Incident Management process; ensure SLAs are met and processes followed. Communicate clear and timely updates to business stakeholders, shielding engineering teams for focused restoration. Provide technical guidance and mentorship during incidents, recommending improvements in troubleshooting and tooling. Partner with Dev, Engineering, and Support teams to manage escalations and ensure platform stability. Identify infrastructure, architectural, and process-level improvements to increase uptime and reduce MTTR. Operationalise new features, scope monitoring needs, produce runbooks, and assess delivery risks before launch