About this role
Remote London/Europe - we do need GMT+0 to GMT+4 of overlap with UK team. Zencargo is looking for a Senior Platform Engineer to support the architectural direction of our cloud platform as our traffic and our AI workloads grow, and to be the team's technical backstop when the hard problems land. This is a hands-on and broad role. Roughly 60% is infrastructure engineering: cloud architecture, a private Kubernetes estate, an event-driven backbone, and the infrastructure as code that describes all of it. Roughly 40% is building software, from our internal AI tooling and platform services to product engineering. You will collaborate closely with our forward deployed engineers and infrastructure team, owning solutions and driving outcomes that solve real business needs in a range of approaches from PoC work to long running projects. Lead the design, implementation and delivery of complex infrastructure projects. Build and scale the infrastructure behind our AI workloads, including cost, API reliability and platform support for LLM-powered features. Debug the hard problems: cluster behaviour under load, service-to-service traffic, consumer lag on the event backbone, and database performance. Own the infrastructure-as-code estate, keeping drift visible and reconciled, and setting module standards the rest of the team builds against. Advance the delivery pipeline, including policy-as-code gates, progressive delivery, and changes that are tested, reviewed and auditable. Contribute to our internal AI tooling, including the MCP server and agentic workflows that remove operational toil. Engineer cost controls through right-sizing, reserved capacity and waste removal, measured against a baseline and advocate for targeted spend. Mentor peers through pairing, code review and knowledge sharing, growing their judgement rather than answering for them. Cover security and governance alongside the platform work in collaboration with the infrastructure team. How success will be measured Success in this role will be measured by the outcomes delivered, including: Reduced time and manual effort in operational workflows, and improved speed, accuracy, cost or service quality across the business. Platform reliability and incident outcomes, including systemic fixes landed rather than repeat incidents patched. Complex infrastructure projects delivered end to end, with the decisions recorded and followable. Infrastructure-as-code and pipeline health, including drift reconciled and standards adopted by the wider team.