Senior Platform Engineer / DevOps
Quick Overview
Job Description
io.tt is growing quickly, and our infrastructure is central to how we deliver a fast, reliable, global platform to enterprise brands. We are looking for a hands-on DevOps Engineer to join our Lead DevOps Engineer and help build, run and evolve the cloud infrastructure that everything else depends on.
Our platform runs on AWS, with infrastructure defined as code in Pulumi (TypeScript). You will work across our AWS estate, covering compute, networking, databases and CI/CD. That means extending our Pulumi codebase, tightening our deployment pipelines and improving the observability and resilience of our services. This is a hands-on engineering role, so expect to spend most of your time writing infrastructure code and automation.
You will partner closely with the Lead DevOps Engineer and the wider engineering team, who work primarily in TypeScript and make heavy use of AI-assisted development. As our teams ship faster, a big part of this role is building the paved paths and guardrails that let engineers provision and deploy safely on their own. You will also make sure infrastructure code, including code generated with AI tooling, is reviewed, policy-checked and production-ready. Because you share a language and toolchain with the developers you support, you will be well placed to make infrastructure a self-service part of how we build.
This role reports to the Lead DevOps Engineer and is based in our central London (EC1V) office, hybrid with 2 days on-site per week. We offer 20 days holiday plus a festive shutdown, with an additional day per year of service (capped at 5), private medical insurance with Vitality after probation, a company pension and sick pay, a cycle to work scheme, and "Work From Anywhere" for one week a year (UK time). The office is dog-friendly and stocked with breakfast, drinks and snacks, with regular team events and a casual dress code.
- Building and maintaining our AWS infrastructure as code in Pulumi (TypeScript), extending existing stacks and creating new ones with clean, reviewable, reusable code.
- Owning and improving CI/CD pipelines, making deployments of our Bun and Node services fast, safe and repeatable.
- Managing core AWS services across compute, networking, storage and databases (including Aurora Serverless Postgres), with an eye on reliability, security and cost.
- Supporting our service topology behind Caddy, including deployment, routing and TLS termination handled at the load-balancer layer.
- Improving observability across metrics, logging, tracing and error monitoring, so issues are caught early and resolved quickly.
- Strengthening reliability and resilience through sensible automation, scaling, backups and disaster-recovery practices.
- Building self-service paved paths, such as reusable Pulumi components and golden-path templates that let engineers stand up services quickly within safe defaults.
- Establishing guardrails for AI-generated infrastructure, including policy-as-code (e.g. Pulumi CrossGuard or OPA), automated security and drift scanning, and PR gates that keep fast-moving IaC safe.
- Owning cost visibility and controls (FinOps), keeping cloud spend efficient as usage and provisioning scale.
- Managing secrets, IAM and infrastructure security in line with the expectations of enterprise customers.
- Participating in incident response and helping drive blameless post-incident improvements.
- Working alongside the Lead DevOps Engineer to shape infrastructure standards and mentor engineers on good operational practice.
- Strong hands-on experience with AWS across compute, networking, databases and IAM.
- Solid experience with infrastructure as code, ideally Pulumi, or strong Terraform or CDK experience with a clear willingness to work in Pulumi.
- Comfort writing infrastructure and automation in TypeScript (or strong general scripting ability and readiness to work in TS).
- Experience building and maintaining CI/CD pipelines for modern web and API services.
- A good grasp of networking, TLS, DNS and load balancing, and how services are exposed securely.
- Experience with observability tooling (metrics, logging, tracing, error monitoring) and using it to improve reliability.
- A pragmatic, ownership-minded approach: you automate the toil, document as you go, and care about cost as well as uptime.
- Comfort working in an AI-assisted engineering environment, using tools like Claude to accelerate your own infrastructure and automation work, and putting sensible review and policy checks around AI-generated code.
- 3 to 6+ years in a DevOps, SRE, platform or cloud engineering role.
- Proven experience running production workloads on AWS at scale.
- Hands-on infrastructure-as-code experience, ideally with Pulumi.
- Experience supporting a SaaS product and working closely with a software engineering team.
- Exposure to container-based and/or serverless architectures, and to managed databases such as Aurora Postgres and MongoDB Atlas, is an advantage.
Skills
Similar jobs
DevOps Engineer
Opus Recruitment Solutions · Newcastle Upon Tyne, United Kingdom
33 minutes agoSenior Site Reliability Engineer
VIQU IT · Wavendon, United Kingdom
33 minutes ago£70k/yrDevops Engineer (permanent)
AdRoc Group · United Kingdom
36 minutes ago£85k/yrSystems Engineer
Manpower · Cheltenham, United Kingdom
36 minutes agoSenior AWS Data Platform Engineer (Fixed Term Contract)
Microlise · Nottingham, United Kingdom
38 minutes agoDevOps Software Engineer
Saab UK · Bedford, United Kingdom
38 minutes ago