Job Description

What you will do

  • Design, deploy, and run AWS infrastructure from end to end: from the edge and networking through to the multi-account cell architecture and the EKS platform.
  • Move between operational work, automation, requests from other engineering teams, and greenfield design, depending on what the platform needs that week.
  • Automate as much as sensibly possible: the infrastructure-as-code and GitOps that provision, upgrade, and keep the whole estate consistent and free of drift.
  • Work closely with the product and feature teams, turning what they need into self-service capabilities and helping them stay autonomous.
  • Keep the platform reliable, secure, and cost-effective: upgrades, autoscaling, observability, and security hardening and compliance (FedRAMP, FIPS). You will share the on-call rotation with the rest of the team.
  • Help the architecture grow: design new capabilities and take the platform to the next level of scale, resilience, and new kinds of workload.

How we work

A few things that are true about this team and, we think, make it a good place to build:

  • It is a team with real ownership. You look after your own areas and make the call on what is best for the platform, rather than working through a queue of tickets someone else has written.
  • We do not work in a silo. We spend a lot of time with the teams who build on us, including a good deal of design work together, so the platform grows around what people genuinely need.
  • The scale keeps it interesting. A cell-based setup across several regions on modern AWS tooling is a good deal more involved, and more rewarding, than a single-region estate.
  • We are improving something that already works. The platform is live and doing its job. Our task is to take it from good to really good, so the problems are about pushing things further rather than wiring up the basics.
  • Like everyone here we use AI tools day to day, and we are steadily moving the more repetitive operational work onto automation. We would rather have engineers than operators, and we do the automation carefully, because infrastructure is not somewhere to be careless.