Senior Platform Engineer
Listed on 2026-08-22
-
IT/Tech
Cloud Computing: Infrastructure & Operations, Cybersecurity, Systems Engineer, SRE/Site Reliability
Location: London (hybrid)
Contract: Permanent, full-time
Cloud Gateway is a UK-based managed network security provider delivering secure
connectivity and network services to government and enterprise clients. We operate in a
regulated, security-first environment where reliability is non-negotiable.
About the RoleWe're looking for a Senior Platform Engineer to join our small but impactful platform team. You'll
own and operate a multi-account AWS estate supporting critical government services, build and
maintain CI/CD pipelines, and contribute to our growing observability and network automation
products.
This is a hands-on role with real ownership. You won't be pushing tickets around a queue; you'll
be designing infrastructure, writing code, responding to incidents, and shipping improvements to
production. You'll work closely with our Net Ops, Service Desk, and engineering teams to keep
things running and make them better.
What You'll Be Doing- Managing and evolving a multi-account AWS organisation with Terraform-managed resources across production and development environments.
- Building, maintaining, and optimising CI/CD pipelines for infrastructure deployments, application releases, Docker image builds, and database migrations.
- Operating and improving our observability stack (Grafana, InfluxDB, Prometheus, Open Telemetry, Cloud Watch) monitoring 200+ network devices and all AWS infrastructure.
- Supporting and developing our Observability-as-a-Service product built on Click House Cloud, HyperDX, and Open Telemetry, including customer onboarding and tenant management.
- Participating in on-call rotation, responding to production incidents, writing postmortem reports, and driving reliability improvements.
- Contributing to security and compliance initiatives including Cyber Essentials Plus certification, IAM governance, and vulnerability management.
- Building and maintaining golden AMI pipelines with automated patching, CIS hardening, and vulnerability scanning.
- Supporting the Network Automation Platform initiative, integrating with Forti Gate, Juniper, and Net Box for automated device configuration management.
- Maintaining and supporting legacy infrastructure across established accounts, including investigating issues in systems, understanding how they were built, and planning gradual modernisation where appropriate.
- Strong hands-on experience with AWS (EC2, ECS, Lambda, RDS, VPC, IAM, Cloud Watch, API Gateway, S3, Route
53). This is an AWS-heavy environment and you'll be working with these services daily. - Solid Terraform experience. We manage all infrastructure as code and you'll be writing, reviewing, and maintaining Terraform across multiple accounts and projects.
- Experience with CI/CD pipelines. We use Bitbucket Pipelines but the principles matter more than the specific tool. You should understand automated testing, quality gates, deployment strategies, and production approval workflows.
- Docker experience. We run production container images on ECS with multi-stage builds and ECR.
- Linux administration and troubleshooting. Amazon Linux and Ubuntu are our primary operating systems.
- Experience with monitoring and observability tools. Grafana, Prometheus, InfluxDB, Cloud Watch, or similar.
- Scripting ability in at least one of Type Script, Python, or Bash.
- A security-conscious mindset. Our clients are government bodies and everything we build has to meet strict compliance standards.
- Comfortable with on-call responsibilities and incident response.
- Comfortable working with legacy infrastructure and services.
- Experience with network automation tools (Ansible, NAPALM, Scrapli, NETCONF).
- Familiarity with Click House, Open Telemetry, or time-series data platforms.
- Experience working with government clients or in regulated environments.
- Exposure to Fin Ops practices and cost optimisation.
- AWS certifications (Solutions Architect Associate or similar).
- The chance to own a platform end-to-end, not just a slice of it.
- Real impact, the services you support help millions of people across the UK.
- A small team where your contributions are visible and valued.
- Exposure to a wide range of technologies across infrastructure, serverless, observability, and network automation.
- A growing product engineering function with opportunities to shape the direction of new
To Search, View & Apply for jobs on this site that accept applications from your location or country, tap here to make a Search: