Principal Cloud Engineer
Listed on 2026-08-15
-
IT/Tech
Cloud Computing: Infrastructure & Operations, AWS, Systems Engineer
The Role
Own the cloud infrastructure, CI/CD systems, and deployment automation for Tetra Science's multi-tenant SaaS platform serving global biopharma customers. This is a hands-on technical lead role. You will lead through technical depth and influence across teams. Strong architecture and implementation skills are important for success in this role. You will evolve our cloud architecture, build substantial parts of it in Python, Cloud Formation and Terraform.
You will architect and build deployment pipelines to AWS and Databricks, and drive the engineering practices that determine how fast and safely we ship software.
Own, design, build, and maintain the cloud infrastructure using Cloud formation, Terraform and custom Python glue.. Every environment is provisioned and governed through code. Architect the deployment pipeline infrastructure end to end. Git Hub Actions, Code Build, container image pipelines, code scanning, artifact registries, pre-merge integration environments, promotion gates, and automated rollback. Your goal: engineers merge code and it reaches production safely without manual intervention.
You will partner with other engineering teams to reduce cycle time from commit to production. Instrument pipeline metrics (build time, deployment frequency, change failure rate, MTTR). Identify and eliminate bottlenecks. Build self-service capabilities so product teams are not blocked by infrastructure.
Deep, hands-on AWS experience:
Serverless Architecture, EKS/ECS, VPC/networking, IAM, KMS, Cloud Watch, Lambda, S3, EC2, Kinesis, Athena, Glue, Cloud Trail, Cost Explorer. You understand Well-Architected Framework principles and apply them daily, not as a checklist exercise.
Databricks experience is strongly preferred.
Embed security into the product and pipelins: container image scanning, SAST/DAST integration, secrets management, least-privilege IAM, and compliance-as-code. You work in a GxP-regulated environment where auditability and traceability of deployments are non-negotiable.
Observability and ReliabilityProduction monitoring, alerting, log aggregation, and incident response infrastructure. Support for developer teams. Blameless postmortem culture.
Current Tech Stack- Cloud: AWS
- IaC:
Terraform, Cloud Formation - CI/CD:
Git Hub Actions - Containers:
Docker, ECS - Languages:
Python, Bash - Data:
PostgreSQL / Aurora, S3 data lake, Databricks (Lakehouse)
Tetra Science is building the data and AI platform for drug development. Our customers are global pharma companies running regulated scientific workloads. The infrastructure you build determines whether we ship features weekly or monthly, whether customer environments are secure and compliant by default, and whether the platform scales from tens to hundreds of enterprise deployments. Release velocity is a company-level strategic priority, and this role is at the center of it.
WhatWe Are Not Looking For
To save everyone's time: this role is not for traditional IT operations. If your background is primarily in manual server provisioning, ticketing-system-driven change management, desktop support, or on-prem datacenter administration, if you always deploy someone else’s code via IaC, this is not the right fit. We need someone whose default mode is writing code to solve infrastructure problems.
Required Experience- 7+ years in Dev Ops, Cloud Engineering, or Platform Engineering roles, with at least 2 years in a senior or lead capacity
- Deep, daily-driver coding experience:,programmatically managing infrastructure through Python, APIs and IaC tools is second nature to you. The web console is an afterthought.
- Strong production AWS experience: compute (EKS, ECS, EC2), networking (VPC, Transit Gateway, ALB/NLB, Route
53), storage (S3, EBS, EFS), security (IAM, KMS, Security Hub, Guard Duty) - Designed and built CI/CD pipeline infrastructure (not just consumed existing pipelines). Git Hub Actions, Git Lab CI, or Jenkins at scale.
- Container orchestration: ECS, Docker, Kubernetes (EKS preferred), service mesh concepts
- Scripting and automation:
Python or Go. Bash only is not enough - Git-based…
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).