×
Register Here to Apply for Jobs or Post Jobs. X

DevOps Team Lead

Job in Madison, Dane County, Wisconsin, 53774, USA
Listing for: VAS
Full Time position
Listed on 2026-08-29
Job specializations:
  • IT/Tech
    AWS, SRE/Site Reliability, Cloud Computing: Infrastructure & Operations
Salary/Wage Range or Industry Benchmark: 140000 - 190000 USD Yearly USD 140000.00 190000.00 YEAR
Job Description & How to Apply Below

Job Description

The Dev Ops Team Lead owns the delivery, reliability, and day-to-day operation of our AWS cloud platform, an extensive multi-account cloud environment managed as code and expanding toward globally distributed, multi-region workloads.

Job Description

The Dev Ops Team Lead owns the delivery, reliability, and day-to-day operation of our AWS cloud platform, an extensive multi-account cloud environment managed as code and expanding toward globally distributed, multi-region workloads. This is a highly hands‑on technical leadership role. You'll be one of the strongest technical contributors on the team while also helping shape and prioritize the work. You'll translate product and engineering needs into infrastructure requirements, guide the team's delivery, and ensure what we build is automated, secure, documented, cost-conscious, and operationally reliable.

The role is approximately 60% hands‑on engineering and 40% technical leadership, including requirements, planning, review, coordination, and coaching.

What You'll Do
  • Design, build, review, and operate our Infrastructure-as-Code estate across a multi-account AWS environment, including reusable modules, environment configurations, state management, and account provisioning.
  • Own and improve CI/CD pipelines for infrastructure changes, including automated validation, security and compliance controls, and keyless cloud authentication.
  • Build and maintain the core AWS services supporting our products, including networking, containers, databases, storage, messaging, systems management, backup, monitoring, and cost controls.
  • Establish and reinforce practical engineering standards around IaC, IAM, resource tagging, state management, versioning, security, documentation, and operational readiness.
  • Translate ambiguous requests from product, engineering, security, and data teams into clear requirements, milestones, estimates, dependencies, and acceptance criteria.
  • Own and refine the Dev Ops backlog and facilitate the team's Scrum cadence, including planning, stand‑ups, refinement, reviews, and retrospectives.
  • Strengthen platform reliability through observability, SLOs, actionable alerting, runbooks, disaster recovery, backup and restore practices, and continuous reduction of operational toil.
  • Participate in a rotating night and weekend on‑call schedule and provide technical leadership during platform incidents, including incident coordination and post‑incident follow‑through. Rotation frequency will depend on team size and coverage needs but is shared equitably across the team.
  • Partner with security to embed identity, least‑privilege access, network segmentation, encryption, secrets management, audit logging, and compliance requirements into the platform.
  • Lead cloud cost visibility and optimization, including tagging, right‑sizing, commitment planning, anomaly response, and the cost implications of multi‑region architecture.
  • Help evolve the platform toward globally distributed workloads, including regional rollout, data residency and replication, latency‑aware routing, cross‑region failover, and associated operational tradeoffs.
  • Use AI tools effectively in engineering workflows and help design and operate the cloud infrastructure needed to support emerging AI capabilities.
  • Set technical direction in partnership with architecture and coach engineers through code review, pairing, design discussions, documentation, and shared platform ownership.
What You'll Bring
  • Extensive experience in software engineering, infrastructure, or Dev Ops, with demonstrated experience operating at a senior or technical lead level in complex cloud environments.
  • Expert‑level, current hands‑on experience with Terraform or equivalent Infrastructure as Code, including managing large‑scale cloud environments through IaC.
  • Experience with IaC at scale, including reusable module design and versioning, state architecture, safe refactoring, upgrades, drift detection and remediation, and bringing legacy infrastructure under IaC management.
  • Deep hands‑on AWS experience across compute, storage, networking, IAM, and security, ideally within a multi‑account AWS Organization.
  • Proven ownership of…
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary