×
Register Here to Apply for Jobs or Post Jobs. X

Sr. Cloud Resilience Engineer - Security

Job in Atlanta, Fulton County, Georgia, 30383, USA
Listing for: UKG
Full Time position
Listed on 2026-07-22
Job specializations:
  • IT/Tech
    Cloud Computing: Infrastructure & Operations, SRE/Site Reliability
Salary/Wage Range or Industry Benchmark: 140000 - 190000 USD Yearly USD 140000.00 190000.00 YEAR
Job Description & How to Apply Below
Position: Sr. Staff Cloud Resilience Engineer - Security

Architectural Leadership & Consulting

Act as the primary resilience advisor to multiple distributed product and enterprise teams, guiding them on best practices for building high availability (HA) and redundancy into their SaaS applications.

Resilient Cloud Design

Design and recommend fast-failover solutions and highly available infrastructure primarily on Google Cloud Platform (GCP), while also providing oversight for workloads in Azure and AWS.

Infrastructure Validation

Leverage your strong background in Infrastructure as Code (IaC) to review, validate, and guide the implementation efforts of engineering teams.

Container & Compute Resilience

Design redundancy strategies for workloads running on Google Kubernetes Engine (GKE) and virtual machines, ensuring self‑healing deployments.

Cross‑Functional Collaboration

Partner closely with Dev Ops, SRE, and Product Engineering teams to champion resilience engineering principles, chaos testing, and failover validations across tier‑0 mission‑critical systems.

Cloud Platform Expertise

Deep, practical technical knowledge of Google Cloud Platform (GCP) core services, specifically GKE, Compute Engine, and CloudSQL. Familiarity with AWS and Azure is highly desirable.

Technical Practitioner Background

Proven past experience as a hands‑on engineer who has deployed complex infrastructure. You should understand the implementation details well enough to effectively guide engineering teams.

High Availability Architecture

Demonstrated success in architecting active‑active or active‑passive fast‑failover mechanisms for high‑volume, data‑intensive SaaS applications.

Database Resilience

Strong understanding of database clustering, replication, and migration strategies (especially migrating legacy RDBMS like MS SQL Server to cloud‑native solutions like CloudSQL).

Advisory Skills

Excellent communication and consulting skills, with the ability to influence technical teams, explain complex architectural concepts, and foster a culture of resilience without having direct reporting authority over the engineering teams.

Resilient Network Services

Practical design experience managing high‑availability network topologies, including load balancing, DNS & name resolution, firewalls/gateways, identity/authentication systems, and centralized logging/SIEM.

User Session Management

Deep understanding of user session replication, session state persistence, and failover routing strategies in high‑traffic, multi‑region application architectures.

Broad HA Domain Exposure

Familiarity assessing or designing resilience across a comprehensive range of critical SaaS failure domains, such as API gateways, caching layers, messaging/queuing systems, and CI/CD pipelines.

#J-18808-Ljbffr
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary