×
Register Here to Apply for Jobs or Post Jobs. X

Site Reliability Engineer II

Job in Atlanta, Fulton County, Georgia, 30383, USA
Listing for: Talanto
Full Time position
Listed on 2026-09-04
Job specializations:
  • IT/Tech
    Cloud Computing: Infrastructure & Operations, SRE/Site Reliability, Systems Engineer
Salary/Wage Range or Industry Benchmark: 113000 - 171600 USD Yearly USD 113000.00 171600.00 YEAR
Job Description & How to Apply Below

Important: if an employer asks you to log into their system via iCloud or Google, send a code, an SMS or Telegram password, run some code, or install software — refuse. These are signs of fraud.

••••••••• (NYSE:PD) is a leader in Digital Operations Management. In an always-on world, organizations of all sizes trust ••••••••
• to help them deliver a perfect digital experience to their customers, every time. Teams use ••••••••
• to identify issues and opportunities in real time and bring together the right people to fix problems faster and prevent them in the future. Over 13,000 organizations (including 60 of Fortune 100) rely on ••••••••
• to succeed with Digital Transformation, Cloud Migration, and Dev Ops Modernization. Notable customers include GE, Cisco, Genentech, Electronic Arts, Cox Automotive, Netflix, Shopify, Zoom, Door Dash, Lululemon and more. We are expanding rapidly as a platform for Digital Operations Management using AI/ML and Automation and growing our adoption by Development, IT, Customer Service, Security, and other teams across the organization.

As a Site Reliability Engineer II on the Core Infrastructure team in our Atlanta office, you'll help build and operate the foundational infrastructure that powers ••••••••• 's real-time digital operations platform. Our systems support millions of events and alerts daily, enabling customers to detect, respond to, and resolve incidents quickly and reliably. You'll work at the intersection of platform evolution and operational excellence, building and evolving foundational network, compute, and ingress infrastructure while scaling and hardening existing systems.

Your work will directly impact the reliability, scalability, and security of the services our customers rely on to keep their businesses running as ••••••••
• continues to grow across products, regions, and customer use cases.

Key Responsibilities
  • Support and improve foundational infrastructure, including networking, compute platforms, Kubernetes clusters, and ingress/traffic management systems.
  • Contribute to the reliability and scalability of ••••••••• 's core platform by hardening existing systems and supporting the rollout of new infrastructure capabilities.
  • Participate in agile rituals (standups, planning, retros) and communicate progress/risks early.
  • You stay current on technical trends to suggest innovative tools and approaches to interesting problems.
  • Monitor system health using metrics, logs, and alerts, and participate in 24/7 on‑call rotations to help detect, respond to, and resolve incidents.
Basic Qualifications
  • 3+ years of experience in Site Reliability Engineering, Dev Ops, or Platform Engineering roles
  • Hands-on experience operating Linux-based systems in production environments
  • Working knowledge of networking fundamentals, such as load balancing, DNS, TLS, and ingress traffic flow
  • Experience with container orchestration (e.g., EKS, Kubernetes)
  • Experience working on cloud-native infrastructure (e.g., AWS, GCP, Azure), including networking and compute concepts
  • Proficiency in at least one programming language (e.g., Python, Ruby, Go, etc.)
  • Experience with Infrastructure as Code (e.g., Terraform, Cloud Formation)
Preferred Qualifications
  • Experience with AWS cloud networking concepts such as VPCs, subnets, routing, security groups, and load balancers
  • Experience operating or contributing to production Kubernetes platforms (e.g., EKS), including cluster upgrades, networking, or ingress configuration
  • Experience with monitoring, observability, and logging platforms (e.g., Data Dog, New Relic, Sumo Logic, Splunk, Prometheus, Grafana)
  • Familiarity with service meshes, ingress controllers, or API gateways (e.g., Envoy, Istio, NGINX)

Salary Range: $113,000 to $171,600

Where we work

••••••••
• operates a hybrid work model with offices in 8 major cities:
Atlanta, Lisbon, London, San Francisco, Santiago, Sydney, Tokyo, and Toronto. While we offer flexibility within our established locations, we cannot employ candidates residing in:
Location restrictions:
Australia: Northern Territory, Queensland, South Australia, Tasmania, Western Australia Canada: Alberta, Manitoba, Newfoundland, Northwest…

To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary