×
Register Here to Apply for Jobs or Post Jobs. X

Senior Site Reliability Engineer

Job in El Segundo, Los Angeles County, California, 90245, USA
Listing for: GoGuardian
Full Time position
Listed on 2026-08-11
Job specializations:
  • IT/Tech
    Cloud Computing: Infrastructure & Operations, SRE/Site Reliability, Systems Engineer
Salary/Wage Range or Industry Benchmark: 140000 - 190000 USD Yearly USD 140000.00 190000.00 YEAR
Job Description & How to Apply Below

We're looking for a Senior Site Reliability Engineer (SRE) to help design, scale, and maintain the infrastructure that powers our core products and services. In this role, you'll collaborate with engineering teams to drive operational excellence, optimise system performance, and ensure high availability across production environments. This position sits on Tech Foundation, a team that manages core cloud infrastructure, shared data services, and developer tooling to empower our product teams to deliver software efficiently and securely.

The ideal candidate brings a strong background in cloud infrastructure, automation, and modern reliability practices, with a passion for solving complex operational challenges in a collaborative environment.

Responsibilities:

  • Architect and maintain scalable, secure cloud infrastructure to ensure high availability for core products.
  • Enhance observability and monitoring frameworks to deliver highly accurate alerts, minimising noise and improving incident detection.
  • Participate in on-call rotations and lead incident response, ensuring comprehensive post-mortems and RCAs are completed to drive systemic improvements.
  • Optimise and modernise deployment pipelines and automation workflows to maximise engineering velocity and operational safety.
  • Partner with product development teams to provide infrastructure support, review architectural changes, and promote reliability best practices.
  • Implement and uphold robust security standards and compliance controls across all managed cloud infrastructure.

Requirements:

  • 5+ years of professional experience in Site Reliability Engineering, Infrastructure, or Dev Ops roles supporting production SaaS applications.
  • Strong proficiency with AWS core services (including EC2 VPC, S3) along with experience in Serverless frameworks and managed Kubernetes environments like EKS.
  • Extensive experience writing and managing Infrastructure as Code (IaC) using Terraform.
  • Familiarity with configuring, troubleshooting, and maintaining data layers such as MongoDB, Redshift, and Open Search.
  • Experience with GCP environments or technologies like Firestore is a plus.
  • Experience managing or modernising CI/CD pipelines and deployment workflows utilising systems like Jenkins, AWS Code Build/Code Pipeline, or Git Hub Actions.
  • Deep understanding of Linux operating system fundamentals and Unix shell scripting.
  • Ability to read and debug code written in JavaScript/Type Script, Python, or Go to effectively troubleshoot underlying service errors.
  • Strong communication and collaboration skills, with a track record of driving technical decisions and establishing team-wide operational standards.
  • Eager to take initiative in a fast-paced, ever-changing, dynamic environment.
  • Fueled by the opportunity to truly impact the education landscape.
#J-18808-Ljbffr
Position Requirements
10+ Years work experience
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary