×
Register Here to Apply for Jobs or Post Jobs. X

Site Reliability Engineer

Job in Mountain View, Santa Clara County, California, 94039, USA
Listing for: EarnIn
Full Time, Part Time position
Listed on 2026-07-17
Job specializations:
  • IT/Tech
    SRE/Site Reliability
Salary/Wage Range or Industry Benchmark: 189000 - 232000 USD Yearly USD 189000.00 232000.00 YEAR
Job Description & How to Apply Below

About Earn In

As one of the first pioneers of earned wage access, Earn In delivers real-time financial flexibility for those living paycheck to paycheck. Members access their earnings as they earn them, with options to spend, save, and grow money without mandatory fees, interest rates, or credit checks.

Earn In has an experienced leadership team and funding partners including A16Z, Matrix Partners, DST, Ribbit Capital. We are growing fast and are excited to continue bringing world-class talent onboard to help shape the next chapter of our growth journey.

WHY this role exists

Earn In’s community members rely on our products to deliver reliability and trust when they need them most. Reliability shapes the product experience; every noisy alert, unclear runbook, fragile deployment, or repeated incident undermines customer trust and hinders engineering teams.

This role enables Earn In to build and run production systems with greater resilience, clarity, and confidence. As a Site Reliability Engineer, you will strengthen infrastructure, optimize tooling, deepen observability, streamline incident response, and elevate reliability standards. These actions empower teams to ship quickly and safely.

The base salary range for this full-time position is $189,000 - $232,000 plus equity and benefits. Our salary ranges are determined by role, level, and location. This is a hybrid position in Mountain View (Headquarters) and will require in-office work 2 days a week.

HOW you will create impact
  • You will operate as a well-rounded SRE practitioner across production operations, observability, incident response, infrastructure-as-code, automation, and software engineering.
  • You will demonstrate growing independence in reliability work. You will not only follow existing playbooks, but also refine them. You will transform production learnings into better alerts, clearer runbooks, safer deployments, stronger observability, and more reliable services.
  • You will harness AI-assisted development and operational workflows to minimize toil, accelerate investigation, enhance documentation, and streamline infrastructure and reliability work. You will meticulously validate AI-generated output before applying it to production systems or operational workflows.
  • You will collaborate with product engineering and platform teams to implement, explain, and support reliability practices, ensuring they are practical, understandable, and actionable.
WHAT you'll do
  • Design and improve systems with resilience and graceful degradation in mind. Plan for capacity and possible failure modes.
  • Define and measure SLOs and SLIs that reflect customer experience and help teams make better reliability tradeoffs.
  • Use observability tools such as Datadog, Cloud Watch, logs, metrics, traces, and APM. Build signal-heavy, noise-light visibility into production systems.
  • Configure and improve alerting and routing through incident management workflows. Make sure pages are actionable, well-routed, and worth human attention.
  • Participate in incident response from detection and triage through communication, resolution, postmortems, and follow-up.
  • Continuously improve the incident lifecycle. Focus on better detection, clearer runbooks, stronger postmortems, and concrete remediations.
  • Construct or optimize infrastructure, reliability tooling, and automation that eliminate toil and ensure operational consistency.
  • Use AI-assisted tools to accelerate coding and documentation. Speed up root-cause exploration, runbook improvement, infrastructure-as-code workflows, and operational tasks.
  • Help engineering teams improve production readiness and deployment safety. Support service ownership and operational clarity.
  • Communicate reliability concepts clearly across technical and non-technical teams.
  • Document operational knowledge to reduce silos. Make it easier for engineers to respond with confidence.
  • Contribute to a culture where reliability is shared by SRE and product engineering teams.
WHAT we're looking for
  • Bachelor’s or master’s degree in Computer Science, Engineering, or a related field, or equivalent industry experience.
  • 3+ years of experience in SRE, Software Engineering, Infrastructure…
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary