×
Register Here to Apply for Jobs or Post Jobs. X
More jobs:

Distinguished Site Reliability Engineer - Cloud

Job in Vancouver, Clark County, Washington, 98662, USA
Listing for: Segment (Twilio)
Full Time position
Listed on 2026-07-16
Job specializations:
  • Software Development
    Unix/Linux
Salary/Wage Range or Industry Benchmark: 320000 - 488750 USD Yearly USD 320000.00 488750.00 YEAR
Job Description & How to Apply Below

Overview

Site Reliability Engineering (SRE) at NVIDIA is an engineering discipline to design, build and maintain large‑scale production systems with high efficiency and availability using software and systems engineering practices. It requires knowledge across systems, networking, coding, database, capacity management, continuous delivery and deployment, and open source cloud enabling technologies such as Kubernetes and Open Stack. SRE ensures internal and external GPU cloud services run with maximum reliability and uptime while enabling developers to make changes safely, focusing on capacity, latency, and performance.

Responsibilities
  • Lead, design, implement and support operational and reliability aspects of large‑scale Kubernetes clusters, focusing on performance at scale, real‑time monitoring, logging, and alerting.
  • Engage in and improve the whole lifecycle of services—from inception and design through deployment, operation and refinement.
  • Support services before they go live through activities such as system design consulting, developing software tools, platforms and frameworks, capacity management and launch reviews.
  • Maintain services once they are live by measuring and monitoring availability, latency and overall system health.
  • Scale systems sustainably through mechanisms like automation, and evolve systems by pushing for changes that improve reliability and velocity.
  • Practice sustainable incident response and blameless postmortems.
  • Participate in an on‑call rotation to support production systems.
Qualifications
  • BS degree in Computer Science or a related technical field involving coding (e.g., physics or mathematics), or equivalent experience.
  • 16+ years of experience with infrastructure automation, distributed systems design, and experience designing and developing tools for operating large‑scale private or public cloud systems in production.
  • Experience in one or more of the following:
    Python, Go, Perl, or Ruby.
  • In‑depth knowledge of Linux, networking and containers.
Ways to Stand Out
  • Interest in crafting, analyzing and fixing large‑scale distributed systems.
  • Systematic problem‑solving approach, coupled with strong communication skills and a sense of ownership and drive.
  • Ability to debug and optimize code and automate routine tasks.
  • Experience in using or running large private and public cloud systems based on Kubernetes, Open Stack and Docker.
Compensation

Base salary will be determined based on location, experience, and market comparable positions. The base salary range is 320,000

USD–488,750

USD. You will also be eligible for equity and benefits.

Application Process

Applications will be accepted at least until July
14,2026.

Equal Opportunity Employer

NVIDIA is committed to fostering an inclusive work environment and proud to be an equal‑opportunity employer. We do not discriminate on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

#J-18808-Ljbffr
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary