Senior AI Cloud SRE | Kubernetes & Observability
Job in
Zürich, 8081, Zurich, Kanton Zürich, Switzerland
Listed on 2026-08-20
Listing for:
NVIDIA
Full Time
position Listed on 2026-08-20
Job specializations:
-
IT/Tech
Cloud Computing: Infrastructure & Operations, SRE/Site Reliability
Job Description & How to Apply Below
NVIDIA in Zürich seeks an experienced Senior Site Reliability Engineer to maintain and scale DGX Cloud clusters for AI researchers and enterprise clients worldwide.
You will define SLOs, build observability stacks, automate infrastructure, and lead on-call incidents across multiple clouds, ensuring high reliability and velocity.
The role requires 10+ years of production experience, deep Kubernetes expertise, and strong programming skills in Python or Go.
#J-18808-LjbffrPosition Requirements
10+ Years
work experience
Note that applications are not being accepted from your jurisdiction for this job currently via this jobsite. Candidate preferences are the decision of the Employer or Recruiting Agent, and are controlled by them alone.
To Search, View & Apply for jobs on this site that accept applications from your location or country, tap here to make a Search:
To Search, View & Apply for jobs on this site that accept applications from your location or country, tap here to make a Search:
Search for further Jobs Here:
×