Site Reliability Engineer- FedRAMP
Listed on 2026-09-14
-
IT/Tech
About The Team
The Rubrik Engineering team is comprised of people who produce extraordinary results. Our engineers are driven to build efficient, reliable, and cost effective products. We believe in empowering our teams, giving engineers autonomy and responsibility, not just tasks. Our goal is to motivate and challenge you to do the best work of your career. As part of the Rubrik Engineering team, you will work closely with product managers, designers, and other engineers to define the next generation of products for Rubrik.
At Rubrik, nothing will stop you from thinking big. We are looking for individuals who are comfortable with ambiguity and excited by the prospect of a challenge. If you have a positive attitude, high energy and limitless drive, and like to win, we want to talk to you!
Site Reliability Engineers at Rubrik are systems/software engineers who ensure that Rubrik's infrastructure services run smoothly and have the capacity for future growth.
What you’ll do:- Ensure we maintain high availability and durability of our databases
- Establish best practices for internal teams to write performant SQL queries
- Perform periodic database upgrades minimizing downtime for our customers
- Design, implement and maintain relational database systems for performance and reliability
- Manage and run backend systems like Kubernetes, MySQL and everything in between
- Drive reliability, availability and efficiency improvements to Rubrik's Polaris Cloud Platform
- Good mix of software and system engineering skills
- Participate on-call rotations across continents, using a follow-the-sun model
- Write and review code, plan and execute upgrades, develop documentation and capacity plans, and debug production issues
- Work cross-functionally with various engineering teams
- Build monitoring tools and automation to increase efficiency of all teams
- Good written and verbal communication skills
- Drive Fed Ramp Certification process
- Good written and verbal communication skills
- 2+ years of experience designing and managing relational databases with a focus on performance, scalability, reliability, high-availability, and disaster recovery.
- Experience in database design and architecture supporting large enterprise customers with high SLO and SLA requirements
- Experience operating database layer of a large scale SaaS product
- Experience in one or more of the following:
Golang, Python, Java, Scala, C++ - Systematic problem-solving approach, coupled with strong communication skills and a sense of ownership and drive
- Expertise in designing, analyzing and troubleshooting large-scale distributed systems
- Ability to debug and optimize code and automate routine tasks
- Strong operational experience with Unix/Linux operating systems and networking
- Experience with Google Cloud Platform or other public cloud technologies
- Minimum 1-3 years of experience as a Development, Dev Ops or Site Reliability Engineer Willing to provide 24/7 coverage
- Strong Documentation skills
- Experience working with multiple departments and divisions within an organization
- Strong understanding of Databases is a definite plus
- Experience leading support personnel
- Experience with FedRAMP certification is strongly desired
- U.S. citizenship at the time of hire.
- Residence within the contiguous United States (i.e., the lower 48 states and the District of Columbia); and
- Willingness to undergo a Single Source Background Investigation if required.
This position carries special Security and Privacy Responsibilities for protecting the U.S. Federal Government’s interests:
- Know, acknowledge, and follow system-specific security policies and procedures;
- Protect data and individual privacy per requirements and…
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).