Site Reliability Engineer
Listed on 2025-12-02
-
Software Development
Cloud Engineer - Software, DevOps
Red Hat is looking for a Platform Engineer to join its Platform Engineering team! In this role, you will help architect, implement, improve, and support the Open Shift-based platform that runs many of Red Hat’s most important multi-tenant Software-as-a-Service (SaaS) and Managed-service offerings. Using your expertise in SRE principles, you will help create an environment where reliability, scalability, and security come first, and are not treated as an afterthought.
In this role, you will spend a portion of your time working across teams to define and iterate upon processes for onboarding new managed services at Red Hat and demonstrate good judgment in employing onboarding methods and techniques that can be repeated and iterated upon. You will also contribute to the codebase of command-and-control software that automates the building, deployment, monitoring, and alerting of Red Hat managed services.
The remainder will be spent on various other tasks, such as diagnosing issues, planning, documenting, and mentoring.
Design, write, and maintain software (primarily in Python and Golang) that automates the deployment, monitoring, and maintenance of Red Hat managed services.
Onboarding of new services onto our Open Shift-based platform
Adhering to cloud-native design principles & best practices to ensure reliability, scalability, and security
Contribute to documents, like standard operating procedures (SOPs) and playbooks, that assist in issue resolution and new-service onboarding.
Proactively utilize AI-assisted development tools (e.g., Git Hub Copilot, Cursor, Claude Code) for code generation, auto-completion, and intelligent suggestions to accelerate development cycles and enhance code quality.
Participate in an Agile Scrum team that scopes, prioritizes, and allocates work items.
Participate in an on-call rotation that is responsible for responding to service incidents.
Background writing object-oriented automation software in Python, experience with Golang is only plus
Background administering production cloud-native services, preferably containerized and deployed via a container-orchestration system like Kubernetes or Open Shift
Experience diagnosing service failures and carrying out incident response procedures
Familiarity with Linux operating system and its configuration
Ability to effectively work in a globally distributed team
Understanding of computer networking and protocols, including TCP/IP and DNS
Understanding of computer security and cryptography basics, including certificates, TLS, and credential-storage systems like Vault is a plus
Familiarity with CI/CD pipeline concepts and systems, like Jenkins and Tekton/Argo is a plus
Familiarity with observability tools like Prometheus and Grafana, and how to define metrics that can be used to measure service health and reliability is a plus
The salary range for this position is $94,550.00 - $. Actual offer will be based on your qualifications.
Pay TransparencyRed Hat determines compensation based on several factors including but not limited to job location, experience, applicable skills and training, external market value, and internal pay equity. Annual salary is one component of Red Hat’s compensation package. This position may also be eligible for bonus, commission, and/or equity. For positions with Remote-US locations, the actual salary range for the position may differ based on location but will be commensurate with job duties and relevant work experience.
AboutRed Hat
Red Hat () is the world’s leading provider of enterprise open source () software solutions, using a community-powered approach to deliver high-performing Linux, cloud, container, and Kubernetes technologies. Spread across 40+ countries, our associates work flexibly across work environments, from in-office, to office-flex, to fully remote, depending on the requirements of their role. Red Hatters are encouraged to bring their best ideas, no matter their title or tenure.
We're a leader in open source because of our open and inclusive environment. We hire creative, passionate people ready to contribute their ideas, help solve complex problems, and make an…
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).