Senior Site Reliability Engineer
Job in
Raleigh, Wake County, North Carolina, 27601, USA
Listed on 2026-07-19
Listing for:
Red Hat, Inc.
Full Time
position Listed on 2026-07-19
Job specializations:
-
IT/Tech
SRE/Site Reliability, Cloud Computing: Infrastructure & Operations, Systems Engineer
Job Description & How to Apply Below
Hybrid locations:
Raleigh time type:
Full time posted on:
Posted Todayjob requisition :
R-056327
** Job Summary
** The Red Hat IT Open Shift team is looking for a Senior Site Reliability Engineer (SRE) to design, develop, scale, and operate our Red Hat Hybrid Open Shift Platforms (on-prem & cloud). As a Senior Engineer, you will contribute to running Red Hat Open Shift at scale by enabling customer self-service, making our monitoring system more sustainable, and eliminating toil through the IT Open Shift team you will have the opportunity to influence the complex challenges of scale which are unique to Red Hat IT managed cloud platform services, while using your skills in coding, operations, and large-scale distributed system design.
We develop, deploy, and maintain Red Hat's next-generation mission critical platform across hybrid cloud infrastructures. We are a global team operating on-premise and in the public cloud, using the latest technologies from Red Hat and beyond. Red Hat relies on teamwork and openness for its success. We learn from our failures in a blameless environment to support the continuous improvement of the team.
At Red Hat, your individual contributions have more visibility than most large companies, and visibility means career opportunities and growth. Successful applicants must reside in a state where Red Hat is registered to do business.
At Red Hat, our commitment to open source innovation extends beyond our products - it’s embedded in how we work and grow. Red Hatters embrace change – especially in our fast-moving technological landscape – and have a strong growth mindset. That's why we encourage our teams to proactively, thoughtfully, and ethically use AI to simplify their workflows, cut complexity, and boost efficiency.
This empowers our associates to focus on higher-impact work, creating smart, more innovative solutions that solve our customers' most pressing challenges.
** What You Will Do
*** Design, build, and manage our large scale infrastructure and platform services, including public cloud, private cloud, and datacenter-based
* Automate cloud infrastructure through use of technologies (e.g. auto scaling, load balancing, etc.), scripting (python and golang), monitoring and alerting solutions (e.g. Splunk, Splunk IM, Prometheus, Grafana, Catchpoint, Data Dog etc)
* Design, develop, and become expert in IT’s Red Hat Open Shift offerings by leveraging emerging industry standards
* Build & support standardized CI/CD platform components using Open Shift Pipelines and Tekton, Git Lab to enable multiple application deployments
* Apply Infrastructure as Code methodologies using Git Ops practices with ArgoCD for declarative platform management
* Breakdown complex engineering efforts into consumable chunks while working with teams to understand deliverables
* Design and development of software like Kubernetes operators, webhooks, cli-tools
* Implement and maintain intelligent infrastructure and application monitoring designed to enable application engineering teams
* Ensure the production environment is operating in accordance with established procedures and best practices
* Lead escalation support for high severity and critical platform-impacting events
* Provide feedback around bugs and feature improvements to the various Red Hat Product Engineering teams
* Design software tests and lead peer reviews to increase the quality of our codebase
* Help and develop peers’ capabilities through knowledge sharing, mentoring, and collaboration
* Participate in a regular on-call schedule, supporting the operation needs of our tenants
* Drive sustainable incident response and lead blameless postmortems
* Work within a small agile team to develop and improve SRE methodologies, support your peers, plan and self-improve
** What You Will Bring
*** 5+ years of experience operating production services on Kubernetes/Open Shift
* 3+ years of programming experience in Python, Go
* 2+ years of experience of using cloud providers and technologies (Google, Azure, Amazon, etc.)
* Hands-on experience with Kubernetes/Open Shift, Linux AWS
* Experience with Git Ops…
Position Requirements
10+ Years
work experience
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
Search for further Jobs Here:
×