Site Reliability Engineer
Listed on 2026-09-12
-
IT/Tech
Cloud Computing: Infrastructure & Operations, SRE/Site Reliability, Network Engineer, Systems Engineer
Senior Site Reliability Engineer (SRE)
Location:
San Francisco, CAWork Model:
Onsite Industry: Renewable Energy Comp: $200,000 - $240,000
We’re partnering with a fast-growing energy technology company looking for a Senior Site Reliability Engineer to take ownership of critical production infrastructure spanning cloud, networking, and physical environments. This is not a traditional SaaS SRE role. You’ll work across AWS infrastructure and real-world systems supporting mission-critical energy operations, including infrastructure deployed in data centers and at commercial and industrial sites.
What You’ll Do- Own and improve highly available production infrastructure across AWS and physical environments.
- Build, manage, and automate infrastructure using Terraform and Infrastructure as Code best practices.
- Design and troubleshoot complex networking across cloud, data center, and on-prem environments.
- Help lead the migration and productionization of critical infrastructure into a colocation environment.
- Improve reliability, monitoring, observability, incident response, and operational processes as the platform scales.
- Support infrastructure that connects cloud services with locally deployed hardware and edge systems.
- Work closely with software and engineering teams to ensure reliable deployment and operation of production services.
- Occasionally travel to customer sites to support infrastructure deployments and troubleshoot real-world systems.
- Strong experience owning production infrastructure in AWS.
- Hands-on expertise with Terraform and Infrastructure as Code.
- Deep understanding of networking fundamentals and experience troubleshooting complex network environments.
- Experience working with Linux-based production systems.
- Comfortable operating across both modern cloud infrastructure and legacy or on-prem environments.
- Strong troubleshooting skills and the ability to work independently in high-impact production environments.
- Experience supporting mission-critical, highly available, or regulated systems is highly valuable.
- Experience with colocation facilities or data center infrastructure.
- Experience supporting edge computing, industrial systems, energy infrastructure, or other environments where software interacts with physical systems.
- Exposure to Rust or experience supporting Rust-based production services.
- Background in defense, aerospace, financial infrastructure, utilities, industrial technology, or similarly high-reliability environments.
You’ll join a small, rapidly scaling engineering organization where infrastructure is central to the product. You’ll have significant ownership from day one, including the opportunity to shape how critical systems are deployed, operated, monitored, and scaled as the company grows.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).