×
Register Here to Apply for Jobs or Post Jobs. X

Senior Site Reliability Engineer

Job in Santa Barbara, Santa Barbara County, California, 93190, USA
Listing for: Umbra
Full Time position
Listed on 2026-08-06
Job specializations:
  • IT/Tech
    Systems Engineer, Cloud Computing: Infrastructure & Operations, SRE/Site Reliability
Salary/Wage Range or Industry Benchmark: 150000 - 180000 USD Yearly USD 150000.00 180000.00 YEAR
Job Description & How to Apply Below

Umbra is an American space technology company delivering advanced systems, from sensors to spacecraft, that empower customers worldwide with unmatched access to critical information from space. Our mission is simple and ambitious: redefine space—for people, systems, and missions in every domain. Umbra’s ecosystem operates through three business units:
Remote Sensing (the data), Space Systems (the components), and Mission Solutions (the platforms).Together, our teams develop capabilities that deliver persistent access, resilient performance, and mission-ready solutions, advancing U.S. space leadership while keeping the world safe and informed.

About the Team

Remote Sensing – The Data

Remote Sensing is where Umbra got its start, and our agile Synthetic Aperture Radar (SAR) constellation remains the most capable on the market. We transform satellite data into real-world, actionable insights that strengthen U.S. national security and intelligence, support disaster response, and advance scientific discovery. Our team delivers data at scale with unmatched quality, persistence, and the speed and responsiveness our partners demand.

If you want to work on cutting-edge space technology that’s redefining what’s possible in remote sensing, you belong here at Umbra.

About the Job

We are seeking an experienced Senior
Site Reliability Engineer to help design, build, operate, and scale the mission- and business-critical infrastructure that powers Umbra's systems. In this role, you will leverage a deep understanding of modern infrastructure, distributed systems, and the broader technology stack to drive technical excellence, make thoughtful architectural decisions, and balance long-term scalability with operational reliability.

You'll partner closely with engineering teams to improve processes, champion best practices, evaluate emerging technologies, and implement solutions that enhance the performance, resilience, and efficiency of our platforms. The ideal candidate is a collaborative technical leader who communicates effectively across technical and non-technical teams and drives meaningful improvements that have a lasting impact across the organization.

This position is based on-site in either our Arlington, VA office, Reston, VA office or Santa Barbara/Goleta, CA office.

Key Responsibilities

  • Ensure the reliability and scalability of critical systems, meeting SLAs through proactive monitoring and effective incident response.
  • Develop and promote new technologies and tools, conducting research and creating proofs of concept to introduce solutions that enhance the team's capabilities.
  • Lead by example in fostering a culture of excellence and reliability.
  • Continuously evaluate and improve team processes and workflows to increase efficiency and reduce complexity.
  • Collaborate closely with cross-functional teams, product managers, and stakeholders to align on technical strategy and provide expert guidance.
  • Participate in on-call rotations, providing support and resolving complex technical issues.
Required Qualifications
  • Bachelor’s degree in Computer Science or a related technical field.
  • 5-8+ years in a Site Reliability Engineer or Dev Ops role supporting a SaaS platform, with demonstrated expertise managing distributed systems.
  • Extensive experience with AWS services (EC2, S3, Lambda, VPC Networking) and deep knowledge of cloud infrastructure, networking, and security best practices.
  • Proficiency running, optimizing, and scaling Kubernetes clusters in production environments.
  • Experience using and writing Terraform to architect and manage production infrastructure.
  • Ability to create and utilize Infrastructure-as-code (IaC), Git Ops practices, and automation tools to increase reliability and reduce manual tasks.
  • Proven success in leading teams or projects using Agile/Scrum methodologies.
  • Expertise in infrastructure and software architecture, capable of designing and implementing large-scale, reliable systems with minimal guidance.
  • Experience developing and managing comprehensive infrastructure monitoring and alerting strategies.
Desired Qualifications
  • 10+ years in a Site Reliability Engineer or Dev Ops role supporting a SaaS platform, with…
Position Requirements
10+ Years work experience
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary