Senior/Site Reliability Engineer
Listed on 2026-07-01
-
IT/Tech
Systems Engineer, Cloud Computing: Infrastructure & Operations, SRE/Site Reliability, Network Engineer
Senior/Staff Site Reliability Engineer
Mountain View, CA
About the RoleWe are seeking an experienced Senior/Staff Site Reliability Engineer to support the operation, monitoring, and scaling of our growing fleet of autonomous vehicles. In this role, you will work closely with our infrastructure and platform teams to manage rollouts of both on-premises and cloud infrastructure in support of expansions to new customer sites. You will be directly involved in the setup and monitoring of our data offload systems, remote supervision stations, and on-prem continuous integration (CI) environments, ensuring our infrastructure is highly reliable, secure, and optimized for performance.
This position plays a critical role in keeping our autonomy operations running smoothly while supporting the rapid growth of our fleet and customer base. This role is onsite 5 days a week at our Mountain View, CA office!
- Upgrade and maintain both physical and cloud infrastructure used for offloading data from our autonomous vehicle fleet.
- Partner with the infrastructure and platform engineering teams to monitor, maintain, and troubleshoot our on-premises data offload and CI systems.
- Design, develop, and maintain business intelligence (BI) dashboards and ETL (extract, transform, load) pipelines to provide actionable insights into our infrastructure performance and health.
- Architect and deploy test environments to validate internal and customer-facing infrastructure solutions.
- Automate deployment, scaling, and upgrading of our remote monitoring software to ensure operational efficiency.
- Perform ongoing analysis of infrastructure performance, identifying opportunities for optimization in latency, throughput, and reliability.
- 5+ years of experience in a related role such as Site Reliability Engineer, Dev Ops Engineer, or Infrastructure Engineer.
- Strong knowledge of networking fundamentals, including protocols, troubleshooting, and optimization.
- Hands-on experience with Docker and related ecosystem tools (e.g., Docker Compose, Kaniko).
- Expertise in Kubernetes deployments and package management via Helm.
- Proficiency with relational and time-series databases (e.g., Postgres, Timescale DB, InfluxDB).
- Familiarity with workflow orchestration tools such as Argo and Airflow.
- Proven experience managing upgrades and rollbacks for customer-facing SaaS environments.
- Scripting experience in Python and Bash for automation and tooling.
- Experience building and maintaining dashboards with tools like Grafana.
Salary Range - $180,000- $260,000
More About GatikFounded in 2017 by experts in autonomous vehicle technology, Gatik has rapidly expanded its presence to Mountain View, Dallas-Fort Worth, Arkansas, and Toronto. As the first and only company to achieve fully driverless middle-mile commercial deliveries, Gatik holds a unique and defensible position in the AV industry, with a clear trajectory toward sustainable growth and profitability.
We have delivered complete, proprietary AV technology - an integration of software and hardware - to enable earlier successes for our clients in constrained Level 4 autonomy. By choosing the middle mile – with defined point-to-point delivery, we have simplified some of the more complex AV challenges, enabling us to achieve full autonomy ahead of competitors. Given extensive knowledge of Gatik's well-defined, fixed route ODDs and hybrid architecture, we are able to hyper-optimize our models with exponentially less data, establish gate-keeping mechanisms to maintain explainability, and ensure continued safety of the system for unmanned operations.
Visit us at Gatik for more company information and Careers at Gatik for more open roles.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).