×
Register Here to Apply for Jobs or Post Jobs. X

Senior Site Reliability Engineer

Job in San Jose, Santa Clara County, California, 95199, USA
Listing for: Archer
Full Time position
Listed on 2026-09-12
Job specializations:
  • IT/Tech
    Cloud Computing: Infrastructure & Operations, Systems Engineer, SRE/Site Reliability
Salary/Wage Range or Industry Benchmark: 160000 - 210000 USD Yearly USD 160000.00 210000.00 YEAR
Job Description & How to Apply Below
  • We are seeking a highly experienced and passionate Sr. Staff Site Reliability Engineer (SRE) to join our growing team. In this critical role, you will be responsible for the reliability, scalability, performance, and security of our core systems and services
  • You will leverage your extensive expertise in various technologies to design, implement, and maintain robust infrastructure and automation solutions
  • Implement and maintain the infrastructure and pipeline required for an internal LLM-powered chat service, potentially leveraging platforms like Open Router or similar alternatives
  • Implement and maintain highly available, scalable, and secure cloud-native infrastructure on Amazon Elastic Kubernetes Service (EKS)
  • Develop and implement comprehensive observability strategies, including monitoring, logging, and alerting, to ensure the health and performance of our systems
  • Architect and optimize data pipelines to ensure efficient and reliable data flow across various platforms
  • Drive the continuous improvement of our CI/CD pipelines, promoting best practices for automated testing, deployment, and release management
  • Champion cloud-first strategies, leveraging the full capabilities of cloud platforms for infrastructure, services, and operations
  • Implement and enforce robust security practices across our infrastructure, applications, and data
  • Design and maintain Docker-based containerization solutions for our applications
  • Develop and maintain automation scripts and tools using Python, Bash, and Power Shell
  • Collaborate with development teams to ensure reliability is built into the software development lifecycle from inception
  • Troubleshoot complex production issues across various layers of the stack, identifying root causes and implementing preventative measures
  • Participate in on-call rotations to support production systems

Proven track record in designing and implementing robust data pipelines (e.g., Kafka, Airflow, Spark)
Ability to work independently and as part of a highly collaborative team Expert-level knowledge of cloud platforms (AWS preferred), including infrastructure-as-code principles

Strong background in CI/CD methodologies and tools (e.g., Jenkins, Git Lab CI, ArgoCD)
Comprehensive understanding of security best practices for cloud environments, applications, and data Solid understanding of networking concepts, distributed systems, and operating systems
12+ years of experience in Site Reliability Engineering, Dev Ops, or a similar role with a strong focus on operational excellence

Excellent problem-solving, analytical, and communication skills

Advanced scripting and programming skills in Python, Bash, and Power Shell Proficiency  in Docker for containerization and orchestration

Extensive experience with observability tools and practices, including Prometheus, Grafana, ELK stack, or similar

Bachelor’s degree in Computer Science, Engineering, or a related field, or equivalent practical experience

Deep expertise in Amazon EKS, including cluster provisioning, management, and troubleshooting

Successful candidates must be able to demonstrate U.S. citizenship, permanent residency, or status as a protected individual to satisfy ITAR, contractual, and/or regulatory requirements

Certifications in AWS, Kubernetes, or other relevant technologies

Experience with other Kubernetes distributions or cloud providers

Familiarity with compliance frameworks (e.g., SOC 2, HIPAA, GDPR)

Position Requirements
10+ Years work experience
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary