×
Register Here to Apply for Jobs or Post Jobs. X

Site Reliability Engineer

Job in New York, New York County, New York, 10261, USA
Listing for: The Cypress Group
Full Time position
Listed on 2026-07-10
Job specializations:
  • IT/Tech
    SRE/Site Reliability, Cloud Computing: Infrastructure & Operations, Systems Engineer, Network Engineer
Salary/Wage Range or Industry Benchmark: 140000 - 180000 USD Yearly USD 140000.00 180000.00 YEAR
Job Description & How to Apply Below
Location: New York

Senior Site Reliability Engineer (SRE / Infrastructure)

Role Overview

We’re hiring a Senior SRE to build and scale the infrastructure behind a high-growth, production system. You’ll ensure reliability, performance, and scalability as the platform grows from early traction to large-scale usage.

This role focuses on designing resilient systems, improving observability, and automating operations so engineering teams can move quickly and safely.

What You’ll Do

  • Own reliability, scalability, and performance of production systems
  • Build and manage cloud infrastructure (primarily AWS/GCP + Linux)
  • Design and operate Kubernetes clusters and containerized workloads
  • Improve CI/CD pipelines and deployment workflows
  • Lead incident response, on-call practices, and root cause analysis
  • Build observability systems (monitoring, logging, alerting)
  • Partner with engineers to design resilient systems (databases, pipelines, async systems)
  • Automate infrastructure and operational workflows using IaC

Requirements

  • 5+ years in SRE, Dev Ops, or infrastructure-focused engineering
  • Strong experience with cloud platforms (AWS/GCP) and Infrastructure as Code (e.g., Terraform)
  • Production experience with Kubernetes
  • Experience with monitoring/observability tools (e.g., Prometheus, ELK, Datadog)
  • Strong understanding of distributed systems, networking, and reliability best practices
  • Comfortable coding/scripting (e.g., Python, Go, or similar)

Nice to Have

  • Experience scaling high-availability systems
  • Familiarity with CI/CD and modern deployment strategies (canary, blue/green)
  • Background in data pipelines, async systems, or large-scale applications
  • Exposure to Go, Rust, C++, or Type Script
  • Interest in applying AI to infrastructure or operations
#J-18808-Ljbffr
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary