×
Register Here to Apply for Jobs or Post Jobs. X

SRE

Job in Chicago, Cook County, Illinois, 60290, USA
Listing for: Benton Partners
Full Time position
Listed on 2026-08-13
Job specializations:
  • IT/Tech
    Cloud Computing: Infrastructure & Operations, SRE/Site Reliability, Systems Engineer
Salary/Wage Range or Industry Benchmark: 175000 - 225000 USD Yearly USD 175000.00 225000.00 YEAR
Job Description & How to Apply Below

Senior Site Reliability Engineer - Platform

Chicago New York

We arelookingfora Site Reliability Engineer, to join our growing Platform Engineering team,who can cultivate our SRE philosophy, processes, and technologies from the ground up.

This roleentailsdriving standards and fostering adoption across our technology teams, whilstcloselypartnering with our Dev Ops and Cloud teams.

With a hands-on approach,you'llwork across both cloud and on-premises hosting platforms, ensuring the reliability and scalability of ourtradingsystemsand production environments. This is a chance to play a pivotal role in transforming our operational capabilities and enhancing performance across a wide array of environments and platforms.

Key Responsibilities:

  • Develop and promote our SRE philosophy,establishing best practices and processes that will be instrumental in scaling our infrastructure.
  • Implement and scaleend-to-end observability and monitoring solutions using Prometheus, Grafana, Loki, and Tempo, ensuring high visibility into application performance and infrastructure health.
  • Participate in on-call rotation with approximately 1 week per month of on-call time shared equally across members of the team
  • Review and define standards for application reliability requirements within our Kubernetesenvironment, ensuring application configuration isoptimizedfor performance,costand reliability.
  • Develop automation and tooling to improve efficiency and reliability of deployment pipelines, system health checks, and recovery procedures.
  • Collaborate with development teams to enhance service stability, scalability, and fault tolerance through SRE best practices like blameless post-mortems and service level objectives(SLOs).

To be considered a good fit, you musthave:

  • 8+ years of experience in SRE or similar roles within complex, distributed systems environments.
  • SMEwith key SRE technologies such as Prometheus, Grafana, Loki,Tempo,andOpenTelemetry.
  • Extensive knowledge of container orchestration using Kubernetes and containerization with Docker.
  • Hands-on experience with both cloud (AWS preferred) and on-premises hosting platforms.
  • Proven ability to script in languages like Python, Bash, or Go, to automate routine tasks and deployment pipelines.
  • Strong understanding of CI/CD principles, agile methodologies, and Dev Ops culture.
  • High levelof initiative, passion for reliability engineering, detail orientation, and follow-through capabilities.
  • Exceptional interpersonal and communication skills, with the ability to explain complex technical concepts to a diverse audience.

With respect to NY, CA, and IL based applicants, the starting base pay range for this role is between USD 175000 and USD 225000 annually. The actual base pay is dependent upon several factors, including, but not limited to, relevant experience, business needs and market demands. This role may also be eligible for bonus compensation and employee benefits.

#J-18808-Ljbffr
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary