×
Register Here to Apply for Jobs or Post Jobs. X

Site Reliability Engineer

Job in South Naperville Area, Will County, Illinois, 60564, USA
Listing for: SCIGON
Full Time position
Listed on 2026-08-22
Job specializations:
  • IT/Tech
    SRE/Site Reliability, Cloud Computing: Infrastructure & Operations, Systems Engineer, IT Project Manager
Salary/Wage Range or Industry Benchmark: 110000 - 170000 USD Yearly USD 110000.00 170000.00 YEAR
Job Description & How to Apply Below

The Site Reliability Engineer (SRE) – Release & Operations is a hybrid technical and operational role responsible for bridging software development and technology operations. This position focuses on managing and optimizing the software release lifecycle, driving change management governance, and facilitating communication across technical and business teams.

In addition to release engineering responsibilities, this role requires a hands-on technical professional who can develop automation solutions to reduce operational overhead, actively support production environments, and help ensure the availability, security, scalability, and reliability of cloud-based infrastructure and applications.

Responsibilities Release Management & Change Governance
  • Lead and coordinate the end-to-end software release lifecycle, including planning, scheduling, staging, deployment, validation, and post-release activities.
  • Participate in and facilitate change management processes, evaluating release readiness, assessing risks, and ensuring governance and compliance requirements are met prior to production deployment.
  • Serve as a primary point of contact for engineering, quality assurance, product, operations, and business stakeholders regarding deployment schedules, release status, risk assessments, and rollback strategies.
  • Continuously improve release management processes by transitioning manual activities into automated, repeatable workflows and CI/CD controls.
  • Drive release standardization and promote best practices across development and operations teams.
Site Reliability & Production Operations
  • Provide hands-on production support, ensuring operational stability and participating in incident response and on-call support activities.
  • Design, develop, and maintain automation tools, scripts, and workflows to reduce operational effort and improve system reliability.
  • Respond to service interruptions, outages, and operational incidents, performing root cause analysis and implementing long-term corrective actions.
  • Implement and maintain Infrastructure-as-Code (IaC) solutions to ensure consistent, scalable, and repeatable infrastructure deployments.
  • Configure, manage, and optimize monitoring, logging, and alerting systems to provide visibility into application and infrastructure health.
  • Identify opportunities to improve system reliability, scalability, performance, and operational efficiency.
Security, Compliance & Collaboration
  • Ensure deployment and operational processes adhere to applicable security, governance, compliance, and risk management requirements.
  • Collaborate closely with software engineering, quality assurance, platform, infrastructure, and security teams to establish reliable deployment and operational practices.
  • Maintain accurate and audit-ready documentation, including operational procedures, deployment records, runbooks, incident reports, and change logs.
  • Support continuous improvement initiatives related to operational excellence, reliability engineering, and service delivery.
Required Skills & Experience
  • Experience in Site Reliability Engineering (SRE), Dev Ops, system administration, cloud operations, release engineering, or related technical roles.
  • Demonstrated experience managing software releases and coordinating deployments within structured change management processes.
  • Knowledge of cloud infrastructure platforms and services.
  • Experience building, deploying, and maintaining scalable cloud-based systems and applications.
  • Solid understanding of Infrastructure-as-Code (IaC) principles and tools.
  • Strong scripting and automation skills using languages such as Python, Bash, Power Shell, or similar technologies.
  • Experience supporting production systems and participating in incident management and root cause analysis activities.
  • Understanding of monitoring, observability, alerting, and operational support practices.
  • Strong communication and stakeholder management skills, with the ability to collaborate across technical and non-technical teams.
  • Ability to balance operational stability with delivery speed and business objectives.
Preferred Qualifications
  • Experience with change management, IT service management, or governance…
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary