×
Register Here to Apply for Jobs or Post Jobs. X

Senior DevOps​/SRE Engineer

Job in Chicago, Cook County, Illinois, 60290, USA
Listing for: Sei-Investments
Full Time position
Listed on 2026-07-21
Job specializations:
  • IT/Tech
    SRE/Site Reliability, Cloud Computing: Infrastructure & Operations, AWS, Systems Engineer
Salary/Wage Range or Industry Benchmark: 140000 - 170000 USD Yearly USD 140000.00 170000.00 YEAR
Job Description & How to Apply Below
We are looking for a Senior Site Reliability Engineer to work as part of a lean, product ‑ focused engineering organization. This role is about building and operating reliable cloud ‑ based systems by writing code, automating infrastructure and delivery workflows, and reducing friction for developers and users. You will work closely with product and application engineers to design, deploy, and operate systems with clear ownership and practical engineering judgment.

We expect you to use modern tooling, including AI ‑ assisted tools where appropriate , to speed up automation, troubleshooting, and operations while remaining accountable for correctness, security, and reliability. This role favors simple, effective solutions, hands ‑ on ownership, and continuous improvement within small Agile teams.

What you will do:

Design , build, and operate cloud infrastructure for critical production and non ‑ production applications with reliability and simplicity as primary goals

Architect and evolve multi-account AWS foundations (organizations, accounts, IAM boundaries, guardrails, and environment separation) to enable secure, scalable delivery

Design and operate cloud networking architecture (VPCs, routing, segmentation, ingress/egress, connectivity patterns) to support reliability, security, and compliance requirements

Treat reliability, security, and compliance as first ‑ class design concerns throughout the system lifecycle

Build tooling and automation that reduces errors, shortens recovery time, and improves day ‑ to ‑ day operations

Implement monitoring, logging, and alerting that make system behavior observable and actionable

Use AI ‑ assisted tools to accelerate infrastructure delivery, automation, troubleshooting, and root ‑ cause analysis, applying engineering judgment to validate outcomesI mplement reliability guardrails for releases (progressive delivery, safe rollbacks, change risk controls) and provide production support during deployments.

Participate in incident response, perform root cause analysis, and drive durable improvements that prevent recurrence

Work closely with application engineers to co ‑ own system design, operation, and continuous improvement

Maintain clear, lightweight documentation that supports shared ownership and effective on ‑ call operations

What we need from you:

BA/BS, in a related technical field; or the equivalent in education and work experience8+ years of experience in Dev Ops, SRE, platform engineering, or similar roles supporting application teams running production services

Strong CI/CD experience (Jenkins and Git-based workflows preferred), including building secure, reliable pipelines and enabling teams to ship safely

Experience implementing and operating observability platforms (logging/metrics/alerting);
Elastic Stack/Open Search experience is a plus Hands-on, demonstrable experience designing and operating AWS environments, and enabling application teams to adopt AWS correctly (networking, IAM, security, reliability, and cost awareness)
Infrastructure as Code experience (Terraform preferred; Cloud Formation acceptable), including building reusable modules/patterns and managing changes through review and automation

Experience supporting CI/CD builds and deployment patterns for common application stacks (for example Java, NodeJS, and .NET)
Experience scripting in Bash, Python, or Power Shell Experience  working on large scale cloud-based web applications

What we would like from you:

Ability to clearly communicate both verbally and in writing with client and team members, including experience documenting and presenting findings

Excellent analytical skills, organizational abilities, and problem-solving skills

Familiarity with AI agentic development tools (e.g., Claude Code, Git Hub Copilot, Windsurf) and practical experience applying them to infrastructure and operations workflows

Self-starter who works efficiently in a fast-paced environment with changing priorities and a geographically distributed team Ability to think creatively and seek optimum solutions

Ability to grasp loosely defined concepts and transform them into tangible results and key deliverables

Diagnostic skills…
Position Requirements
10+ Years work experience
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary