×
Register Here to Apply for Jobs or Post Jobs. X

Senior Reliability Engineer; Remote

Remote / Online - Candidates ideally in
Menomonee Falls, Waukesha County, Wisconsin, 53051, USA
Listing for: Kohl's
Remote/Work from Home position
Listed on 2026-06-18
Job specializations:
  • IT/Tech
    Systems Engineer, Cloud Computing: Infrastructure & Operations, SRE/Site Reliability
Salary/Wage Range or Industry Benchmark: 100000 - 125000 USD Yearly USD 100000.00 125000.00 YEAR
Job Description & How to Apply Below
Position: Senior Reliability Engineer (Remote)

About the Role

As Senior Reliability Engineer, you will ensure the resilience and availability of Kohl’s systems and applications, collaborate closely with development teams, contribute to architectural designs, conduct risk assessments and design for failure, and implement robust monitoring and failover mechanisms.

What You’ll Do
  • Drive error budget and Service Level Objective (SLO) adoption across products
  • Drive incident response efforts, perform root cause analysis and implement preventative measures to enhance system reliability
  • Establish consistent practices that elevate Kohl’s operational excellence through automation and process improvements
  • Follow software lifecycle and drive reliability, observability, and efficiency across product teams within an assigned domain
  • Identify repeated toil and find opportunities for automation and risk reduction
  • On-call on a rotation to respond to production incidents and conduct blameless retrospectives and root‑cause analyses (RCAs) to drive a culture of continuous improvements
  • Proactively identify failures before they cause outages using chaos engineering techniques such as edge cases, failure modes and design review
  • Advise on capacity planning and provide continuous assessments on systems behavior and consumption
  • Work with product managers to identify and prioritize work for reliability best practices (i.e., leveraging SLIs/SLOs/Error Budgets)
  • Mentor and assist engineers on the team
  • Additional tasks may be assigned
Required What Skills You Have
  • Bachelor’s Degree or equivalent in MIS, Computer Science or related field
  • 4+ years of experience in software development
  • Strong programming skills in one or more languages (Java, Python, Go or Node.js)
  • In-depth knowledge of systems architecture, operating system internals and network fundamentals
  • In-depth knowledge of application design patterns, event‑driven architecture, database schemas, and testing strategies
  • Experience with multi‑region application troubleshooting and performance tuning
  • Working experience with one cloud platform (GCP, AWS, or Azure)
  • Working experience with monitoring techniques and tools (e.g., Cloud Watch, Grafana, Prometheus, Open Telemetry, Tracing)
Preferred
  • In‑depth knowledge of containerization and container orchestration (e.g., Docker, Kubernetes, Rancher)
  • Experience with one or more configuration management systems (e.g., Chef, Ansible, Puppet)
  • Passion for and experience with AI and ML methodologies (MLOps)
  • Experience writing Infrastructure as code (e.g., Terraform, Open Tofu)
#J-18808-Ljbffr
Position Requirements
10+ Years work experience
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary