Site Reliability Engineer
Listed on 2026-07-25
-
IT/Tech
Cloud Computing: Infrastructure & Operations, SRE/Site Reliability, IT Support, Systems Engineer
Job Posting End Date:
August 04
Our Fortune 500 company is driving a digital transformation and looking for forward-thinking innovators to disrupt how our industry thinks about and uses technology. As one of the world's leading employee benefits providers, we help millions of people gain affordable access to benefits that help them protect their families, their finances and their futures.
Are you an asker of questions, a solver of problems, and a challenger of the status quo? Our mission is to provide a differentiated customer experience and exceed the expectations people have of technology at any company — not just insurers.
We are seeking individuals to join our team of talented IT professionals who share never-ending passion and an unwavering focus on our customer experience. Team members comfortable working in an agile, fast-paced, and delivery-focused environment thrive in our environment where we value an entrepreneurial spirit and those who challenge the status-quo.
Unum is changing, and we’re excited about what’s next. Join us.
General Summary- Design, build, and maintain observability, monitoring, and alerting capabilities across consumer and client-facing digital platforms
- Develop and maintain dashboards that measure availability, latency, error rate, throughput, capacity, MTTR/MBTI, and other reliability metrics
- Diagnose and troubleshoot distributed systems issues across cloud-based and on-prem services
- Partner with engineering teams to improve service reliability, reduce operational toil, and mature incident response practices
- Implement automation for service health checks, performance monitoring, and remediation
- Manage CI/CD pipeline reliability and deployment quality controls
- Conduct root-cause analysis and drive long-term corrective actions
- Collaborate with Run teams to transition monitoring, dashboards, and operational insights into production support processes
- Provide guidance on service re-platforming, performance improvements, and architectural decisions based on reliability data
Requires a Bachelor’s degree in Computer Science, Engineering, or related field plus 5 years of experience. Requires 5 years of experience with the following:
Observability and monitoring platforms used to monitor application performance and system health using Dynatrace, AWS Cloud Watch, Datadog, Grafana, or Amplitude;
Working with containerized and cloud-native architectures, including deployment, configuration, and operational support in cloud environments, using AWS;
Supporting incident response processes, including participation in on-call rotations, post-incident reviews, and implementation of service-level objectives (SLOs), service-level indicators (SLIs), or service-level agreements (SLAs);
Developing scripts or automation to improve system reliability or operational efficiency using Python, Bash, or Power Shell;
Troubleshooting distributed systems and analyzing performance bottlenecks across multi-tier or microservices-based architectures;
Collaborating with cross-functional engineering teams, including software engineering, platform, infrastructure, or operations teams, within a Dev Ops or reliability-focused environment; working with version control systems and collaborative development workflows using Git Hub, Git Lab, or Bitbucket. Requires 4 years of experience with the following:
Designing, implementing, or maintaining logging, metrics, and distributed tracing pipelines for enterprise or cloud-based systems;
Hands-on experience with continuous integration and continuous deployment (CI/CD) tools and pipelines, using Git Hub Actions, Jenkins, or Azure Dev Ops. Requires 3 years of experience with using infrastructure-as-code or configuration management tools to provision, manage, or maintain environments, including Terraform, AWS Cloud Formation, or Ansible. Telecommuting w/i worksite. Up to 5% domestic travel.
40 hours/week; $152,131 - $162,131 per year. This wage range supersedes the base salary range listed below, due to the salary range below reflecting a national range.
~IN1
Our company is built on helping individuals and families, and this starts with our employees. We want employees to…
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).