×
Register Here to Apply for Jobs or Post Jobs. X

Sr. Site Reliability Engineer

Job in Northern, Floyd County, Kentucky, USA
Listing for: Far Coder
Full Time position
Listed on 2026-08-13
Job specializations:
  • IT/Tech
    SRE/Site Reliability, Cloud Computing: Infrastructure & Operations, Systems Engineer
Salary/Wage Range or Industry Benchmark: 25000 - 40000 USD Yearly USD 25000.00 40000.00 YEAR
Job Description & How to Apply Below
Location: Northern

# Remote Sr. Site Reliability Engineer Job at Meridian

LinkUSA5 hours agoFull TimeUSA $25000 - $40000 USDKubernetes

TerraformCI/CDPythonAWSAzure Bash Compliance Dev Ops Excel MEANPredictive Analytics“When applying, mention the word Far Coder to show you’ve read the job post completely. Employers can look for these words to identify genuine, thoughtful applicants and avoid spam.”## Job Overview Meridian Link  is hiring a remote candidate for Sr. Site Reliability Engineer. This is a full time position.

Work location:

USA.The role typically involves technologies such as Kubernetes, Terraform, CI/CD, Python, AWS, Azure.## Required Skills### Primary Skills
* Kubernetes* Terraform
* CI/CD
* Python* AWS
* Azure### Secondary Skills
* Bash* Compliance
* Dev Ops* Excel
* MEAN* Predictive Analytics Skills required for this role include Kubernetes, Terraform, CI/CD, and related tools for day-to-day development.## Job Details
* *
* Employment Type:

** Full Time
* *
* Location:

** USA
* ** Salary:** $25000 – $40000 USD## Tech Stack Kubernetes, Terraform, CI/CD, Python, AWS, Azure, Bash, Compliance, Dev Ops, Excel, MEAN, Predictive Analytics## Role details

About the Role We are seeking a Senior Site Reliability Engineer to join our cloud engineering team. You will own the reliability, scalability, and observability of our critical financial SaaS applications and infrastructure, working across cloud platforms to ensure our customers experience is seamless, secure, and performant services. This is a high-impact role for someone who is passionate about building resilient systems and preventing outages before they happen.

Key Responsibilities
* Design, implement, and maintain Service Level Objectives (SLOs) and Service Level Indicators (SLIs) across all critical systems; ensure we meet or exceed targets consistently
* Lead observability strategy by designing comprehensive monitoring, logging, and tracing architectures; select and deploy observability tools that provide deep visibility into system behavior
* Build and own runbooks, incident response procedures, and post-incident review processes; mentor the team on incident management and blameless postmortems
* Architect and deploy cloud infrastructure on AWS or Azure; implement infrastructure-as-code practices and ensure high availability, disaster recovery, and business continuity
* Develop automation and AIOps capabilities to reduce toil, accelerate incident detection, and enable self-healing systems; implement intelligent alerting to minimize false positives
* Drive reliability improvements through load testing, chaos engineering, and failure scenario analysis; identify and eliminate single points of failure
* Partner with application and backend teams to design reliable systems from inception; conduct architecture reviews and reliability assessments
* Write production-grade Python tooling for automation, metrics collection, alert management, and operational workflows
* Champion security and compliance in infrastructure; implement defense-in-depth principles for a regulated fintech environment

Required Qualifications
* 7+ years in Site Reliability Engineering, Dev Ops, platform engineering, or closely related roles with significant responsibility for production systems
* Expert-level experience with Azure or AWS (or both); deep knowledge of compute, networking, storage, and managed services; experience managing infrastructure at scale
* Demonstrated expertise in observability: designing and implementing monitoring, alerting, logging, and distributed tracing solutions; hands-on with observability platforms (e.g., Prometheus, Grafana, ELK, Datadog, New Relic, or similar)
* Strong background in SLOs, SLIs, and SLAs; experience defining meaningful objectives and building systems to meet them; understanding of error budgets and their role in prioritization
* Proven experience designing and troubleshooting highly available, resilient, and scalable systems; deep understanding of distributed systems concepts and failure modes
* Proficiency in Python, Power Shell, bash, etc. scripting languages for production automation, tooling, and systems programming; ability to write clean, maintainable code…
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary