×
Register Here to Apply for Jobs or Post Jobs. X

Site Reliability Engineer

Job in Barrington, Bristol County, Rhode Island, 02806, USA
Listing for: Workiy
Full Time position
Listed on 2026-08-29
Job specializations:
  • IT/Tech
    SRE/Site Reliability, Cloud Computing: Infrastructure & Operations
Job Description & How to Apply Below

Site Reliability Engineer (SRE)

We are seeking an experienced Site Reliability Engineer (SRE) to design, implement, and maintain highly available, scalable, secure, and reliable production systems. The ideal candidate will have strong software engineering skills combined with hands-on experience in cloud infrastructure, Kubernetes, automation, monitoring, and observability. The SRE will work closely with development, infrastructure, and operations teams to improve system reliability, automate operational processes, and resolve complex production issues.

Location:

Barrington, Rhode Island, United States, 02806

Industry: IT Services

Job Type: Contract

About Us:

Workiy is a leading IT solutions and staff augmentation company committed to driving business success through talent acquisition and technology expertise. With a proven track record of excellence, Workiy empowers organizations across industries by connecting them with top-tier IT professionals and delivering innovative technology solutions.

Requirements:

  • Design, deploy, and maintain highly reliable and scalable production systems.
  • Develop automation and reliability tooling using Go, Python, Java, or Rust.
  • Manage and support cloud environments across AWS, Azure, or GCP.
  • Deploy and manage containerized applications using Docker and Kubernetes.
  • Implement and maintain monitoring, logging, metrics, and distributed tracing solutions.
  • Build and enhance observability using Open Telemetry (OTel) and related technologies.
  • Troubleshoot complex infrastructure, application, and production issues.
  • Participate in incident response, root-cause analysis, and post-incident reviews.
  • Automate repetitive operational tasks and improve engineering efficiency.
  • Implement Infrastructure as Code using Terraform and/or Ansible.
  • Develop and maintain CI/CD pipelines for reliable and automated deployments.
  • Monitor system performance, availability, capacity, and overall reliability.
  • Identify reliability risks and implement proactive solutions.
  • Collaborate with software developers to improve application reliability and performance.
  • Establish and improve SRE best practices, operational procedures, and reliability standards.

Required

Skills & Experience:

  • Proven experience as a Site Reliability Engineer, Production Engineer, Dev Ops Engineer, or similar role.
  • Strong programming experience with at least one of:
    Go/Golang, Python, Java, Rust.
  • Hands-on experience with AWS, Azure, or GCP.
  • Strong experience with Kubernetes and Docker.
  • Strong Linux/Unix administration and troubleshooting skills.
  • Experience with Open Telemetry and observability.
  • Knowledge of monitoring and visualization tools such as Prometheus and Grafana.
  • Experience with Terraform, Ansible, or similar Infrastructure-as-Code tools.
  • Strong understanding of CI/CD pipelines and Dev Ops practices.
  • Experience with production incident management and Root Cause Analysis (RCA).
  • Strong knowledge of automation, scripting, networking, and distributed systems.
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary