Senior AWS Site Reliability Engineer/Infrastructure Engineer
Listed on 2026-09-02
-
IT/Tech
AWS, Cloud Computing: Infrastructure & Operations, SRE/Site Reliability, Systems Engineer
AWS Site Reliability Engineer (SRE) / Infrastructure Engineer
One of Insight Global's largest Payroll Clients is seeking a highly experienced AWS Site Reliability Engineer (SRE) / Infrastructure Engineer to join their Public Cloud Engineering team. This individual will play a critical role in supporting and enhancing their AWS cloud infrastructure while driving infrastructure automation, reliability, and operational excellence across enterprise-scale cloud environments. This is a hands-on engineering role requiring deep expertise in AWS, Terraform, Jenkins, and Python.
The ideal candidate will be capable of independently owning technical initiatives, collaborating with cross-functional stakeholders, and delivering cloud infrastructure solutions with minimal oversight.
Day to Day Responsibilities:
Cloud Infrastructure & Reliability Engineering
- Design, implement, and maintain scalable, secure, and highly available AWS infrastructure.
- Support ongoing cloud engineering initiatives and operational activities within a large enterprise environment.
- Act as a technical lead for infrastructure-related projects, partnering with engineering, security, networking, and application teams.
- Troubleshoot complex infrastructure, reliability, and performance issues across AWS services.
- Implement best practices around resiliency, observability, scalability, and operational excellence.
Infrastructure as Code (Terraform)
- Build new Terraform modules from scratch to support enterprise cloud initiatives.
- Enhance, refactor, and maintain existing Terraform codebases.
- Establish reusable infrastructure patterns and standards across AWS environments.
- Ensure infrastructure deployments are automated, version-controlled, and repeatable.
CI/CD & Automation
- Design, build, and maintain Jenkins pipelines supporting infrastructure and application deployments.
- Develop new CI/CD workflows while also optimizing existing pipelines.
- Drive automation initiatives to reduce manual effort and improve operational efficiency.
- Collaborate with development teams to streamline deployment and release processes.
Python Development & Infrastructure Automation
- Develop automation tools and scripts using Python.
- Build solutions supporting:
Infrastructure provisioning and configuration management AWS Lambda functions and serverless automation Monitoring and operational tooling Infrastructure health checks and remediation workflows Data processing and cloud operations automation - Create scalable solutions to improve platform reliability and operational efficiency.
Stakeholder Management
- Partner directly with technical and business stakeholders to gather requirements and deliver solutions.
- Take ownership of projects from initial design through implementation and production support.
- Provide technical guidance and recommendations related to AWS cloud best practices.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).