Sr DevOps Engineer
Listed on 2026-09-20
-
IT/Tech
Cloud Computing: Infrastructure & Operations, SRE/Site Reliability, AWS, Systems Engineer
Dev Ops Engineer
Indianapolis, IN
Hybrid Role
$65 - $75 per hour
THE ROLE
Eight Eleven Group, LLC is seeking a skilled Dev Ops Engineer to join our team in Indianapolis, IN. In this hybrid role, you will be responsible for designing, deploying, and maintaining scalable AWS cloud environments, leveraging a wide range of AWS services to ensure high availability and performance. You will develop Infrastructure as Code solutions, build and maintain CI/CD pipelines, and manage containerized workloads in Kubernetes clusters.
The role also involves implementing robust monitoring and alerting solutions, supporting large-scale cloud-native microservices, and troubleshooting complex production issues. You will work within Agile teams, create operational documentation, and provide services to clients across the US, with the potential for travel and relocation. This is an excellent opportunity for a motivated engineer with a strong background in cloud infrastructure, automation, and Dev Ops best practices to work with advanced technologies in a collaborative, client-focused environment.
YOU'LL DO
- Design, deploy, and maintain scalable AWS cloud environments using services such as EC2, S3, Lambda, VPC, Route 53, Cloud Watch, Cloud Trail, Redshift, and Load Balancers, implementing auto-scaling and high availability strategies
- Develop Infrastructure as Code solutions with Terraform, integrating AWS Lambda and Code Pipeline for automated deployments and resilient architectures
- Build and maintain CI/CD pipelines using Git, Docker, Kubernetes, Terraform, and Ansible to support automated application delivery
- Manage containerized workloads by building Docker images and orchestrating deployments in Kubernetes clusters, including AWS EKS cluster management and workload deployments
- Participate in Agile development processes (Scrum, Kanban), manage work items in JIRA, and maintain documentation in Confluence
- Implement monitoring, logging, and alerting solutions using Grafana, Prometheus, and Datadog for infrastructure and application health across cloud and on-prem environments
- Support large-scale cloud-native microservices environments by implementing monitoring frameworks, log analysis procedures, and proactive alerting mechanisms to detect anomalies early
- Troubleshoot complex production issues, including Tier IV escalations, system outages, and connectivity problems affecting cloud services or WiFi network infrastructure
- Create operational documentation, procedures, and deployment MOPs to standardize application rollout processes and improve operational efficiency
- Deploy and manage Datadog agents across cloud and on-prem servers, build monitoring dashboards, and conduct capacity planning to maintain system performance and stability
- Implement infrastructure and network monitoring exporters, enabling real-time metrics collection for WiFi infrastructure and services
- Utilize network analysis tools (Net Scout nGeniusONE, nGeniusPULSE) for traffic diagnostics, performance monitoring, and identifying abnormal traffic patterns including potential DDoS or malicious activity
- Generate analytical reports and dashboards using Net Scout reporting tools to support capacity planning, network optimization, and performance improvements
- Provide services to clients throughout the US, with willingness to travel and relocate as needed
- Master’s degree in Computer Science, Technology Management, or a related field
- Two years of experience performing Dev Ops engineering duties as described above
- Proficiency with AWS services (EC2, S3, Lambda, VPC, Route 53, Cloud Watch, Cloud Trail, Redshift, Load Balancers), including auto-scaling and high availability strategies
- Experience developing Infrastructure as Code solutions using Terraform, and integrating AWS Lambda and Code Pipeline for automated deployments
- Hands-on experience building and maintaining CI/CD pipelines with Git, Docker, Kubernetes, Terraform, and Ansible
- Experience managing containerized workloads, building Docker images, and orchestrating deployments in Kubernetes clusters (including AWS EKS)
- Familiarity with Agile methodologies (Scrum, Kanban), and experience using JIRA and Confluence for work management and documentation
- Experience implementing monitoring, logging, and alerting solutions with Grafana, Prometheus, and Datadog for both cloud and on-prem environments
- Ability to support and monitor large-scale cloud-native microservices environments, including log analysis and proactive alerting
- Strong…
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).