Sr. DevOps Engineer
Listed on 2026-08-10
-
IT/Tech
Cloud Computing: Infrastructure & Operations, Systems Engineer, SRE/Site Reliability, AWS
Senior Dev Ops Engineer
At EKN Engineering, we solve challenging problems with innovative engineering and configurable software solutions. We use an engineering-first approach, combined with data-driven strategies, to improve overall compliance, risk management, and design accuracy tailored to the specific needs of each client.
With decades of experience in engineering, our team of over 170 professionals in Irvine, California is dedicated to building a safer and more efficient tomorrow through engineering and technological innovation.
Role OverviewAs a Senior Dev Ops Engineer, you’ll join EKN’s software development team — responsible for architecting, implementing, testing, deploying, and supporting SaaS solutions that align with customer and product requirements — to design and maintain scalable cloud infrastructure and implement best practices for CI/CD, Infrastructure as Code (IaC), logging, monitoring, and automation. Bringing strong technical expertise and a product mindset, you’ll take a hands‑on approach to driving engineering best practices and staying current on emerging technologies, platforms, and applications to help keep our business secure and efficient.
Key ResponsibilitiesInfluence the development and architecture of EKN's Dev Ops platform, providing technical leadership and guidance to development teams
Assess and identify risks that may affect the team's larger goals
Document best practices and strategies for application deployment and infrastructure maintenance
Partner with software development teams and leaders to identify and implement optimal cloud-based solutions for the company
Drive implementation of functionally appropriate and technically sound solutions that meet all quality and security standards
Identify and resolve technical problems using strong analytical and problem-solving skills
Participate in all aspects of the software development lifecycle for cloud-based solutions, including planning, defining requirements, developing, and testing
Evaluate new technologies, tools, and platforms to manage, control cost, and secure the cloud platform
Contribute to collective team goals by providing regular feedback and driving continuous improvement
Respond quickly to production outages, performance drops, and pipeline failures
Participate in after‑hours on‑call rotation to triage and fix system alerts
Execute planned off‑hours maintenance, database migrations, and critical security patching
Bachelor's degree in Computer Science, Engineering, or a related field, or equivalent practical experience
5+ years of hands‑on Dev Ops or Site Reliability Engineering experience, primarily supporting production workloads in AWS
Strong experience designing, deploying, and managing cloud infrastructure using Infrastructure as Code (IaC) tools such as Terraform or AWS Cloud Formation
Solid working knowledge of core AWS services (e.g., EC2, EKS, S3, RDS, SQS, SNS, Lambda) and how they are securely operated at scale
Proven experience with containers, containter orchestration, and deployment technologies, such as Docker, Kubernetes and ArgoCD
Experience building, operating, and improving CI/CD pipelines using platforms such as Git Hub Actions or equivalent
Working knowledge of Git-based workflows and version control best practices
Strong understanding of cloud networking fundamentals, security best practices, and identity and access management (IAM)
Ability to identify, troubleshoot, and resolve complex infrastructure, deployment, and reliability issues
Practical scripting and automation experience (e.g., shell scripting) to support infrastructure and operational workflows
Strong written and verbal communication skills, with the ability to collaborate effectively across engineering teams
Prior experience as a software engineer or cloud‑native application developer, enabling strong collaboration with product and development teams
Hands‑on experience building or supporting modern applications using languages such as Type Script or Python
Experience designing and operating monitoring, logging, and observability solutions (e.g., Prometheus, Grafana, Datadog, Sentry) to improve…
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).