More jobs:
Senior AWS Infrastructure/Dev Ops Engineer
Job in
Carlsbad, San Diego County, California, 92002, USA
Listed on 2026-06-21
Listing for:
Full Swing Simulators
Full Time
position Listed on 2026-06-21
Job specializations:
-
IT/Tech
AWS, SRE/Site Reliability, Cloud Computing: Infrastructure & Operations, Disaster Recovery IT
Job Description & How to Apply Below
AWS Infrastructure / Dev Ops Engineer (Senior)
As an AWS Infrastructure / Dev Ops Engineer (Senior), you will help own the reliability, security, scalability, and operational maturity of Full Swing's AWS environment. This role exists to eliminate AWS single-resource dependency, increase capacity for infrastructure automation, and strengthen incident response, observability, identity, data protection, and compliance-aligned controls. The AWS Infrastructure / Dev Ops Engineer will partner with engineering, QA, support, security, and business stakeholders to improve uptime, reduce mean time to repair (MTTR), reduce toil through automation, and ensure cloud operations remain documented, measurable, and resilient.
PrimaryFunctions
- Own and improve AWS infrastructure supporting Full Swing production, staging, and internal environments, with a focus on availability, scalability, security, and cost-conscious operations.
- Design, implement, and maintain AWS architecture across accounts, networking, compute, container or serverless services, databases, storage, load balancing, DNS/CDN, security tooling, backup, and disaster recovery capabilities.
- Build and maintain infrastructure as code using Terraform, AWS CDK, Cloud Formation, or similar tools, ensuring repeatable deployments, peer review, version control, and environment parity.
- Automate operational workflows including provisioning, configuration, CI/CD integrations, patching, backup validation, certificate or secret rotation, access reviews, and routine maintenance to reduce toil.
- Strengthen AWS identity and security posture through IAM least privilege, MFA/SSO patterns, encryption, secrets management, logging, vulnerability remediation, guardrails, and secure network controls.
- Maintain observability across AWS environments, including logs, metrics, traces, dashboards, alerts, service-level indicators, and actionable runbooks that improve detection, escalation, and MTTR.
- Lead or participate in production incident response, on-call escalation, root cause analysis, post-incident reviews, corrective actions, and stakeholder communication.
- Develop and test resilience practices including high availability patterns, capacity planning, scaling policies, backup and restore testing, disaster recovery plans, and peak-season readiness.
- Partner with engineering to improve release pipelines, environment management, deployment safety, rollback procedures, and production readiness standards.
- Maintain documentation, SOPs, runbooks, architecture diagrams, control evidence, and cross-training materials to reduce single points of failure and improve operational continuity.
- Review AWS costs and resource utilization; recommend rightsizing, lifecycle policies, savings inputs, and automation to control spend without compromising reliability.
- Support compliance-aligned controls related to change management, access, logging, encryption, data retention, backup/recovery, evidence collection, and audit readiness.
- Reduced AWS single-resource dependency through shared ownership, runbooks, cross-training, and documented operational procedures.
- Improved reliability and MTTR through better observability, alerting, incident response, rollback, and recovery practices.
- Increased engineering delivery capacity by automating recurring cloud operations toil and improving platform self-service.
- Improved security and compliance posture through customer-owned controls for identity, data protection, logging, backup/recovery, and change management.
- 7+ years of experience in infrastructure engineering, Dev Ops, site reliability engineering, cloud operations, systems engineering, or a related technical role; 5+ years of hands‑on AWS experience in production environments preferred.
- Proven experience operating business‑critical AWS environments with uptime expectations, production support, on-call participation, incident management, and documented recovery processes.
- Deep AWS expertise across core services such as IAM, Organizations/Control Tower, VPC, EC2, ECS/EKS, Lambda, RDS/Aurora, S3, ELB/ALB, Route 53, Cloud Front, WAF, Cloud Watch, Cloud Trail, Config, Guard…
Position Requirements
10+ Years
work experience
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
Search for further Jobs Here:
×