Cloud Engineer
Listed on 2026-10-01
-
IT/Tech
Cloud Computing: Infrastructure & Operations, AWS, Systems Administrator, Systems Engineer
Cloud Engineer
K is seeking a highly qualified Cloud Engineer, AWS Administrator and Cloud Watch Specialist to support the Aviation Product Lifecycle Management program. The selected candidate will provide engineering, administration, monitoring, automation, and cybersecurity support for secure Amazon Web Services Gov Cloud environments supporting Navy and Department of Defense aviation systems. The Cloud Engineer will design, administer, secure, and sustain AWS infrastructure operating at Impact Levels 5 and 6.
The position will serve as a technical specialist for Amazon Cloud Watch monitoring, Windows and Linux environments, containerized applications, Kubernetes orchestration, disaster recovery, and infrastructure automation. The successful candidate will help ensure that AvPLM systems remain secure, resilient, observable, recoverable, and aligned with applicable Navy and DoD requirements, including established Recovery Time Objectives and Recovery Point Objectives.
- Administer the lifecycle of AWS Gov Cloud resources, including Amazon EC2 instances, Linux-hosted containers, Amazon S3 storage buckets, Amazon FSx storage shares, and associated network configurations.
- Configure and maintain secure Virtual Private Cloud environments, including subnets, route tables, security groups, Network Access Control Lists, and transit gateways.
- Implement and maintain Identity and Access Management policies, roles, and permissions in accordance with least-privilege principles.
- Use Infrastructure as Code tools, including Terraform and AWS Cloud Formation, to automate infrastructure deployment and configuration.
- Develop Python and Bash scripts to automate system deployment, patch management, and recurring administrative activities.
- Use AWS Systems Manager to support configuration management, patching, automation, and administration of cloud-hosted systems.
- Design, implement, and maintain enterprise monitoring solutions using Amazon Cloud Watch, Cloud Watch Logs, and AWS Event Bridge.
- Develop and automate deployment of the Cloud Watch unified agent across virtual-machine fleets to collect operating-system metrics and system logs.
- Create custom Cloud Watch dashboards that display key performance indicators, infrastructure health, system availability, and disaster recovery readiness.
- Configure Cloud Watch log groups, metric filters, alarms, and alerting capabilities.
- Establish event-driven remediation and notification workflows by integrating Cloud Watch Alarms with AWS Lambda, Amazon Simple Notification Service, and applicable IT service management tools.
- Implement monitoring capabilities for the network traffic, servers, databases, and supporting systems that comprise Commercial Off-the-Shelf Product Lifecycle Management applications.
- Monitor system performance and identify conditions that may affect availability, reliability, security, or mission operations.
- Deploy, configure, administer, and sustain Windows Server environments hosted on Amazon EC2.
- Support the availability, security, and performance of cloud-hosted Windows systems.
- Administer Active Directory, Group Policy Objects, Domain Name System services, and Internet Information Services.
- Integrate Windows Server environments with DoD Public Key Infrastructure to support Common Access Card authentication.
- Perform operating-system patching, upgrades, vulnerability remediation, and recurring maintenance using AWS Systems Manager, Windows Update Services, or other approved tools.
- Design, deploy, configure, and administer Kubernetes clusters using Amazon Elastic Kubernetes Service or self-hosted Kubernetes on Amazon EC2, based on AvPLM and COTS application requirements.
- Build, maintain, secure, and optimize Docker and other container images in alignment with DoD Dev Sec Ops practices.
- Administer Kubernetes components, including pods, deployments, services, ingress controllers, and Role-Based Access Control.
- Manage secure container registries, including Amazon Elastic Container Registry.
- Use Helm and other approved package-management tools to support application deployment, configuration, upgrades, and lifecycle management.
- Configure, test, and validate AWS disaster recovery capabilities using AWS Backup, AWS Elastic Disaster Recovery, multi-region replication, and other approved…
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).