Platform Engineer Senior Principal
Listed on 2026-08-22
-
IT/Tech
Cloud Computing: Infrastructure & Operations, Systems Engineer, SRE/Site Reliability, Cybersecurity
Type of Requisition
Regular
Clearance Level Must Currently PossessTop Secret
Clearance Level Must Be Able To ObtainTop Secret/SCI
Public Trust/Other RequiredNone
Job FamilyIT Infrastructure and Operations
SkillsAWS Cloud Computing, Cloud Engineering, Cloud Infrastructure, Cloud Platform, Dev Ops
CertificationsNone
Experience7 + years of related experience
US Citizenship RequiredYes
Job DescriptionOUR IMPACT
Own your opportunity to support one of the largest government organizations in the nation. Make an impact by advancing the Department of War mission through secure, reliable, and scalable cloud platforms that enable mission teams to operate more effectively.
Our CompanyIron EagleX (IEX), a wholly owned subsidiary of General Dynamics Information Technology (GDIT), delivers agile IT and intelligence solutions. Combining small-team flexibility with global scale, IEX leverages emerging technologies to provide innovative, user-focused solutions that empower organizations and end users to operate smarter, faster, and more securely in dynamic environments.
Job DescriptionIron EagleX is seeking an experienced Platform Engineer Sr Principal to lead the architecture, implementation, and continuous improvement of secure, scalable cloud-native platforms.
This position requires extensive experience with Amazon Web Services, Kubernetes, cloud networking, infrastructure automation, Dev Sec Ops , and Site Reliability Engineering. The successful candidate will serve as a technical authority across multiple projects, establish platform engineering standards, and help development teams deliver secure and reliable applications more efficiently.
Experience with software development, customer-facing applications, and Internal Developer Platforms is highly desirable.
MEANINGFUL WORK AND PERSONAL IMPACTWe are seeking a senior Platform Engineering professional who will work closely with software engineers, infrastructure teams, cybersecurity personnel, program leadership, and government stakeholders to modernize enterprise platforms and improve the reliability, security, and efficiency of mission-critical systems.
The Platform Engineer Sr Principal will provide technical leadership throughout the full service lifecycle, from architecture and development through deployment, monitoring, incident response, optimization, and continuous improvement.
Job DutiesINCLUDE BUT ARE NOT LIMITED TO
- Serve as the technical authority for Platform Engineering, Dev Ops, and Dev Sec Ops architecture across multiple projects.
- Lead the design, implementation, modernization, and continuous improvement of secure, scalable, and highly available cloud-native platforms.
- Design and operate enterprise Kubernetes platforms supporting multiple applications, teams, and deployment environments.
- Architect secure AWS solutions with an emphasis on cloud networking, routing, load balancing, ingress, identity management, resiliency, and zero-trust principles.
- Capture, analyze, and implement customer requirements while ensuring alignment with mission objectives, program priorities, security requirements, and performance expectations.
- Engage throughout the full lifecycle of platform services, including design, development, deployment, operations, monitoring, incident response, and optimization.
- Design, implement, and maintain enterprise CI/CD pipelines using Git Lab or similar technologies and industry-recognized best practices.
- Define, establish, document, and maintain Dev Sec Ops standards, processes, governance, and security controls throughout the software development lifecycle.
- Design and implement Infrastructure as Code solutions using Terraform, Crossplane, and comparable technologies.
- Standardize the creation, deployment, configuration, and lifecycle management of Helm charts, Kustomize overlays, and Kubernetes-based applications.
- Develop and implement automated testing frameworks, deployment validation processes, and release strategies that improve platform reliability and reduce operational risk.
- Establish and support platform observability using metrics, logs, dashboards, alerting, tracing, and service-level objectives.
- Monitor platform availability, performance, scalability, capacity, and security while proactively identifying opportunities for improvement.
- Apply Site Reliability Engineering principles to improve system resilience, reduce operational toil, and increase service availability.
- Troubleshoot complex infrastructure, networking, security, and application issues while providing root cause analysis and implementing durable corrective actions.
- Conduct architecture reviews, infrastructure reviews, code reviews, and technical design sessions.
- Review and validate application architectures, infrastructure designs, deployment patterns, and technical implementation plans.
- Onboard application development teams to standardized cloud platforms, CI/CD pipelines, Kubernetes deployment patterns, and operational processes.
- Develop automation and platform capabilities using Python, Bash, or…
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).