Manager, Sre & Platform Engineering
Listed on 2026-07-18
-
IT/Tech
SRE/Site Reliability, Systems Engineer, Cloud Computing: Infrastructure & Operations, IT Infrastructure
Locations
- Plano, Texas:
Hybrid schedule with mandatory in-office three days on Tuesday, Wednesday, and Thursday. - Pune, India:
Hybrid schedule with mandatory in-office four days on Monday, Tuesday, Wednesday, and Thursday.
The Manager, SRE & Platform Engineering reports to the Director of Platform Engineering and provides technical leadership for Armor’s SRE and infrastructure engineering team. This position is responsible for the operational reliability, availability, and performance of Armor’s production infrastructure, including managed security services, enterprise cloud, and MDR platforms. The role exercises independent judgment and discretion in directing incident response, establishing reliability strategy, making infrastructure architecture decisions, and developing engineering talent.
This is a hands‑on technical leadership position that combines people management with direct operational accountability.
- Lead, hire, develop, and manage the SRE and infrastructure engineering team, including setting technical direction, conducting performance evaluations, and building engineering capability.
- Own operational coverage including on-call rotations, incident command, escalation procedures, and blameless postmortem processes to drive continuous improvement and reduce incident recurrence.
- Drive automation of infrastructure operations across CI/CD pipelines, infrastructure‑as‑code, self‑healing systems, and automated remediation to systematically reduce manual operational burden.
- Manage and improve infrastructure spanning VMware/Proxmox private cloud, public cloud platforms (AWS, Azure, GCP), and hybrid environments.
- Own monitoring, alerting, and observability across the production environment, including designing actionable dashboards, meaningful alert thresholds, and operational runbooks.
- Plan and execute infrastructure migrations, platform transitions, and hardware refresh programs in coordination with the Director of Platform Engineering, ensuring minimal customer disruption.
- Maintain compliance and audit readiness (PCI‑DSS, HIPAA, SOC
2) as an integrated operational discipline across all infrastructure operations. - Define and track SLIs/SLOs that reflect customer impact. Systematically identify and reduce operational toil.
- Implement AI‑assisted operations including automated triage, root cause analysis, predictive alerting, and intelligent escalation to improve mean time to resolution.
- Coordinate with the Director of Platform Engineering and Product Engineering teams on release readiness, production handoffs, change management, and capacity planning.
- 8+ years of experience in SRE, Dev Ops, or Infrastructure Engineering in production environments, including 2+ years leading or managing engineering teams.
- Hands‑on production experience with Kubernetes and container orchestration, cloud infrastructure (AWS preferred; Azure or GCP acceptable), Infrastructure as Code (Terraform or equivalent), and CI/CD/Git Ops practices.
- Strong automation and observability experience using scripting (Python, Bash, or Power Shell) and monitoring platforms such as Datadog, Prometheus, Grafana, or equivalent.
- VMware vSphere, Proxmox, NSX‑T, Zerto, and Rubrik (expected to learn within first six months).
- Advanced networking (DNS, VPNs, firewalls, load balancing).
- Git proficiency and AI‑assisted development tools (Git Hub Copilot, Claude Code, or similar).
- Bachelor's degree in Computer Science, Information Technology, or equivalent experience.
- Experience operating in regulated environments with compliance and audit responsibilities (PCI‑DSS, HIPAA, SOC 2, or similar), including supporting infrastructure through audit readiness and compliance activities.
The work environment characteristics described here are representative of those an employee encounters while performing the essential functions of this job. The noise level in the work environment is usually low to moderate. The work environment can be either in an office setting or remotely from anywhere.
Equal opportunity employerIt is the policy of the company to comply with all employment laws and to afford equal employment opportunity to individuals in all aspects of employment, including in selection for job opportunities, without regard to race, color, religion, sex, national origin, age, disability, genetic information, veteran status, or any other consideration protected by federal, state or local laws.
#J-18808-Ljbffr(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).