Manager, Cloud and DevOps Engineering
Listed on 2026-08-05
-
IT/Tech
Cloud Computing: Infrastructure & Operations, SRE/Site Reliability, Systems Engineer, IT Infrastructure
Manager, Cloud And Dev Ops Engineering
The Manager, Cloud and Dev Ops Engineering is responsible for leading the design, implementation, operation, security, automation, and continuous improvement of the organization's cloud, platform, infrastructure, Dev Ops, observability, and platform engineering capabilities. Operating within an Azure-first, Databricks-centric environment, this role provides technical and people leadership for cloud infrastructure, networking, Dev Ops, Infrastructure-as-Code (IaC), platform engineering, cloud security, observability, reliability engineering, Fin Ops, and enterprise platform operations supporting business applications, data platforms, analytics, artificial intelligence (AI), machine learning (ML), automation, and enterprise technology services.
This is a hands-on leadership role responsible for managing Cloud Engineers and Dev Ops Engineers while actively participating in architecture reviews, platform engineering initiatives, cloud modernization programs, infrastructure automation, operational excellence efforts, and enterprise technology strategy.
The Manager, Cloud and Dev Ops Engineering partners closely with Data Engineering, Enterprise & Product Engineering, Security, Compliance, Infrastructure Operations, and business stakeholders to deliver secure, scalable, resilient, highly automated, and cost-optimized technology platforms aligned with organizational objectives and regulatory requirements. The role serves as the technical owner for Azure platform services, cloud governance, Dev Ops standards, infrastructure automation, cloud security controls, networking architecture, disaster recovery capabilities, observability platforms, and enterprise cloud operating models.
Responsibilities include leading, mentoring, coaching, and developing Cloud Engineers and Dev Ops Engineers; managing staffing, workforce planning, recruiting, employee development, performance management, succession planning, and technical career growth across the Cloud and Dev Ops Engineering organization; establishing cloud engineering, platform engineering, Dev Ops, Infrastructure-as-Code, observability, and operational excellence standards across the enterprise; defining and maintaining Azure landing zone architecture, cloud governance frameworks, subscription strategies, management group structures, security baselines, and platform standards;
leading architecture reviews for cloud infrastructure, networking, Dev Ops platforms, automation solutions, security controls, data platform infrastructure, and AI platform services; serving as the technical owner for Azure cloud architecture, platform engineering, cloud networking, cloud security, Infrastructure-as-Code, and enterprise automation initiatives; establishing and governing Infrastructure-as-Code standards utilizing Terraform, Bicep, Azure Policy, policy-as-code, and automated platform provisioning frameworks; leading cloud modernization, infrastructure transformation, platform consolidation, automation, resiliency, and operational maturity initiatives;
partnering with Data Engineering teams to establish secure, scalable cloud foundations supporting Azure Databricks, Delta Lake, Unity Catalog, Azure Data Lake Storage Gen2, analytics platforms, and AI workloads; supporting cloud infrastructure requirements for machine learning, generative AI, Azure OpenAI, Azure AI Services, Databricks ML, vector databases, intelligent automation platforms, and emerging AI technologies; establishing Dev Ops standards supporting CI/CD, Git Ops, release automation, environment management, platform automation, and software delivery governance;
driving adoption of platform engineering principles, self-service infrastructure, reusable cloud services, golden templates, automation frameworks, and developer enablement capabilities; leading enterprise observability initiatives utilizing Azure Monitor, Log Analytics, Application Insights, KQL, telemetry platforms, alerting frameworks, and operational dashboards; establishing reliability engineering practices including SLOs, SLIs, capacity planning, availability management, disaster recovery, business continuity, and operational resilience; serving as escalation point for complex cloud infrastructure, networking, platform engineering, Dev Ops, automation, security, and operational incidents;
partnering with Security and Compliance teams to implement Zero Trust architecture, cloud security controls, identity governance, threat monitoring, vulnerability management, and regulatory compliance requirements; leading Fin Ops initiatives including cloud cost governance, Azure Cost Management, chargeback/showback models, optimization strategies, reserved capacity planning, and resource utilization management; evaluating emerging cloud-native technologies, AI platform capabilities, automation solutions, and engineering tools supporting organizational growth and innovation;
participating in enterprise technology strategy, roadmap…
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).