×
Register Here to Apply for Jobs or Post Jobs. X

Sr Manager

Job in Cupertino, Santa Clara County, California, 95014, USA
Listing for: Apple Inc.
Full Time position
Listed on 2026-08-30
Job specializations:
  • IT/Tech
    Cloud Computing: Infrastructure & Operations, Systems Engineer, SRE/Site Reliability
Salary/Wage Range or Industry Benchmark: 237600 - 401700 USD Yearly USD 237600.00 401700.00 YEAR
Job Description & How to Apply Below

Cupertino, California, United States Software and Services

We are looking to hire a Senior Infrastructure, SRE & AI Platforms Manager to help set the long-term technical strategy, organizational structure, and operational roadmap for global, mission-critical infrastructure platforms on the Services Special Projects team.

This position requires a rare blend of deep technical domain expertise—spanning distributed systems, Kubernetes, and AI workload orchestration—and proven organizational leadership managing large, globally distributed engineering teams.

Description

In this role, you will be responsible for defining and building infrastructure strategy that balances continuous innovation with high reliability, performance, and cost efficiency. You will lead a growing, multi-tiered team of engineers who are responsible for foundational platforms that power large-scale consumer and enterprise workloads.

Beyond operational delivery, you will establish standards for operational excellence, Site Reliability Engineering (SRE), and capacity planning. You will be a key strategic partner, translating complex business imperatives into scalable platform designs while cultivating a strong engineering culture focused on automation, technical ownership, accountability, and continuous improvement.

Responsibilities
  • Strategic Leadership & Architecture
  • Multi-Year Roadmap & Strategy:
    Define and execute the long-term technical vision and capital investment strategy for global compute, storage, network, observability, and AI infrastructure.
  • Management & Organizational Alignment:
    Partner with Leadership to align platform capabilities, risk management, capacity investments, and architectural decisions with overarching business goals.
  • Technical Tradeoffs:
    Evaluate emerging infrastructure technologies, and make strategic platform trade-off decisions.
  • AI Compute & Modern Infrastructure Platforms
  • AI Infrastructure at Scale:
    Architect, scale, and optimize large-scale environments for training and inference, resolving complex challenges in cluster design, scheduling, interconnect performance, storage throughput, and capacity planning.
  • Hybrid & Multi-Cloud Compute:
    Oversee internal Kubernetes compute environments as well as managed public cloud platforms (AWS EKS, GCP GKE) and large bare-metal footprints to provide seamless developer experiences.
  • Data & Storage Platform Management:
    Direct the strategy and maintenance for distributed block/object storage alongside managed database and data streaming platforms (e.g., Cassandra, Foundation DB, Redis, PostgreSQL, MongoDB, Kafka).
  • Networking, Traffic & Security:
    Ensure reliable global traffic management, load balancing, cloud networking architectures, and enterprise security compliance across all environments.
  • SRE, Operational Excellence & Engineering Culture
  • Site Reliability Engineering (SRE):
    Cultivate a mature SRE culture focusing on high availability, automated fault recovery, telemetry, logging, metrics, and rigorous post-incident analysis.
  • Global Team & Leadership Development:
    Build, mentor, and lead a globally distributed organization comprising engineers, managers, and managers-of-managers across all levels (interns through senior principal staff).
  • Culture of Ownership & Automation:
    Establish an environment characterized by strong technical ownership, clear accountability, continuous operational refinement, and aggressive automation of manual processes.
Minimum Qualifications
  • MS Degree in Computer Science or related degree and 12+ years of experience of progressive engineering leadership experience building, scaling, and operating mission-critical infrastructure platforms and global services.
  • Management & Leadership Scope: 6+ years managing multi-layered engineering organizations (manager-of-managers) with a proven track record of hiring, developing, and retaining top-tier technical talent across global sites.
  • Cloud & Distributed Compute Expertise:
    Demonstrated hands-on and architectural mastery of cloud-native infrastructure, Kubernetes platform engineering, and hybrid cloud operations (AWS, GCP, private data centers).
  • Accelerated Computing & AI Infrastructure:
    Direct operational and…
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary