×
Register Here to Apply for Jobs or Post Jobs. X

Sr. Manager, DevOps & SRE – Platform Reliability & Global Operations

Job in Santa Clara, Santa Clara County, California, 95053, USA
Listing for: Socket.dev
Full Time position
Listed on 2026-07-26
Job specializations:
  • IT/Tech
    SRE/Site Reliability, Cloud Computing: Infrastructure & Operations
Salary/Wage Range or Industry Benchmark: 180000 - 260000 USD Yearly USD 180000.00 260000.00 YEAR
Job Description & How to Apply Below

Description POSITION DESCRIPTION

The Senior Dev Ops & SRE Manager – Platform Reliability & Global Operations is a senior technical leader responsible for the reliability, scalability, security, and operational excellence of a complex, multi‑platform ecosystem spanning applications, workflows, event streaming, and data platforms. This role leads a blended Dev Ops and SRE organization, manages third‑party service providers, and ensures 24x7 follow‑-the‑-sun operations. The ideal candidate has deep experience running global ‑large scale 24x7 production platforms, adhering to (service and incident) SLAs, and operating confidently during ‑off-hours‑ while delegating effectively across global teams.

LOCATION

& WORK ARRANGEMENT

Candidates must be able to work primarily within Pacific or Central Time Zone business hours to support collaboration with global teams. Employees located within 50 miles of a Qcells office (e.g., Irvine, San Francisco, Houston, or South Carolina locations) are expected to follow the company’s hybrid work policy of at least three in‑office days per week.

Responsibilities
  • Lead and scale a global, multi‑tier (L1/L2/L3) Dev Ops and SRE organization
  • Design and operate follow‑the‑-sun ‑on call‑ and support models
  • Own incident management, including Sev‑1/Sev‑2 incident command and executive communication
  • Define and operate SLOs, SLIs, and error budgets across apps, workflows, events, and data pipelines
  • Oversee Dev Ops practices for CI/CD, Kubernetes, IaC, automation, and cost optimization
  • Ensure reliable operation of event driven‑and telemetry pipelines
  • Govern and manage third‑party‑ Dev Ops and SRE vendors, including SLAs and escalations
  • Drive operational maturity: postmortems, automation, reliability improvements
  • Partner with security on secure operations, incident response, and compliance readiness
PLATFORMS IN SCOPE
  • Application Platforms:
    Kubernetes, containerized, EMS telemetry & control
  • Workflow Orchestration:
    Fleet Manager, Power Automate, cross-system‑ workflows
  • Event & Streaming:
    Microsoft Event Hub, Event streams, Kafka, RabbitMQ
  • Data & Telemetry:
    Microsoft Fabric, Kusto, PostgreSQL, Timescale DB, Cassandra
  • CI/CD & Infrastructure:
    Git Hub Actions, Jenkins, Terraform, Helm, Ansible (Azure & AWS)
  • IAM across Azure and AWS
  • Experience with Sales Force, Snowflake will be a preferred plus
TECHNICAL STRENGTHS
  • Kubernetes and container platforms in production
  • Azure (required), AWS
  • Event streaming and messaging systems
  • Data pipelines and telemetry platforms
  • Power pages, Power Automate
  • CI/CD, Infrastructure as Code, and automation
  • Observability and incident troubleshooting at scale
OPERATIONAL EXPECTATIONS
  • Escalation management for on call‑ and major incidents
  • Willingness to work off‑hours when ‑required
  • Comfortable making high impact‑ decisions under pressure
Required Qualifications
  • 15+ years in Dev Ops, SRE, Platform Engineering, or Production Operations
  • 5+ years leading globally distributed engineering teams
  • Proven ownership of 24x7, mission critical‑production platforms
  • Strong experience managing third‑party‑vendors / managed service providers
  • Deep hands‑on experience with Kubernetes, cloud platforms, and ‑event driven‑ systems
Preferred Qualifications
  • Solar Industry experience (Renewable)
  • Not requirements but good to have (optional)
USE OF AI TOOLS

As a technology organization, Qcells expects team members to leverage AI models and AI‑assisted tools in their daily workflows where appropriate. Candidates should be comfortable working in an AI‑augmented environment and applying sound judgment when using AI‑generated outputs.
During the interview process, candidates will be asked to share examples of how they have used AI tools or models in their work.
Hanwha Q CELLS America Inc. (“HQCA”) is a Qcells company, one of the world’s largest manufacturers and providers of solar photovoltaic (PV) products and solutions. Headquartered in Irvine, California, HQCA has been rapidly expanding its business in North America through the expansion of products and solutions, including distributed energy solutions, direct‑to‑homeowner solar sales and financing, and EPC services. We provide an opportunity to be part of an exciting and growing…

To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary