×
Register Here to Apply for Jobs or Post Jobs. X

Sr. Director, Design Quality & Reliability - OCI Data Center Infrastructure

Job in Madison, Dane County, Wisconsin, 53786, USA
Listing for: Oracle
Full Time position
Listed on 2026-09-09
Job specializations:
  • Engineering
    Quality Engineering, Systems Engineer
Job Description & How to Apply Below
** Job Description*
* The Sr. Director will partner closely with Construction, Supply Chain, Product Engineering, Operations, and strategic suppliers to ensure OCI infrastructure platforms consistently meet aggressive reliability, availability, and lifecycle performance objectives.

** Key Responsibilities*
* ** Build and Lead the Function*
* + Establish and scale OCI's Design Quality & Reliability organization for AI data center infrastructure.

+ Develop the strategy, operating model, governance, metrics, and execution roadmap for the function.

+ Build and lead a high-performing multidisciplinary team spanning reliability engineering, supplier quality, design assurance and validation,.

+ Define organizational processes and standards for quality and reliability across the infrastructure lifecycle.

** Design Quality & Reliability Leadership*
* + Ensure infrastructure designs meet OCI reliability, resiliency, maintainability, and lifecycle performance requirements.

+ Drive design assurance processes that validate design intent against operational requirements and long-term reliability objectives.

+ Lead cross-functional design reviews focused on reliability risk reduction, failure prevention.

+ Establish reliability engineering methodologies including FMEA, fault tree analysis, accelerated life testing, and design-for-reliability practices.

** Product Quality & Supplier Reliability*
* + Define qualification and acceptance criteria for critical infrastructure products and systems used in OCI data centers.

+ Establish product quality benchmarks and reliability performance targets, including AFR (Annualized Failure Rate), IDR, MTBF, and other key reliability indicators.

+ Develop supplier quality management frameworks and collaborate with strategic suppliers to improve product reliability and manufacturing quality.

+ Support root cause analysis and corrective action processes for field failures and reliability excursions.

** Metrics, Benchmarking & Continuous Improvement*
* + Develop KPI dashboards and measurement systems to benchmark design and product reliability performance across the OCI infrastructure portfolio.

+ Analyze field performance data, warranty trends, operational incidents, and failure modes to identify systemic improvement opportunities.

+ Establish data-driven processes to recommend and implement design, component, or supplier changes that improve quality, reliability, and operational efficiency.

+ Benchmark OCI performance against hyperscale and industry best practices.

** Cross-Functional Partnership*
* + Partner with Infrastructure Capacity Delivery, Operations, Supply Chain, and Product teams to ensure reliability objectives are embedded throughout the lifecycle.

+ Influence strategic technology and supplier selection decisions using quality and reliability data.

+ Provide executive-level reporting on reliability performance, risks, and improvement initiatives.

** Responsibilities*
* ** Required Experience*
* + 15+ years of experience in quality, reliability engineering, critical infrastructure, manufacturing quality, or related technical leadership roles.

+ 7+ years leading large-scale engineering or quality organizations.

+ Experience building or transforming quality and reliability programs in hyperscale infrastructure, cloud, semiconductor, power systems, telecom, or mission-critical environments.

+ Deep expertise in reliability engineering methodologies and statistical analysis techniques.

+ Proven experience with supplier quality management and complex hardware ecosystems.

+ Strong understanding of critical infrastructure systems including power distribution, cooling, controls, mechanical, and electrical systems.

** Preferred Qualifications*
* + Experience in hyperscale data center infrastructure or cloud infrastructure environments.

+ Familiarity with AFR, IDR, MTBF, and reliability growth methodologies.

+

Experience with GW-scale infrastructure deployment programs.

+ Demonstrated success driving measurable reliability improvements across large operational fleets.

+ Advanced degree in Engineering, Reliability Engineering, Mechanical Engineering, Electrical Engineering, or related field preferred.

**…
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary