Senior Manager, Facilities Remote Operations Center
Dallas, Dallas County, Texas, 75201, USA
Listed on 2026-08-22
-
Management
Operations Management
Senior Manager, Facility Remote Operations Center
Crusoe is building the energy and infrastructure foundation for the AI era. We design, build, and operate purpose-built, AI-optimized data centers — from gigawatt-scale campuses like Stargate in Abilene, Texas, to modular and distributed deployments — powered by an energy-first approach that pairs compute with reliable, low-cost power. Our Facility Remote Operations Center is the nerve center for this footprint: a 24/7 command function that provides centralized visibility, monitoring, and rapid response across Crusoe's critical infrastructure nationwide.
Crusoe is seeking a Senior Manager to lead the Facility Remote Operations Center (ROC) based in Dallas, Texas. This role owns the people, processes, and technology behind 24/7/365 remote monitoring of critical facility infrastructure — electrical, mechanical, fire/life safety, and building automation systems — across Crusoe's national data center portfolio.
The Senior Manager will build and lead a team of shift supervisors and operations specialists who serve as the first line of detection, triage, and escalation for facility events, working in close coordination with on-site Data Center Operations, Engineering, and Critical Facilities teams. This is a highly visible role that blends operational leadership, technical fluency in MEP (Mechanical, Electrical, Plumbing) systems, and program-building responsibility, as the ROC scales alongside Crusoe's rapidly growing AI infrastructure footprint.
Key Responsibilities
- Lead 24/7/365 ROC operations, including staffing, scheduling, shift coverage, and performance management for a team of remote operations supervisors and specialists.
- Establish and continuously improve monitoring protocols, escalation procedures, and standard operating procedures (SOPs) for facility alarms, events, and anomalies across the fleet.
- Serve as the senior escalation point for critical facility events outside of normal parameters, coordinating real-time response with on-site teams, vendors, and leadership.
- Drive a culture of accountability, urgency, and continuous improvement within the ROC team.
- Maintain deep working knowledge of MEP systems within AI data centers, including electrical distribution (utility feeds, switch gear, generators, UPS, PDUs, busway), mechanical/cooling systems (CRAH/CRAC units, chillers, cooling towers, liquid cooling/CDUs, air handling), fire detection/suppression, and building management systems (BMS/DCIM/EPMS).
- Partner with Facility/Critical Infrastructure Engineering to understand system design intent, sequences of operation, and normal vs. abnormal operating conditions for each site.
- Ensure the ROC's monitoring platforms (DCIM, EPMS/BMS, ticketing, and alarm management tools) are correctly configured, integrated, and actionable for new and existing sites.
- Own incident management processes for facility-impacting events detected remotely, including detection, notification, escalation, and post-incident documentation.
- Lead or support root cause analysis (RCA) and after-action reviews for significant events, driving corrective and preventive actions.
- Develop and maintain emergency response runbooks and escalation matrices in partnership with site teams, ensuring readiness for weather events, utility disturbances, and equipment failures.
- Act as the connective tissue between remote monitoring and on-site Data Center Operations, Critical Facilities Engineering, Construction/Commissioning, and Security teams.
- Support commissioning and turnover of new sites into the ROC monitoring scope as Crusoe's footprint expands.
- Report on ROC performance, uptime-impacting events, and trends to senior leadership.
- Recruit, train, and develop ROC staff, building technical competency in MEP systems and monitoring tools.
- Build training programs and certification paths for remote operations specialists.
- Identify opportunities for automation, tooling improvements, and process standardization to improve detection speed and reduce false-positive alarm fatigue.
Minimum Qualifications
- 5+ years of experience in Data Center Operations, with direct responsibility for critical facility uptime.
- Strong…
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).