Senior Data Center Operations Engineer
Listed on 2026-08-06
-
IT/Tech
Systems Engineer, IT Support
About the Organization
Colovoreis the leader in ultra-high-density, liquid-cooled AI colocation data centers for the deployment of AI inference,Super Pod, and AI factory applications and solutions. With facilities featuring power densities of 17-500+ kilowatts per cabinet, we are expanding quickly and building a national footprint of modern, hyper-efficient, high-density data centers to serve F500 enterprises and startups. The company wasacquiredby King Street in 2024, who provides real estate developmentexpertiseand secured $925 million funding to accelerate its national expansion.
Join our team and play a critical role as we scale for rapid growth. You can bea game-changerat
Colovore.
We’re seeking a Senior Engineer, Data Center Operations to lead site-level operations in a high-density, mission-critical environment. This role blends hands-on technical work with leadership in preventative maintenance, customer service, and systems optimization. The Senior Engineer ensures uptime, efficiency, and compliance while mentoring junior staff and driving continuous improvements across facilities and IT operations.
Key Responsibilities- Lead day-to-day operational support, including equipment installation, cabling, and facility rounds.
- Oversee preventative maintenance cycles and ensure compliance with operational standards and audit requirements.
- Manage and resolve escalations during incidents; act as decision-maker in high-pressure situations.
- Operate and optimize Building Management Systems (BMS), Computerized Maintenance Management Systems (CMMS), and ticketing systems.
- Configure, monitor, and tune infrastructure systems to improve reliability and efficiency.
- Mentor and train junior operations staff; ensure proper documentation and SOP adherence.
- Interface with strategic customers, providing remote hands support, escalation management, and service recovery.
- Partner with facilities, engineering and customer success teams to align on operational needs and improvements.
- Maintain accurate records, logs, and reports for compliance, audits, and performance reviews.
- Support capacity planning and operational readiness for new deployments.
- Deep understanding of electrical, mechanical, and IT infrastructure systems (power distribution, cooling, cabling, server hardware)
- Ability to diagnose, troubleshoot, and resolve complex infrastructure issues under time-sensitive conditions
- Skilled in operating and interpreting data from BMS, CMMS, ticketing, and monitoring tools
- Strong command of SOPs, MOPs, and EOPs with a disciplined approach to execution
- Expertise in designing, managing, and improving preventative maintenance programs
- Detail-oriented mindset for inspections, documentation, compliance, and audit readiness
- Ability to remain calm, clear-thinking, and decisive during outages or operational escalations
- Strong situational awareness and risk assessment capabilities to drive safe, effective resolutions
- Experience coordinating stakeholders and communicating clearly during high-pressure events
- Analytical approach to tuning and optimizing power, cooling, and monitoring systems for high-density environments
- Ability to identify inefficiencies, propose improvements, and drive implementation across teams
- Comfortable working with data to inform operational decisions and capacity planning
- Professional, customer-oriented approach when supporting remote hands, escalations, or service recovery
- Clear written and verbal communication skills, able to translate technical updates for both technical and non-technical audiences
- Skilled at balancing customer needs with site-level operational priorities
- Experience training and mentoring junior technicians, with an emphasis on safety, accuracy, and professional growth
- Ability to model operational discipline, set expectations, and ensure adherence to processes
- Strong collaborator who works well across facilities, engineering, and customer-facing teams
- 5+ years of data center or mission-critical facility operations experience, with a strong focus on hands-on field work.
- Proven expertise with power, cooling, and IT infrastructure systems in high-density environments.
- Hands-on experience with BMS, CMMS, ticketing, and monitoring tools.
- Demonstrated ability to lead incident resolution and high-pressure operational decisions.
- Prior experience mentoring or leading junior technicians preferred.
- Zero downtime achieved through proactive monitoring, preventative maintenance, and effective incident management.
- Smooth customer experience with quick, accurate, and professional handling of remote hands requests and escalations.
- High-performing systems with optimized power, cooling, and monitoring configurations that support dense AI/HPC…
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).