×
Register Here to Apply for Jobs or Post Jobs. X

Remote - Compute Engineer, Deployment

Remote / Online - Candidates ideally in
Albany, Albany County, New York, 12201, USA
Listing for: Insight Global
Remote/Work from Home position
Listed on 2026-09-05
Job specializations:
  • IT/Tech
    Systems Engineer, IT Infrastructure, SRE/Site Reliability, Network Engineer
Job Description & How to Apply Below

Compute Deployment Engineer

Our client is a leading AI infrastructure company building and operating large-scale compute environments that power next-generation AI workloads. Their teams design, deploy, and operate high-performance data center infrastructure at massive scale, with a focus on speed, reliability, and operational excellence.

They are seeking a Compute Deployment Engineer to own server and accelerator cluster bring-up from facility readiness through production deployment. This role is ideal for someone who thrives in fast-paced environments, takes ownership of large-scale deployments, and enjoys solving complex infrastructure challenges.

Responsibilities:
This Compute Engineer will bring gigawatts of accelerators from first power-on to production. Facility availability to ready-for-service across thousands of racks per site, with a new data hall landing every few weeks.

They will make rack qualification faster than the fleet grows. Firmware baselines, burn-in, and cluster validation proven on every rack before a customer workload touches it, at a pace that never becomes the critical path.

They must be able to scale by tooling, not headcount. Deployed megawatts grow several fold next year while the team stays near-flat, because anything done twice by hand becomes software.

Must own compute turn-up from facility availability to ready-for-service: the stretch after the network hands off and before customers run workloads.

Ability to qualify racks at scale: establish firmware baselines, configure BMC and BIOS, run burn-in, and validate at node and cluster level across hundreds of racks per site on GPU and custom accelerator platforms.

Drive qualification through the base-management Kubernetes platform and provisioning stack (discovery, imaging, firmware updates, shared services), burning down qual queues with tooling rather than manual runs.

Triage hardware failures found in qualification: isolate to component, drive RMA and vendor escalation, and feed failure patterns back into the qual gates.

Run turn-up remotely by default, with on-site pulses of roughly a week per data hall as new halls reach facility availability, plus occasional overlapping-site weeks.

Partner with network deployment, ICT, data center operations, and hardware teams during turn-up windows, and support incident response on freshly-live capacity.

Ability to travel 20-30% of the time to our Data Centers and Labs, as needed.

To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary