×
Register Here to Apply for Jobs or Post Jobs. X

HPC Systems Administrator

Remote / Online - Candidates ideally in
Springfield, Clark County, Ohio, 45502, USA
Listing for: 5C Group
Remote/Work from Home position
Listed on 2026-10-07
Job specializations:
  • IT/Tech
    Systems Administrator, Unix/Linux, IT Infrastructure, Network Administrator
Salary/Wage Range or Industry Benchmark: 120000 - 135000 USD Yearly USD 120000.00 135000.00 YEAR
Job Description & How to Apply Below

5C DATA CENTERS

Join the Future of Digital Infrastructure

Are you a passionate HPC Systems Administrator looking to make a meaningful impact? We're building the next generation of digital infrastructure powering hyperscalers, AI innovation, and high-performance computing across North America.

JOB TITLE HPC Systems Administrator

DEPARTMENT Cloud Operations

LOCATION OH, USA

WORK ARRANGEMENT Remote

SALARY RANGE $120,000 – $135,000

Role Summary

As an HPC Systems Administrator, you will support the day-to-day operation, maintenance, and troubleshooting of HPC and AI infrastructure across data center and cloud environments.

This role focuses primarily on the infrastructure and hardware layer, including Linux operating systems, GPU servers, high-speed networking, storage connectivity, firmware, drivers, and out-of-band management.

You will work alongside senior Systems Administrators, Network Engineering, Data Center Operations, Deployment Engineering, vendors, and other technical teams to troubleshoot infrastructure issues, restore systems to service, perform routine maintenance, and improve the reliability of production HPC environments.

This position is well suited for someone with a foundation in Linux systems administration or data center infrastructure who wants to develop deeper expertise in HPC, GPU computing, high-performance networking, and large-scale AI infrastructure.

How We Work at 5C

Our core values guide how we collaborate, make decisions, support one another, and serve our customers. We're looking for people who embrace them and help us build something great.

What You Will Do HPC Infrastructure Operations
  • Support and maintain Linux-based HPC and AI computing environments, including large-scale NVIDIA GPU clusters, while troubleshooting hardware, operating system, networking, storage, driver, and firmware issues across bare-metal, virtualized, containerized, and cloud-hosted platforms.
  • Investigate and resolve infrastructure and compute-node issues, document technical findings, maintain operational runbooks and support procedures, and collaborate with senior engineers to
    ** escalate
    * * complex problems and improve overall system reliability.
GPU and Accelerated Computing Systems
  • Assist with the installation, configuration, validation, monitoring, and troubleshooting of NVIDIA GPU infrastructure, including drivers, firmware, system software, and GPU health management tools, while diagnosing performance, communication, and hardware-related issues.
  • Perform post-maintenance system validation, collect and analyze diagnostic data for GPU and infrastructure faults, and coordinate hardware replacement and RMA activities for GPUs, system boards, NICs, power supplies, and other server components.
Linux System Administration
  • Administer and support enterprise Linux environments, including system services, file systems, networking, storage, authentication, permissions, SSH, DNS, and NTP, while ensuring compliance with established security and operating system standards.
  • Troubleshoot operating system, CPU, memory, device-discovery, and kernel-related issues, and collect diagnostic data to investigate system crashes, hardware failures, and performance-related incidents.
Server Hardware, Firmware, and Out-of-Band Management
  • Troubleshoot hardware issues across enterprise server platforms, including memory, PCIe devices, GPUs, NICs, storage systems, power supplies, fans, system boards, and cabling, while supporting firmware updates for BIOS, BMC, GPUs, drives, and other infrastructure components.
  • Utilize out-of-band management tools to perform remote administration, system health monitoring, hardware inventory and log collection, and collaborate with Data Center technicians on hardware replacements, break-fix…
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary