Software Engineer - Server Infrastructure & Automation
Listed on 2026-09-02
-
IT/Tech
IT Infrastructure, Unix/Linux, Systems Administrator, Systems Engineer
Software Engineer - Server Infrastructure & Automation Fort Worth:
2650 Lou Menk Dr, Fort Worth, TX 76131
Topeka:100 NE Jefferson St, Topeka, KS 66607
Position SummaryWe are seeking a highly skilled Software Engineer responsible for the engineering, lifecycle management, automation, and operational support of enterprise server infrastructure platforms. This role owns the underlying server hardware, firmware, drivers, and base operating system environments that power critical enterprise services across multiple data centers.
The ideal candidate possesses a strong understanding of server architecture and low-level systems, including hardware components, firmware, device drivers, and operating system kernels. This individual will leverage software engineering practices, Infrastructure-as-Code, and automation to streamline infrastructure operations, reduce manual effort, improve reliability, and accelerate hardware lifecycle management.
This role is focused on the foundation of the infrastructure stack, including physical servers, data center operations, operating systems, and hardware support processes. It does not focus on managing application platforms, PaaS services, or cloud overlay technologies.
Key Responsibilities- Serve as the technical owner for enterprise server infrastructure platforms throughout their entire lifecycle, including evaluation, deployment, maintenance, refresh, and retirement.
- Manage and support physical server infrastructure across multiple data center locations.
- Develop automation solutions to reduce manual operational activities related to server deployment, provisioning, maintenance, monitoring, and hardware replacement.
- Build and maintain Infrastructure-as-Code and configuration management solutions to standardize infrastructure operations and lifecycle activities.
- Monitor server health and proactively identify potential hardware failures, performance degradation, and infrastructure risks.
- Lead troubleshooting and resolution efforts for complex hardware, firmware, driver, operating system, and platform-related issues.
- Develop automated workflows to detect, triage, and elevate hardware failures to remote hands and data center support personnel.
- Coordinate with data center operations teams, remote technicians, vendors, and support providers to resolve infrastructure incidents and perform hardware replacements.
- Manage server firmware, BIOS, driver, and operating system lifecycle activities, ensuring platform stability, security, and supportability.
- Establish standards and engineering practices for server platform deployment, hardware management, and operational support.
- Partner with network, security, storage, and application teams to ensure platform reliability and operational excellence.
- Create and maintain technical documentation, operational procedures, architecture standards, and support runbooks.
- Drive continuous improvement initiatives focused on automation, operational efficiency, resiliency, and infrastructure modernization.
- Bachelor's degree in Computer Science, Engineering, Information Technology, or equivalent practical experience.
- 10+ years of experience supporting enterprise server infrastructure in large-scale production environments.
- Strong understanding of server hardware architecture, including compute, memory, storage, networking, and out-of-band management technologies.
- Experience troubleshooting hardware failures and coordinating repairs, replacements, and vendor support activities.
- Deep understanding of system firmware, BIOS/UEFI, device drivers, and operating system kernel fundamentals.
- Experience supporting enterprise Windows and Linux base operating system environments.
- Strong software development and automation experience using technologies such as Power Shell, Python, Go, C#, or similar languages.
- Experience implementing Infrastructure-as-Code and automation frameworks such as Terraform, Ansible, Bicep, Puppet, Chef, or similar technologies.
- Experience managing server lifecycle activities, including hardware refreshes, firmware management, operating system upgrades, and decommissioning.
- Experience working in enterprise data center environments and coordinating with remote hands support organizations.
- Strong analytical and troubleshooting skills with the ability to diagnose complex issues across hardware, firmware, drivers, operating systems, and infrastructure components.
- Familiarity with monitoring, alerting, and observability platforms.
- Excellent communication, documentation, and collaboration skills.
- Experience with large-scale enterprise data centers and mission-critical infrastructure environments.
- Experience developing self-healing infrastructure and automated remediation workflows.
- Experience integrating infrastructure operations with ITSM and incident management platforms.
- Knowledge of server vendor management platforms such as Dell iDRAC, HPE iLO, Lenovo XClarity, or similar technologies.
- Experience with Git, CI/CD pipelines, and…
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).