×
Register Here to Apply for Jobs or Post Jobs. X

Test Engineer, AI Data Center Infrastructure; Austin

Job in Austin, Travis County, Texas, 78716, USA
Listing for: Celestica Inc.
Full Time position
Listed on 2026-07-19
Job specializations:
  • Software Development
Salary/Wage Range or Industry Benchmark: 120000 - 150000 USD Yearly USD 120000.00 150000.00 YEAR
Job Description & How to Apply Below
Position: Staff Test Engineer, AI Data Center Infrastructure (Austin)

10 - Staff Engineer, Software

Date: May 19, 2026

Location: Austin, TX, US

General Overview

Job Title: Staff Test Engineer, Server Compute Firmware - AI Data Center

Functional Area: Engineering (ENG)

Career Stream: Design - Software Engineering

Job Code: SEN-ENG-DSE

Job Band: 10

Direct/Indirect Indicator: Indirect

Summary

Define and implement test strategies for all storage and server hardware, firmware, and software components within the AI data center environment.

Lead the definition and development of holistic test strategies, test plans and test cases for complex data center solutions, including functional, performance, reliability, stress, and endurance testing.

Mentor and provide technical guidance to junior test engineers, fostering a culture of technical excellence and continuous improvement.

Design and implement automated test frameworks and scripts using languages like Python, Go, or similar, to improve efficiency and coverage of testing.

Conduct in-depth performance analysis and bottleneck identification for server platforms (e.g., CPU, GPU, memory, PCIe, networking), security (e.g., secure boot, Root of Trust, Platform Firmware Resilience) and OpenBMC interfaces/features.

This includes debugging issues related to BIOS, BMC functionality and its interaction with server hardware.

Develop and maintain robust testbeds and infrastructure for continuous integration and validation.

Utilize open‑source and commercial test tools relevant to server, BIOS, OpenBMC and storage validation.

Collaborate closely with hardware design, software development, infrastructure, and AI/ML engineering teams to understand requirements and integrate testing throughout the product lifecycle.

Communicate test progress, results, and critical issues effectively to stakeholders, including executive leadership.

Develop specialized test methodologies to validate performance and reliability under heavy AI/ML workloads (e.g., large model training, inference at scale, data ingestion).

Understand and test the interactions.

Required Qualifications

Bachelor's or Master's degree in Computer Science, Electrical Engineering, or a related technical field.

10+ years of experience in hardware and/or software testing, with at least 5 years focused on enterprise‑level storage and server systems.

5+ years of experience in a lead or senior technical role, mentoring junior engineers or leading test initiatives.

Deep expertise in server architectures (x86, ARM, GPU servers), CPU/memory subsystems, PCIe, and power management.

Extensive experience in server architectures (x86, ARM, GPU servers), BIOS, CPU/memory subsystems, PCIe, power management, and Baseband Management Controllers (BMC) functionality.

Strong understanding of enterprise software security (e.g., secure boot, Root of Trust, Platform Firmware Resilience).

Proficiency in scripting languages (e.g., Python, Bash) for test automation and data analysis.

Experience with Linux operating systems (e.g., Ubuntu, CentOS, RHEL) and command‑line tools.

Familiarity with networking concepts (Ethernet, TCP/IP, Infini Band) and network testing methodologies.

Experience with test methodologies such as performance testing, reliability testing, stress testing, and fault injection.

Excellent problem‑solving, analytical, and debugging skills.

Strong communication and interpersonal skills, with the ability to collaborate effectively across diverse teams.

Preferred Qualifications

Familiarity with OCP (Open Compute Project).

Experience with cloud environments (AWS, Azure, GCP) and virtualization technologies.

Knowledge of containerization technologies (Docker, Kubernetes).

Familiarity with AI/ML frameworks (e.g., Tensor Flow, PyTorch) and their infrastructure requirements.

Experience with performance profiling tools (e.g., fio, Iometer, Perf, VTune).

Contributions to open‑source projects related to storage, servers, or testing.

Certifications in relevant technologies (e.g., Net App, Dell EMC, HPE, NVIDIA).

Notes

This job description is not intended to be an exhaustive list of all duties and responsibilities of the position. Employees are held accountable for all duties of the job. Job duties and the % of time identified for any…

To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary