×
Register Here to Apply for Jobs or Post Jobs. X

Systems Engineer, AI​/ML

Job in Richardson, Dallas County, Texas, 75080, USA
Listing for: GlobalFoundries inc.
Full Time position
Listed on 2026-08-28
Job specializations:
  • IT/Tech
    Systems Engineer, AI Engineer (Applied/Software)
Salary/Wage Range or Industry Benchmark: 106000 - 205000 USD Yearly USD 106000.00 205000.00 YEAR
Job Description & How to Apply Below
Position: Staff Systems Engineer, AI/ML

About Global Foundries:

Global Foundries is a leading full-service semiconductor foundry providing a unique combination of design, development, and fabrication services to some of the worlds most inspired technology companies. With a global manufacturing footprint spanning three continents, Global Foundries makes possible the technologies and systems that transform industries and give customers the power to shape their markets. For more information, visit

Summary of Role:

We are looking for a seasoned Staff AI/ML Systems Engineer to lead workload-driven architecture strategy across hardware and software boundaries. You will define how we study, model, and optimize AI/ML workloads for current and next-generation products, drive alignment across HW and SW engineering organizations, and serve as a technical authority on performance and architecture tradeoffs. This is a senior individual contributor role with significant cross-functional scope and organizational influence.

Essential Responsibilities:

You will own the end-to-end process of workload characterization and hardware performance analysis for AI/ML systems - from selecting the right representative workloads and defining measurement methodology, to building analytical models that project system-level KPIs against candidate architectures. Your findings will directly inform SoC architecture decisions, memory subsystem design, and HW/SW co-optimization strategy. You will lead architectural discussions with hardware teams (CPU, SoC, memory, interconnect) and software teams (compilers, runtimes, ML frameworks), serving as the connective tissue between workload reality and design decisions.

You will identify where the critical bottlenecks lie - whether in compute throughput, DRAM bandwidth, on-chip memory capacity, data movement latency, or software overhead - and build the case for specific architectural changes or optimization investments. A core part of this role is defining the performance KPI framework for AI/ML workloads across the product portfolio: what metrics matter, how to measure them accurately, how to estimate them pre-silicon, and how to use them to make architectural bets.

You will set the standard for how the team does this work and mentor more junior engineers in applying it. You will regularly present findings and recommendations to senior engineering leadership and product stakeholders. Your communication must bridge deep technical content and strategic implication - you should be as comfortable writing a one-page architectural recommendation as a detailed technical memo.

Other Responsibilities:

Perform all activities in a safe and responsible manner and support all Environmental, Health, Safety & Security requirements and programs.

Required Qualifications:

A BS or MS (MS preferred) in Electrical Engineering, Computer Engineering, Computer Science, or equivalent, with 4+ years of industry experience in systems engineering, hardware architecture, ML systems, or performance engineering, with a track record of technical leadership. Exceptional mathematical reasoning is a core requirement at this level. You should be able to derive and defend analytical performance models from first principles, reason rigorously about the numerical behavior of quantized and sparse models, construct bandwidth-latency tradeoff curves across memory hierarchy levels, and identify when an approximation in a model is safe versus misleading.

You will also be expected to evaluate the mathematical soundness of others' models and call out gaps in cross-functional reviews. Deep expertise in CPU and SoC architecture is expected - you should be fluent in how modern processors handle memory hierarchies, out-of-order execution, vector/SIMD pipelines, and power management, and understand how these interact with AI/ML workloads. You should have strong command of memory bandwidth constraints at the system level (DDR/LPDDR bandwidth, channel configuration, utilization efficiency) and know how to reason quantitatively about when workloads are memory-bound vs.

compute-bound. You have built and validated analytical performance models (roofline, bandwidth-latency, first-principles throughput models) and know their limits. You have experience with AI/ML acceleration on edge devices - NPUs, dedicated inference accelerators, DSP-based pipelines - and understand the HW/SW co-design challenges involved. Experience with model quantization, sparsity, or other efficiency techniques and their interaction with hardware capabilities is a strong plus.

Familiarity with AI compiler infrastructure is preferred and increasingly important in this role. Experience with MLIR-based tool chains, IREE, TVM, or equivalent compilation and lowering pipelines - understanding how high-level graph representations are transformed, tiled, scheduled, and lowered to hardware - will meaningfully improve your ability to engage with software teams and identify where compiler strategy and hardware architecture must be…

To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary