×
Register Here to Apply for Jobs or Post Jobs. X

AI Accelerator Software Principal Engineer – NPU Full-Stack Integration

Job in Portland, Multnomah County, Oregon, 97201, USA
Listing for: Ampere
Full Time position
Listed on 2026-07-10
Job specializations:
  • Software Development
    AI Engineer (Applied/Software), Software Engineer
Salary/Wage Range or Industry Benchmark: 195000 - 292000 USD Yearly USD 195000.00 292000.00 YEAR
Job Description & How to Apply Below

AI Accelerator Software Principal Engineer – NPU Full-Stack Integration

Invent the future with us. Ampere is a semiconductor design company for a new era, leading the future of computing with an innovative approach to CPU design focused on high-performance, energy efficient AI compute. As a pioneer in the new frontier of energy efficient high-performance computing, Ampere is part of the Softbank Group of companies driving sustainable computing for AI, Cloud, and edge applications.

Join us at Ampere and work alongside a passionate and growing team - we'd love to have you apply!

As an AI Accelerator Software Principal Engineer – NPU Full-Stack Integration, you will lead the design and delivery of high-performance, low-latency deep learning inference solutions on the Arm® Ethos™-U85. You'll help advance Ampere's AI software stack by enabling models with performance and efficiency requirements.

You will operate at the intersection of software engineering, performance engineering, and hardware-aware optimization, contributing to the full stack from model execution to accelerator-ready kernel performance.

What You'll Achieve:

  • End-to-end deep learning performance acceleration Go deep into the full software/hardware execution stack, including:
    • framework integration layers
    • compiler and graph/runtime support
    • runtime libraries and user-mode execution paths
    • compute kernel development
    • profiling, benchmarking, and performance tuning
  • Model enablement with quality and speed Improve both performance and accuracy for models using popular frameworks, helping deliver production-ready inference behavior in edge devices.
  • Hardware/software co-design and optimization Partner with hardware and platform teams to co-optimize AI execution for better outcomes:
    • increased throughput
    • reduced latency
    • improved scalability
    • better resource utilization (compute/memory/IO)
    • higher sustained performance under realistic workloads
  • Build state-of-the-art AI software components Contribute to the development of software and hardware AI co-processors/accelerators, delivering reusable libraries, optimized execution paths, and robust integration with existing tooling.
  • Cross-functional collaboration Work closely with cross-functional teams (compiler/runtime, kernels, platform, and product engineering) to integrate AI capabilities into Ampere's cloud-native processor platforms and accelerators.

About You:

  • Education and Experience:

    BS Computer Science, Computer Engineering, Electrical Engineering, or Software Engineering or related technical field & 8 years of related experience; or MS degree & 6 years; or PhD & 3 years.
  • Understands AOT (Ahead-Of-Time) compilation path in popular frameworks like PyTorch and deployment path like execu

    Torch in edge environment
  • Linux + accelerator/runtime expertise (preferred):
    Experience with developing user-mode drivers, runtime libraries, or low-level integration for GPUs or deep learning accelerators in Linux is a plus.
  • Strong systems programming & performance skills:
    • Expert in Python and C/C++
    • Strong background in performance profiling and tuning (latency/throughput, memory behavior, kernel efficiency)
  • Deep ML understanding:
    Solid understanding of AI/ML concepts including neural networks and data processing frameworks. Experience with modern deep model architectures such as Transformers and Diffusion models is preferred.
  • Modern AI tooling fluency (preferred):
    Fluent with modern AI programming tools such as Codex or Claude Code, and comfortable accelerating development workflows.

What We'll

Offer:

At Ampere we believe in taking care of our employees and providing a competitive total rewards package that includes base pay, cash long-term incentive, and comprehensive benefits. The full base pay range for this role is between $182,000 and $273,000, except in the San Francisco Bay Area where the range is between $195,000 and $292,000. Our benefits include health, wellness, and financial programs that support employees through every stage of life.

Benefit highlights include:

  • Premium medical insurance, dental insurance, vision insurance, as well as income protection and a 401K retirement plan, so that you can feel secure in your health and…
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary