Senior AI/ML Inference Engineer; Neuron
Listed on 2026-10-06
-
Software Development
AI Engineer (Applied/Software), Machine Learning/ ML Engineer
Annapurna Labs (U.S.) Inc. is seeking an experienced software engineer to design, develop, and optimize ML models for AWS Neuron on Inferentia and Trainium accelerators. You will work across frameworks and kernels, collaborating with compiler, runtime, and customer teams to maximize performance.
You will lead high-performance kernel development, analyze system-wide bottlenecks, and mentor engineers in a startup-like environment with a strong focus on performance and scale.
We invite applications for the Senior AI/ML Inference Engineer (Neuron) position located in Cupertino, CA, United States.
This posting is for the Senior AI/ML Inference Engineer (Neuron) role at Amazon Web Services (AWS), based in Cupertino, CA, United States.
We are looking to fill the Senior AI/ML Inference Engineer (Neuron) position at Amazon Web Services (AWS) in Cupertino, CA, United States.
The Senior AI/ML Inference Engineer (Neuron) role at Amazon Web Services (AWS) is now open for applications in Cupertino, CA, United States.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).