ML Inference Engineer - AWS Neuron & GenAI
Listed on 2026-10-06
-
Software Development
AI Engineer (Applied/Software), Machine Learning/ ML Engineer, AWS
Annapurna Labs (U.S.) Inc. in Cupertino, CA, seeks engineers to design and optimize ML models for AWS Neuron on Inferentia and Trainium accelerators, bridging software and hardware for high-performance inference and training.
You will work across frameworks, kernels and compilers, implement high-performance ML kernels, profile performance, and collaborate with customers to enable their models on AWS accelerators, in a fast-paced startup‑like culture.
This role, ML Inference Engineer - AWS Neuron & GenAI at Amazon Web Services (AWS), could be your next move.
Learn more about the ML Inference Engineer - AWS Neuron & GenAI role in the description above.
We appreciate your interest in this position.
Join Amazon Web Services (AWS) and contribute to our ongoing work.
Take a moment to read everything above and see whether this role is right for you.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).