AI/ML Inference Engineer PyTorch Trainium
Listed on 2026-10-06
-
Software Development
Machine Learning/ ML Engineer, AI Engineer (Applied/Software), AWS
Amazon Web Services (AWS) Annapurna Labs team seeks a Software Engineer II - AI/ML to design and optimize models for AWS Trainium accelerators. You will build distributed inference support for PyTorch, implement high-performance kernels, and collaborate with compiler, runtime, framework, and hardware teams.
You will work across the ML system lifecycle, profiling performance, applying optimizations, and enabling customers to run demanding workloads at scale on AWS accelerators.
We are currently recruiting a AI/ML Inference Engineer for PyTorch on Trainium for our team in Cupertino, CA, United States.
This is a genuine opening to take on the AI/ML Inference Engineer for PyTorch on Trainium role at Amazon Web Services (AWS).
As a AI/ML Inference Engineer for PyTorch on Trainium, you will play an important part at Amazon Web Services (AWS) in Cupertino, CA, United States.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).