Senior ML Inference Engineer Custom Hardware
Listed on 2026-10-06
-
Software Development
AI Engineer (Applied/Software), AWS, Machine Learning/ ML Engineer
Annapurna Labs (U.S.) Inc. is seeking a Senior Software Development Engineer to own the design and implementation of the inference data plane for large models.
You will build software for efficient execution on custom hardware, covering model execution, memory management, data movement, and serving integration. Responsibilities include developing compute kernels, validating end-to-end LLM architectures, and integrating backends into ML serving frameworks.
Join us at Amazon Web Services (AWS) as our next Senior ML Inference Engineer for Custom Hardware in Cupertino, CA, United States.
We appreciate your interest in this position.
Join Amazon Web Services (AWS) and contribute to our ongoing work.
Take a moment to read everything above and see whether this role is right for you.
This posting is for the Senior ML Inference Engineer for Custom Hardware role at Amazon Web Services (AWS), based in Cupertino, CA, United States.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).