ML Inference Engineer - Custom Accelerator Data Plane
Listed on 2026-10-07
-
Software Development
Machine Learning/ ML Engineer, AI Engineer (Applied/Software), Data Engineering
Amazon Cupertino is seeking a Software Development Engineer for the MLIL Data Plane team to own design and implement inference data plane for large models on custom hardware. You will work on model execution, memory management, data movement, and serving integration.
You will develop and optimize compute kernels for a custom ML accelerator, validate architectures end-to-end, and build profiling infrastructure while driving performance across the stack from simulation to production on GPUs/ASICs.
This posting is for the ML Inference Engineer
- Custom Accelerator Data Plane role at Amazon, based in Cupertino, CA, United States.
We invite applications for the ML Inference Engineer
- Custom Accelerator Data Plane position located in Cupertino, CA, United States.
The following opening is for a ML Inference Engineer
- Custom Accelerator Data Plane with Amazon.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).