LLM Inference Systems Engineer — Custom Accelerator
Listed on 2026-10-06
-
Software Development
AI Engineer (Applied/Software)
Annapurna Labs (U.S.) Inc. in Cupertino, CA is seeking a Software Development Engineer to own the design and implementation of our inference data plane for large models running on custom hardware.
You will shape model execution, memory management, data movement, and serving integration across the full inference path. You will write and optimize low-level code, validate architectures end-to-end, and build profiling and test infra to drive performance across the stack, from simulation to hardware
We have an opening for a LLM Inference Systems Engineer — Custom Accelerator in Cupertino, CA, United States within IT & Technology.
The following position is for a LLM Inference Systems Engineer — Custom Accelerator with Amazon Web Services (AWS).
Our team is growing, and we are hiring a LLM Inference Systems Engineer — Custom Accelerator in Cupertino, CA, United States.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).