Software Development Engineer , ML Infra Services, Annapurna Labs
Listed on 2026-07-06
-
Software Development
Cloud Engineer - Software, DevOps, AWS, Software Engineer
Annapurna Labs was a startup acquired by AWS in 2015 and is now fully integrated. If AWS is an infrastructure company, think of Annapurna Labs as the infrastructure provider of AWS. Our org spans silicon engineering, hardware design and verification, software, and operations. We've delivered AWS Nitro, ENA, EFA, Graviton, F1 EC2 Instances, AWS Neuron, Inferentia and Trainium ML Accelerators, and scalable NVMe storage.
AWS Neuron is the complete software stack for AWS Inferentia and Trainium cloud‑scale machine learning accelerators and the Trn1 and Inf1 servers that use them.
We’re looking for a Software Development Engineer to help build and evolve machine learning tools that run, optimize, and analyze ML workloads on custom AI accelerators. You'll work across the stack, from infrastructure orchestration to developer‑facing tooling – alongside hardware engineers, system architects, and ML researchers both within and outside Amazon.
Key job responsibilities- Design and implement tooling for profiling, optimization, and resource management of ML workloads on custom accelerators.
- Build high‑impact solutions that ship to a large and growing customer base.
- Participate in design discussions, code reviews, and cross‑functional collaboration with hardware, software, and customer‑facing teams.
- Create metrics, implement automation, and resolve root causes of software defects.
- Work in a startup‑like environment where you’re always focused on the most important problems.
This is a high‑impact, high‑visibility team where your work directly accelerates every Neuron team’s ability to ship, effectively multiplying the output of 100+ engineers. We are a small, senior group actively building greenfield capabilities, which means significant design ownership for SDEs and the opportunity to own major components and drive architectural decisions. You’ll work at the cutting edge of AI infrastructure, at the intersection of Kubernetes, custom silicon, and large‑scale ML workloads.
BasicQualifications
- 18 years of age or older.
- Experience with at least one modern language such as Java, Python, C++, or C# including object‑oriented design.
- Experience with at least one general‑purpose programming language such as Java, Python, C++, C#, Go, Rust, or Type Script.
- Experience with data structure implementation, basic algorithm development, and/or object‑oriented design principles.
- Proficiency in Java and at least one of Go, Python, or Type Script.
- Familiarity with Git and CI/CD pipelines.
- Experience from a technical internship.
- Experience in optimization mathematics such as linear programming and nonlinear optimization.
- Experience with distributed, multi‑tiered systems, algorithms, and relational databases.
- Experience with Cloud platforms (preferably AWS), database systems (SQL and No
SQL), AI tools for development productivity, contributing to open‑source projects, and/or version control systems. - Internship or project experience with AWS services (EKS, EC2, Lambda, S3, DynamoDB, or SQS).
- Familiarity with distributed systems or big data architectures.
- Experience with Linux systems and performance profiling.
- Exposure to compiler tool chains, code generation, or instruction set architectures (CPU, NPU, GPU).
Amazon is an equal‑opportunity employer and does not discriminate on the basis of protected veteran status, disability, or other legally protected status.
Location:
Cupertino, CA, USA
Base salary range: $ – $ USD annually.
Learn more about our benefits at (Use the "Apply for this Job" box below)..
#J-18808-Ljbffr(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).