Software Engineer, Model Inference & LLM Deployment
Listed on 2026-10-07
-
Software Development
Machine Learning/ ML Engineer, AI Engineer (Applied/Software), Software Engineer
Google Deep Mind seeks a Software Engineer for Model Inference to advance production-grade ML serving systems. You will collaborate with ML researchers and engineers to deploy large language models and optimize inference on GPUs/TPUs within Google's infra.
You'll work across model hosting, testing, and performance tuning, contributing to scalable, low-latency serving infrastructure and end-to-end deployment pipelines in a mission-driven team.
For the Software Engineer, Model Inference & LLM Deployment position at Google LLC, we are reviewing applications now.
The position is based in Mountain View, CA, United States.
This opportunity is part of our work in IT & Technology.
The advertised compensation is 180..
We aim to respond to suitable candidates as soon as possible.
Full responsibilities and requirements are described in the listing above.
Learn more about the Software Engineer, Model Inference & LLM Deployment role in the description above.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).