Software Engineer, Model Inference, DeepMind
Listed on 2026-09-30
-
Software Development
AI Engineer (Applied/Software), Machine Learning/ ML Engineer, Software Engineer
Share Software Engineer, Model Inference, Deep Mind
Deep Mind London, UK ;
Mountain View, CA, USA
X
Note:
By applying to this position you will have an opportunity to share your preferred working location from the following:
London, UK;
Mountain View, CA, USA
.
- Bachelor’s degree or equivalent practical experience.
- 8 years of experience in software development.
- 2 years of experience in deploying and maintaining machine learning models in a live production environment.
- Experience in profiling, configuring, or executing ML workloads directly on hardware accelerators (e.g., GPU or TPU).
- Experience designing, building, or optimizing model serving infrastructure or inference backends.
- Experience with developing serving infrastructure.
- Experience programming hardware accelerators (GPUs, TPUs) via ML frameworks (e.g., JAX, PyTorch) or low-level programming models (e.g., Pallas, CUDA, OpenCL).
- Experience profiling software to identify performance bottlenecks.
- Experience with distributed ML systems optimization and parallelism (e.g., data, model, or pipeline parallelism).
- Familiarity with writing performance-optimized kernels.
- Understanding of LLM architecture and inference performance dynamics (e.g., Transformer models, memory bandwidth and compute bounds, KV cache scaling).
At Google Deep Mind our mission is to build the world's first general-purpose learning agent. Central to this mission is the complex task of measuring the intelligence of our prototypes. As a Software Engineer, you will be working with the cutting edge AI agents developed by our exceptional team of Machine Learning and Neuroscience research scientists. Your responsibilities will include everything from creating systems for agent testing using 2D and 3D games to developing test problems within physics simulators.
You will create graphical visualization of results, build competitive agent leaderboards and test new algorithms on robots. To succeed in this role you will need to have a strong foundation in software engineering and enjoy working on a wide range of challenging problems within a mission-driven team.
At Google Deep Mind our mission is to build the world's first general-purpose learning agent. Central to this mission is the complex task of measuring the intelligence of our prototypes. As a Software Engineer, you will be working with the cutting edge AI agents developed by our exceptional team of Machine Learning and Neuroscience research scientists. Your responsibilities will include everything from creating systems for agent testing using 2D and 3D games to developing test problems within physics simulators.
You will create graphical visualization of results, build competitive agent leaderboards and test new algorithms on robots. To succeed in this role you will need to have a strong foundation in software engineering and enjoy working on a wide range of challenging problems within a mission-driven team.
In this role, you will be at the forefront of bringing AI research to life. You'll work directly with researchers and engineers to optimize and deploy large language models (LLMs) like Gemini onto Google's production infrastructure, impacting users across a different range of applications. This involves a blend of technical expertise and collaborative problem-solving to ensure both efficiency and quality throughout the entire LLM deployment lifecycle.
The role includes opportunities for both IC and TL opportunities, and is open to both Software Engineering and Research Engineering backgrounds. There are opportunities across multiple teams, so applicants with both specialist and generalist interests within serving are encouraged to apply.
Ar…
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).