GenAI Infra Engineer: GPU Serving Weights
Listed on 2026-10-10
-
Software Development
AI Engineer (Applied/Software), Machine Learning/ ML Engineer, Backend Developer
Location: Greater London
Deliveroo's GenAI Platform team is hiring a Software Engineer to build production infrastructure for open-weights model serving, inference pipelines, and fine-tuning. You will work across backend services, GPU autoscaling, and observability to deliver scalable, cost-aware AI platforms for internal and partner teams.
Ideal candidates have 3+ years in software engineering, strong Python and distributed systems skills, and hands-on experience with ML infrastructure at scale and production-grade
Step into the GenAI Infra Engineer:
Real-Time GPU Serving & Open-Weights role at deliveroo in Greater London, England, United Kingdom and grow with us.
If your experience matches this role, we encourage you to apply.
All applications are reviewed carefully by our team.
The position is based in Greater London, England, United Kingdom.
This opportunity is part of our work in IT & Technology.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).