Software Engineer, Machine Learning Infrastructure
Deliveroo · London - The River Building HQ
About this role
SOFTWARE ENGINEER, MACHINE LEARNING INFRASTRUCTURE - GENERATIVE AI ABOUT THE TEAM Deliveroo's GenAI Platform team sits within Machine Learning Platform and builds the shared infrastructure that helps DoorDash, Wolt, and Deliveroo teams safely bring GenAI-powered products, agents, automation, and personalization to production. Our mission is to increase the velocity of business impact from GenAI. A central pillar of that work is running frontier open-weight LLMs and VLMs (such as GLM, Qwen, Kimi, and DeepSeek) ourselves - real-time GPU serving, high-throughput batch inference, and fine-tuning on autoscaling GPUs - delivering large cost and latency wins (for example, a billion embeddings produced roughly 20× cheaper and visual models served roughly 72% cheaper). We also own core platform surfaces including the LLM Gateway, Agent Gateway, evals infrastructure, guardrails, and cost attribution. ABOUT THE ROLE You will join a small, high-leverage team building production infrastructure for Generative AI at Deliveroo and DoorDash, with a primary focus on our open-weights model platform spanning inference and fine-tuning: real-time GPU serving, high-throughput batch inference, and model fine-tuning. You’ll work across model serving and inference engines, fine-tuning and training pipelines, GPU autoscaling and utilization, batch pipelines, backend services, and observability. This role is ideal for an engineer who enjoys pushing the cost/performance frontier of GPU inference and fine-tuning in a fast-moving technical area where product needs, model capabilities, vendor ecosystems, and cost/performance tradeoffs are evolving quickly. YOU’RE EXCITED ABOUT THIS OPPORTUNITY BECAUSE YOU WILL… - Build the infrastructure that helps Deliveroo teams move GenAI ideas from prototype to production, increasing the velocity of business impact from AI across the company. - Work on our open-weights serving stack - real-time GPU endpoints, high-throughput batch inference, and fine-tuning (S
Excerpt from the posting Deliveroo published. Read the full description on their site before applying.
Apply to this role, tailored
Queue Software Engineer, Machine Learning Infrastructure at Deliveroo and I will read the posting, rewrite your resume against it, draft the cover letter, and score the fit before you send anything.
Applications finish as drafts. Nothing is sent until you read it and press send. New accounts start with 500 free credits.
More roles at Deliveroo
See all- Today
- Today
- Today
- Today
CRM Executive - Retail
London - The River Building HQ
- Today
- Today
Similar roles elsewhere
See more- 2 weeks ago
- 7 weeks ago
- 8 weeks ago
Listed from the job board Deliveroo publishes. Refolk is not the employer and does not handle their hiring. Applications go to Deliveroo directly.