$ cat jobs/software-engineer-machine-learning-infrastructure-deliveroo-6a926b412023.json
Software Engineer, Machine Learning Infrastructure
SOFTWARE ENGINEER, MACHINE LEARNING INFRASTRUCTURE - GENERATIVE AI ABOUT THE TEAM Deliveroo's GenAI Platform team sits within Machine Learning Platform and builds the shared infrastructure that helps DoorDash, Wolt, and Deliveroo teams safely bring GenAI-powered products, agents, automation, and personalization to production. Our mission is to increase the velocity of business impact from GenAI. A central pillar of that work is running frontier open-weight LLMs and VLMs (such as GLM, Qwen, Kimi, and DeepSeek) ourselves — real-time GPU serving, high-throughput batch inference, and fine-tuning on autoscaling GPUs — delivering large cost and latency wins (for example, a billion embeddings produced roughly 20× cheaper and visual models served roughly 72% cheaper). We also own core platform surfaces including the LLM Gateway, Agent Gateway, evals infrastructure, guardrails, and cost attribution. ABOUT THE ROLE You will join a small, high-leverage team building production infrastructure for Generative AI at Deliveroo and DoorDash, with a primary focus on our open-weights model platform spanning inference and fine-tuning: real-time GPU serving, high-throughput batch inference, and model fine-tuning. You’ll work across model serving and inference engines, fine-tuning and training pipelines, GPU autoscaling and utilization, batch pipelines, backend services, and observability. This role is ideal for an engineer who enjoys pushing the cost/performance frontier of GPU inference and fine-tuning in a fast-moving technical area where product needs, model capabilities, vendor ecosystems, and cost/performance tradeoffs are evolving quickly. YOU’RE EXCITED ABOUT THIS OPPORTUNITY BECAUSE YOU WILL… - Build the infrastructure that helps Deliveroo teams move GenAI ideas from prototype to production, increasing the velocity of business impact from AI across the company. - Work on our open-weights serving stack — real-time GPU endpoints, high-throughput batch inference, and fine-tuning (S
Similar remote roles
Data Scientist II
CommerceIQ · APAC · mid
Staff Software Engineer (Python)
Sia · Worldwide · senior
Machine Learning Engineer
Oak Tree Software · Worldwide · mid
Machine Learning Software Developer
Llnl · Worldwide · mid
Senior Generative AI Engineer
lifelancer · US · senior
Senior AI Engineer – GenAI & Agentic Systems
provectus · EU · senior
Applied AI Engineer
bjak · Worldwide · mid
Machine Learning Engineer, CX Intelligence
Coinbase · Worldwide · mid