Doctolib

Senior Machine Learning Engineer - Orchestration - Applied AI & LLMs (x/f/m)

Paris, Paris, FranceFull timeSeniorPosted 19 days ago
Apply on Doctolib →

Sign into see who you know at Doctolib.

Set a new pulse for healthcare! We are looking for a Senior Machine Learning Engineer to join the Orchestration team in AI Health Assistant. Your mission will be to ensure our AI Health Companion behaves safely, reliably, and helpfully — for millions of patients and practitioners. You will work in a cross-functional team building robust evaluation pipelines for agentic AI systems, contributing directly to improving the quality and safety of AI-powered healthcare. Working in the tech team at Doctolib means building innovative products and features to improve the daily lives of care teams and patients. What you'll do Your responsibilities include but are not limited to: Define and own the evaluation strategy for our AI agentic system — metrics, protocols, datasets, and tooling Implement and maintain automated evaluation pipelines to monitor model quality, safety, and alignment across iterations Run systematic experiments to assess reasoning, factuality, robustness, and user experience Collaborate closely with model developers and research scientists to provide insights and drive iterative improvement Contribute to research and internal knowledge sharing on LLM evaluation methodologies and best practices Life at Doctolib Tech Our solutions are built on a single fully cloud-native platform that supports web and mobile app interfaces, multiple languages, and is adapted to country and healthcare specialty requirements. Our stack is composed of Rails, TypeScript, Java, Python, Kotlin, Swift, and React Native. We leverage AI ethically across our products to empower patients and health professionals. Discover our AI vision here. Want to learn more about our tech culture and environment? Visit the Doctolib Tech site. Who you are Before you read on: if you don't have the exact profile described below, but you feel this job description matches your skill set, we still encourage you to apply. You'll be a great fit if you: Hold an MSc or PhD in Computer Science, Machine Learning, Data Science, or a related field Have 7+ years of hands-on experience working with large language models (e.g., GPT, Claude, Llama, or BERT-like architectures) Have proven experience evaluating agentic or reasoning systems (e.g., autonomous agents, tool-using LLMs, dialogue systems, or task-oriented assistants) Have a strong track record in experiment design, metric definition, and evaluation automation Are able to bridge research and production, influencing both modeling and product decisions Have excellent communication skills, a collaborative mindset, and are fluent in English It would be fantastic if you: Have experience in the clinical or medical domain and sensitivity to ethical or regulatory challenges in healthcare AI What we offer Free comprehensive health insurance (basic package) for you and your children 25 days of paid vacation per year, plus up to 14 days of RTT Free mental health and coaching services through our partner Moka.care Work from abroad for up to 10 d...