JobBobsLa recherche d'emploi en temps réelEn direct

Senior Data Engineer Python/GCP (x/f/m)

Doctolib · Paris, Paris, France · 2026-05-21

Contractexecutive
Postuler sur le site de l'employeur

À propos du poste

<h2 class="s-pb-2 s-pt-4 s-heading-xl s-text-foreground dark:s-text-foreground-night"><strong class="s-font-semibold s-text-foreground dark:s-text-foreground-night">Your Impact</strong></h2> <div class="s-whitespace-pre-wrap s-break-words s-font-normal first:s-pt-0 last:s-pb-0 s-py-1 @md:s-pt-2 @md:s-pb-[10px] @md:s-leading-relaxed s-text-base s-leading-relaxed s-text-foreground dark:s-text-foreground-night">We are looking for a <strong class="s-font-semibold s-text-foreground dark:s-text-foreground-night">Senior Data Engineer</strong> to join the <strong class="s-font-semibold s-text-foreground dark:s-text-foreground-night">AI Team</strong> working on our <strong class="s-font-semibold s-text-foreground dark:s-text-foreground-night">AI Medical Companion</strong>.</div> <div class="s-whitespace-pre-wrap s-break-words s-font-normal first:s-pt-0 last:s-pb-0 s-py-1 @md:s-pt-2 @md:s-pb-[10px] @md:s-leading-relaxed s-text-base s-leading-relaxed s-text-foreground dark:s-text-foreground-night">Your mission will be to build and optimize the data foundations that power safe, scalable, and impactful AI models. You will work on data infrastructure for <strong class="s-font-semibold s-text-foreground dark:s-text-foreground-night">LLM, VLM, and RAG-based systems</strong>, ensuring our engineers and data scientists can <strong class="s-font-semibold s-text-foreground dark:s-text-foreground-night">train, evaluate, and deploy AI models</strong> efficiently on high-quality, well-structured, and compliant data. Your work will directly support health professionals in delivering better care while improving their work-life balance, ultimately impacting 80 million patients and 400,000 healthcare professionals across Europe.</div> <div class="s-whitespace-pre-wrap s-break-words s-font-normal first:s-pt-0 last:s-pb-0 s-py-1 @md:s-pt-2 @md:s-pb-[10px] @md:s-leading-relaxed s-text-base s-leading-relaxed s-text-foreground dark:s-text-foreground-night">Working in the tech team at Doctolib means building innovative products and features to improve the daily lives of care teams and patients.</div> <h2 class="s-pb-2 s-pt-4 s-heading-xl s-text-foreground dark:s-text-foreground-night"><strong class="s-font-semibold s-text-foreground dark:s-text-foreground-night">What you'll do</strong></h2> <div class="s-whitespace-pre-wrap s-break-words s-font-normal first:s-pt-0 last:s-pb-0 s-py-1 @md:s-pt-2 @md:s-pb-[10px] @md:s-leading-relaxed s-text-base s-leading-relaxed s-text-foreground dark:s-text-foreground-night">Your responsibilities include but are not limited to:</div> <ul class="s-list-disc s-pb-2 s-pl-6 s-flex s-flex-col s-gap-1 s-text-foreground dark:s-text-foreground-night s-text-base s-leading-relaxed"> <li class="s-break-words s-text-foreground dark:s-text-foreground-night s-text-base s-leading-relaxed">Design, build, and maintain scalable data pipelines on <strong class="s-font-semibold s-text-foreground dark:s-text-foreground-night">Google Cloud Platform (GCP)</strong> for AI and machine learning use cases</li> <li class="s-break-words s-text-foreground dark:s-text-foreground-night s-text-base s-leading-relaxed">Implement data ingestion and transformation frameworks that power <strong class="s-font-semibold s-text-foreground dark:s-text-foreground-night">Retrieval systems</strong> and <strong class="s-font-semibold s-text-foreground dark:s-text-foreground-night">training datasets</strong> for LLMs and multimodal models</li> <li class="s-break-words s-text-foreground dark:s-text-foreground-night s-text-base s-leading-relaxed">Architect and manage <strong class="s-font-semibold s-text-foreground dark:s-text-foreground-night">NoSQL and Vector Databases</strong> to store and retrieve embeddings, documents, and model inputs efficiently</li> <li class="s-break-words s-text-foreground dark:s-text-foreground-night s-text-base s-leading-relaxed">Collaborate with ML and platform teams to define data schemas, partitioning strategies, and governance rules that ensure privacy, scalability, and reliability</li> <li class="s-break-words s-text-foreground dark:s-text-foreground-night s-text-base s-leading-relaxed">Integrate unstructured and structured data sources (text, speech, image, documents, metadata) into unified data models ready for AI consumption</li> <li class="s-break-words s-text-foreground dark:s-text-foreground-night s-text-base s-leading-relaxed">Optimize performance and cost of data pipelines using GCP native services (BigQuery, Dataflow, Pub/Sub, Cloud Storage, Vertex AI)</li> <li class="s-break-words s-text-foreground dark:s-text-foreground-night s-text-base s-leading-relaxed">Contribute to data quality and lineage frameworks, ensuring AI models are trained on validated, auditable, and compliant datasets</li> <li class="s-break-words s-text-foreground dark:s-text-foreground-night s-text-base s-leading-relaxed">Continuously evaluate and improve our data stack to accelerate AI experimentation and deployment</li> </ul> <h2 class="s-pb-2 s-pt-4 s-heading-xl s-text-foreground dark:s-text-foreground-night"><strong class="s-font-semibold s-text-foreground dark:s-text-foreground-night">Who you are</strong></h2> <div class="s-whitespace-pre-wrap s-break-words s-font-normal first:s-pt-0 last:s-pb-0 s-py-1 @md:s-pt-2 @md:s-pb-[10px]…

Compétences demandées

Offres similaires

Postuler sur le site de l'employeur

Votre prochain poste est déjà ici.

Parcourez les offres en direct de milliers d'employeurs, gardez celles qui méritent un second regard et laissez JobBob surveiller le reste.