Senior Data Platform / Data Engineer (Madrid)

Senior Data Platform / Data Engineer (Madrid)

04 ago
|
Straumann Group
|
Madrid

04 ago

Straumann Group

Madrid

ppStraumann Groupbr/At Straumann Group we’re on an exciting journey of growth, innovation, and impact - driven by our mission to improve oral health and transform millions of lives worldwide. United by purpose, we bring our best selves to work every day, embracing a high-performance, player-learner culture that inspires collaboration, curiosity, and ambition. Here, you’ll have the opportunity to take charge of your own career, harnessing your skills, passion, and enthusiasm for learning to continually grow and progress. Together, we’re not just shaping brighter smiles, we’re unlocking the potential of people everywhere, including our own. /p h3About The Role /h3 pWe are looking for a Senior Data Platform / Data Engineer to join our ML Platform team and help build and scale the data infrastructure that powers our AI products in dentistry. Our platform supports the full AI development lifecycle, from raw data ingestion and annotation workflows to dataset versioning and model training pipelines. You will work closely with Machine Learning Researchers (MLRs), MLOps engineers, and product teams to ensure our data infrastructure is reliable, scalable, and easy to use. /p pA key focus of the role is improving our Data Lakehouse (DLH) and dataset management workflows, including dataset versioning (DVC) and improving how data is prepared, extracted, and consumed across research and production systems. /p h3What You Will Work On /h3 pYou will play a key role in shaping the next generation of our data platform.



/p h3Typical Responsibilities Include /h3 h3Data platform ownership /h3 ul liDesign and evolve the Data Lakehouse (DLH) architecture used across our ML teams. /li liImprove the reliability and structure of data ingestion, extraction, and transformation pipelines. /li liEnsure datasets used for training and evaluation are consistent, reproducible, and well documented. /li /ul h3Dataset lifecycle management /h3 ul liImprove workflows for dataset versioning and reproducibility using tools such as DVC. /li liDesign solutions for managing multiple versions of datasets and annotations across experiments and models. /li liImprove the ability for researchers to retrieve the correct dataset versions reliably. /li /ul h3Data pipelines and infrastructure /h3 ul liBuild and maintain scalable data pipelines in Python. /li liImprove metadata management, dataset validation, and data quality monitoring. /li liOptimize data workflows across AWS-based infrastructure. /li /ul h3Collaboration with ML teams /h3 ul liWork closely with ML researchers and ML engineers to understand their data needs. /li liSupport research workflows with reliable and efficient data access patterns. /li liHelp translate research requirements into robust platform capabilities.



/li /ul h3Data governance and quality /h3 ul liImplement practices for data quality, reproducibility, and traceability across the ML lifecycle. /li liEnsure our data infrastructure meets the requirements of regulated AI development. /li /ul h3Must Have /h3 pWhat we’re looking for: /p ul liStrong Python engineering skills /li liExperience building data pipelines or data platforms /li liExperience working with AWS /li liExperience working with large datasets used in ML workflows /li liStrong software engineering practices (testing, CI/CD, documentation) /li liExperience collaborating with ML teams or working in AI environments /li /ul h3Nice To Have /h3 ul liExperience with dataset versioning tools such as DVC /li liExperience with Kubernetes /li liExperience with data lakehouse architectures /li liExperience working with annotation pipelines or ML training datasets /li liExperience with PostgreSQL, Metabase, or similar data tooling /li liExperience working in regulated environments (medical / healthcare AI) /li /ul h3Our stack /h3 ul liAWS /li liPython /li liKubernetes /li liPostgreSQL /li liMetabase /li liDVC for dataset versioning /li liInternal Data Lakehouse infrastructure /li /ul pAll qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, or disability. /p pEmployment Type: Full Time /p pAlternative Locations: Spain : Madrid /p pTravel Percentage: 0 - 10% /p pRequisition ID: 20071 /p /p #J-18808-Ljbffr

📌 Senior Data Platform / Data Engineer (Madrid)
🏢 Straumann Group
📍 Madrid

Postulate a este anuncio

Muestra tus habilidades a la empresa, rellenar el formulario y deja un toque personal en la carta, ayudará el reclutador en la elección del candidato.

Suscribete a esta alerta:

Recibe por email las nuevas ofertas de trabajo para: senior data platform / data engineer (madrid) / madrid

Suscribete a esta alerta:

Recibe por email las nuevas ofertas de trabajo para: senior data platform / data engineer (madrid) / madrid