04 ago
|
RemoteStar
|
España
pulliBachelors or master's degree in computer science, software engineering, or a related field /lili5+ years of professional experience in data engineering, including ownership of production data platforms or pipelines /liliExpert programming skills in Python and strong command of SQL /liliExpertise in data modeling, ETL development, and database management, with both SQL and NoSQL databases /liliHands-on experience with lakehouse architectures and columnar / open table formats (e.g., Parquet, Apache Iceberg, Delta Lake) /liliExperience with distributed data processing frameworks such as Spark, and with workflow orchestrators such as Airflow or Argo Workflows /liliStrong experience with cloud data platforms (Azure, AWS, or GCP), including object storage, containers, and Kubernetes /liliSolid grounding in data governance: catalogs, metadata, lineage, access control, and dataset versioning /liliComfortable with Git-based workflows, CI/CD, and infrastructure-as-code working models /liliExcellent problem-solving, communication, and collaboration skills; able to lead technical discussions with clients and stakeholders in English /li /ulh3Responsibilities /h3ulliOwn the end-to-end design and delivery of data platform architectures — lakehouse, data catalog, and governance — from initial scoping through production release /liliDesign,
implement, and operate large-scale ETL/ELT pipelines and workflow orchestration to ensure data is clean, accurate, versioned, and accessible /liliDefine data modeling, partitioning, schema evolution, and versioning conventions so datasets remain queryable, interoperable, and reproducible at scale /liliEstablish and maintain authoritative data catalogs, including schemas, metadata, lineage, sensitivity labels, and access policies /liliValidate released datasets against their sources for completeness, correctness, schema consistency, and query performance, defining objective acceptance criteria /liliWork closely with Machine Learning and AI Engineers to make data products directly consumable by analytics, APIs, and AI/agent workflows /liliCollaborate with clients and cross-functional teams to scope requirements, lead technical sessions, and document architectures for knowledge transfer and internal ownership /liliMentor and support other data engineers, reviewing designs and code and raising the team's engineering standards /liliStay up to date with emerging trends in data engineering — open table formats, data catalogs, orchestration — and drive their adoption where they add value /li /ul /p #J-18808-Ljbffr
📌 Senior Data Engineer (España)
🏢 RemoteStar
📍 España