04 ago
|
Sabia Personal
|
San Sebastián
04 ago
Sabia Personal
San Sebastián
ppbOur client /b is a fast-growing deep-tech company founded in 2019 and recognized by CB Insights as one of the 100 most promising AI companies globally. They are the largest quantum software company in the EU, with 250+ employees worldwide and growing, delivering advanced solutions trusted by leading general enterprises across several critical industries, including finance, energy, manufacturing, telecom, and industrial sectors. /p h3Required Qualifications: /h3 ul liBachelors or master's degree in computer science, software engineering, or a related field /li li4+ years of professional experience in data engineering, including ownership of production data platforms or pipelines /li liExpert programming skills in Python and strong command of SQL /li liExpertise in data modeling, ETL development, and database management, with both SQL and NoSQL databases /li liHands-on experience with lakehouse architectures and columnar / open table formats (e.g., Parquet, Apache Iceberg, Delta Lake) /li liExperience with distributed data processing frameworks such as Spark, and with workflow orchestrators such as Airflow or Argo Workflows /li liStrong experience with cloud data platforms (Azure, AWS, or GCP), including object storage, containers, and Kubernetes /li liSolid grounding in data governance: catalogs, metadata, lineage, access control, and dataset versioning /li liComfortable with Git-based workflows, CI/CD, and infrastructure-as-code working models /li liExcellent problem-solving, communication, and collaboration skills; able to lead technical discussions with clients and stakeholders in English /li /ul h3Nice to have: /h3 ul liExperience with scientific or geospatial data formats and tooling (e.g., Zarr, NetCDF, GRIB2, xarray, H3 spatial indexing) /li liExperience preparing and serving data for LLM, RAG, or agent-based applications /li liPrevious experience in consulting or client-facing delivery teams /li /ul h3Perks and Benefits:
/h3 ul liIndefinite contract. /li liSigning bonus. /li liThey offer work visa sponsorship (If applicable) and relocation package (if applicable). /li liPrivate health insurance. /li liEligibility for educational budget according to internal policy. /li liHybrid opportunity in their offices located in San Sebastian. /li liLanguage classes and discounted lunch options. /li liA high-performance, collaborative environment, operating at pace on cutting-edge technologies. /li liCareer plan. Opportunity to learn and teach. /li /ul h3Responsibilities /h3 ul liOwn the end-to-end design and delivery of data platform architectures — lakehouse, data catalog, and governance — from initial scoping through production release /li liDesign, implement, and operate large-scale ETL/ELT pipelines and workflow orchestration to ensure data is clean, accurate, versioned, and accessible /li liDefine data modeling, partitioning, schema evolution, and versioning conventions so datasets remain queryable, interoperable, and reproducible at scale /li liEstablish and maintain authoritative data catalogs, including schemas, metadata, lineage, sensitivity labels, and access policies /li liValidate released datasets against their sources for completeness, correctness, schema consistency, and query performance, defining objective acceptance criteria /li liWork closely with Machine Learning and AI Engineers to make data products directly consumable by analytics, APIs, and AI/agent workflows /li liCollaborate with clients and cross-functional teams to scope requirements, lead technical sessions, and document architectures for knowledge transfer and internal ownership /li liMentor and support other data engineers, reviewing designs and code and raising the team's engineering standards /li liStay up to date with emerging trends in data engineering — open table formats, data catalogs, orchestration — and drive their adoption where they add value /li /ul /p #J-18808-Ljbffr
📌 Senior Data Engineer (San Sebastián)
🏢 Sabia Personal
📍 San Sebastián