04 ago
|
Medium
|
San Salvador de Hornija
04 ago
Medium
San Salvador de Hornija
ppThis is a full-time opportunity for a data/ML Engineer from LATAM. In‑person verification will be conducted. /p pIDT is an American telecommunications company founded in 1990 and headquartered in New Jersey. It is an industry leader in prepaid communication and payment services and one of the world’s largest international voice carriers. The company is listed on the NYSE, employs over 1,300 people across 20+ countries, and has revenues in excess of $1.5 billion. /p pWe are looking for a skilled Data/ML Engineer to join our BI team and take an active role in designing, building, and maintaining the end‑to‑end data pipeline, architecture, and design that powers our warehouse, LLM‑driven applications, and AI‑based BI. /p h3Responsibilities /h3 ul liDesign, develop, and maintain scalable data pipelines to support ingestion, transformation, and delivery into centralized feature stores, model‑training workflows, and real‑time inference services. /li liBuild and optimize workflows for extracting, storing, and retrieving semantic representations of unstructured data to enable advanced search and retrieval patterns. /li liArchitect and implement lightweight analytics and dashboarding solutions that deliver natural language query experience and AI‑backed insights. /li liDefine and execute processes for managing prompt engineering techniques, orchestration flows, and model fine‑tuning routines to power conversational interfaces. /li liOversee vector data stores and develop efficient indexing methodologies to support retrieval‑augmented generation (RAG) workflows.
/li liPartner with data stakeholders to gather requirements for language‑model initiatives and translate them into scalable solutions. /li liCreate and maintain comprehensive documentation for all data processes, workflows, and model deployment routines. /li liStay informed and learn emerging methodologies in data engineering, MLOps, and LLM operations. /li /ul h3Requirements /h3 ul li8+ years of experience as a Data Engineer with 2+ years focused on MLOps. /li liExcellent English communication skills. /li liEffective oral and written communication skills with the BI team and user community. /li liDemonstrated experience in utilizing Python for data engineering tasks, including transformation, advanced data manipulation, and large‑scale data processing. /li liDeep understanding of vector databases and RAG architectures, and how they drive semantic retrieval workflows. /li liSkilled at integrating open‑source LLM frameworks into data engineering workflows for end‑to‑end model training, customization, and scalable inference. /li liExperience with cloud platforms like AWS or Azure Machine Learning for managed LLM deployments. /li liHands‑on experience with big data technologies including Apache Spark, Hadoop, and Kafka for distributed processing and real‑time data ingestion.
/li liExperience designing complex data pipelines extracting data from RDBMS, JSON, API, and flat‑file sources. /li liDemonstrated skills in SQL and PL/SQL programming, with advanced mastery in Business Intelligence and data warehouse methodologies, and hands‑on experience in one or more relational database systems and cloud‑based database services such as Snowflake or Redshift. /li liUnderstanding of software engineering principles and experience working on Unix/Linux/Windows operating systems, and experience with Agile methodologies. /li liProficiency in version control systems, with experience in managing code repositories, branching, merging, and collaborating within a distributed development environment. /li liInterest in business operations and comprehensive understanding of how robust BI systems drive corporate profitability by enabling data‑driven decision‑making and strategic insights. /li /ul h3Pluses /h3 ul liExperience with vector databases such as DataStax AstraDB, and developing LLM‑powered applications using popular open‑source frameworks like LangChain and LlamaIndex – including prompt engineering, retrieval‑augmented generation (RAG), and orchestration of intelligent workflows. /li liFamiliarity with evaluating and integrating open‑source LLM frameworks – such as Hugging Face Transformers or LLaMA‑4 – across end‑to‑end workflows, including fine‑tuning and inference optimization. /li liKnowledge of MLOps tooling and CI/CD pipelines to manage model versioning and automated deployments. /li /ul pOnly accepting applicants from LATAM. /p /p #J-18808-Ljbffr
📌 Data & Machine Learning Engineer (San Salvador de Hornija)
🏢 Medium
📍 San Salvador de Hornija