Data & Machine Learning Engineer

Medium
Municipality of Valladolid, Spain
2 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Working hours
Regular working hours
Languages
English
Experience level
Intermediate

Job location

Municipality of Valladolid, Spain

Tech stack

Microsoft Windows
API
Agile Methodologies
Artificial Intelligence
Amazon Web Services (AWS)
Big Data
Cloud Computing
Information Engineering
Data Stores
Data Warehousing
Relational Databases
Distributed Computing Environment
Hadoop
JSON
Python
Machine Learning
Payment Service Provider
Azure
Software Engineering
PL-SQL
SQL Databases
Systems Integration
Unstructured Data
Software Repository
Data Processing
Datastax
Data Ingestion
Large Language Models
Snowflake
Prompt Engineering
Spark
Deployment Automation
HuggingFace
Kafka
Machine Learning Operations
GPT
Software Version Control
Data Pipelines

Job description

This is a full-time opportunity for a data/ML Engineer from LATAM.In?person verification will be conducted.IDT is an American telecommunications company founded in **** and headquartered in New Jersey.It is an industry leader in prepaid communication and payment services and one of the world's largest international voice carriers.The company is listed on the NYSE, employs over 1,300 people across 20+ countries, and has revenues in excess of $1.5 billion.We are looking for a skilled Data/ML Engineer to join our BI team and take an active role in designing, building, and maintaining the end?to?end data pipeline, architecture, and design that powers our warehouse, LLM?driven applications, and AI?based BI.ResponsibilitiesDesign, develop, and maintain scalable data pipelines to support ingestion, transformation, and delivery into centralized feature stores, model?training workflows, and real?time inference services.Build and optimize workflows for extracting, storing, and retrieving semantic representations of unstructured data to enable advanced search and retrieval patterns.Architect and implement lightweight analytics and dashboarding solutions that deliver natural language query experience and AI?backed insights.Define and execute processes for managing prompt engineering techniques, orchestration flows, and model fine?tuning routines to power conversational interfaces.Oversee vector data stores and develop efficient indexing methodologies to support retrieval?augmented generation (RAG) workflows.Partner with data stakeholders to gather requirements for language?model initiatives and translate them into scalable solutions.Create and maintain comprehensive documentation for all data processes, workflows, and model deployment routines.Stay informed and learn emerging methodologies in data engineering, MLOps, and LLM operations.Requirements8+ years of experience as a Data Engineer with 2+ years focused on MLOps.Excellent English communication skills.Effective oral and written communication skills with the BI team and user community.Demonstrated experience in utilizing Python for data engineering tasks, including transformation, advanced data manipulation, and large?scale data processing.Deep understanding of vector databases and RAG architectures, and how they drive semantic retrieval workflows.Skilled at integrating open?source LLM frameworks into data engineering workflows for end?to?end model training, customization, and scalable inference.Experience with cloud platforms like AWS or Azure Machine Learning for managed LLM deployments.Hands?on experience with big data technologies including Apache Spark, Hadoop, and Kafka for distributed processing and real?time data ingestion.Experience designing complex data pipelines extracting data from RDBMS, JSON, API, and flat?file sources.Demonstrated skills in SQL and PL/SQL programming, with advanced mastery in Business Intelligence and data warehouse methodologies, and hands?on experience in one or more relational database systems and cloud?based database services such as Snowflake or Redshift.Understanding of software engineering principles and experience working on Unix/Linux/Windows operating systems, and experience with Agile methodologies.Proficiency in version control systems, with experience in managing code repositories, branching, merging, and collaborating within a distributed development environment.Interest in business operations and comprehensive understanding of how robust BI systems drive corporate profitability by enabling data?driven decision?making and strategic insights.PlusesExperience with vector databases such as DataStax AstraDB, and developing LLM?powered applications using popular open?source frameworks like LangChain and LlamaIndex - including prompt engineering, retrieval?augmented generation (RAG), and orchestration of intelligent workflows.Familiarity with evaluating and integrating open?source LLM frameworks - such as Hugging Face Transformers or LLaMA?4 - across end?to?end workflows, including fine?tuning and inference optimization.Knowledge of MLOps tooling and CI/CD pipelines to manage model versioning and automated deployments.Only accepting applicants from LATAM.#J-*****-Ljbffr

Requirements

8+ years of experience as a Data Engineer with 2+ years focused on MLOps. Excellent English communication skills. Effective oral and written communication skills with the BI team and user community. Demonstrated experience in utilizing Python for data engineering tasks, including transformation, advanced data manipulation, and large?scale data processing. Deep understanding of vector databases and RAG architectures, and how they drive semantic retrieval workflows. Skilled at integrating open?source LLM frameworks into data engineering workflows for end?to?end model training, customization, and scalable inference. Experience with cloud platforms like AWS or Azure Machine Learning for managed LLM deployments. Hands?on experience with big data technologies including Apache Spark, Hadoop, and Kafka for distributed processing and real?time data ingestion. Experience designing complex data pipelines extracting data from RDBMS, JSON, API, and flat?file sources. Demonstrated skills in SQL and PL/SQL programming, with advanced mastery in Business Intelligence and data warehouse methodologies, and hands?on experience in one or more relational database systems and cloud?based database services such as Snowflake or Redshift. Understanding of software engineering principles and experience working on Unix/Linux/Windows operating systems, and experience with Agile methodologies. Proficiency in version control systems, with experience in managing code repositories, branching, merging, and collaborating within a distributed development environment. Interest in business operations and comprehensive understanding of how robust BI systems drive corporate profitability by enabling data?driven decision?making and strategic insights. Pluses Experience with vector databases such as DataStax AstraDB, and developing LLM?powered applications using popular open?source frameworks like LangChain and LlamaIndex - including prompt engineering, retrieval?augmented generation (RAG), and orchestration of intelligent workflows. Familiarity with evaluating and integrating open?source LLM frameworks - such as Hugging Face Transformers or LLaMA?4 - across end?to?end workflows, including fine?tuning and inference optimization. Knowledge of MLOps tooling and CI/CD pipelines to manage model versioning and automated deployments. Only accepting applicants from LATAM. #J-*****-Ljbffr

About the company

This is a full-time opportunity for a data/ML Engineer from LATAM. In?person verification will be conducted. IDT is an American telecommunications company founded in **** and headquartered in New Jersey. It is an industry leader in prepaid communication and payment services and one of the world's largest international voice carriers. The company is listed on the NYSE, employs over 1,300 people across 20+ countries, and has revenues in excess of $1.5 billion.

Apply for this position