> Markdown version of [/jobs/ext/3144682-data-scientist-recsys](https://www.wearedevelopers.com/jobs/ext/3144682-data-scientist-recsys). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Data Scientist - Recsys - **Company:** Arrise - **Location:** Álava, Spain - **Contract:** Permanent contract - **Skills:** A/B Testing, Amazon Web Services, Unit Testing, Microsoft Azure, Batch Processing, Big Data, BigQuery, Extract Transform Load (ETL), Data Transformation, Distributed Computing Environment, Data Flow Control, Apache Hadoop, Apache Hive, Python (Programming Language), PostgreSQL, Machine Learning, MongoDB, MySQL, Online Analytical Processing, NoSQL, NumPy, Online Transaction Processing, Recommender Systems, Tensorflow, Standard Sql, Data Streaming, Google Cloud, Data Ingestion, Azure Data Factory, Pytorch, Delivery Pipeline, Large Language Models, Snowflake, Apache Spark, Deep Learning, Pandas, Pytest, Containerization, AI Platforms, Scikit Learn, Information Technology, Apache Flink, Cassandra, HuggingFace, AWS Glue, Apache Kafka, Machine Learning Operations, Presto, Software Version Control, Data Pipelines, Docker, Amazon Redshift, Databricks - **Published:** September 29, 2026 - **Apply:** https://www.buscojobs.com.es/data-scientist-recsys-en-alava-ID-373477716 ## About the Role Bachelor's or Master's degree in Computer Science, Engineering, or a related field Strong Python experience with recent production use, including hands-on work with data science and machine learning libraries and frameworks (e.g., Pandas, Polars, NumPy, scikit-learn, PyTorch, TensorFlow, JAX, Hugging Face, ...) Experience building and deploying end-to-end machine learning systems on cloud AI platforms (Azure, GCP, or AWS), from ETL pipelines to deployment and monitoring, including model versioning and experiment tracking, supporting either batch or real-time workflows. Strong understanding of deep learning-based recommender systems for next-item prediction, and analogous NLP architectures that model sequential patterns and context Demonstrated experience building efficient data transformation pipelines for both transactional (OLTP) and analytical (OLAP) workloads, with strong knowledge of SQL and NoSQL databases (e.g., PostgreSQL, MySQL, Redshift, Snowflake, BigQuery, MongoDB, Cassandra) Experience with unit and integration testing (e.g., Pytest), CI/CD pipelines, and Docker-based containerization What Will Set You Up Apart Experience building large-scale recommender systems (e.g., candidate generation, ranking, retrieval, personalization). Track record of publications in deep learning at relevant conferences or journals. Experience with Azure Data Factory / AWS Glue / Google Cloud Dataflow. Experience designing and analyzing A/B tests, with a solid understanding of relevant evaluation metrics. Experience designing and implementing metadata-driven pipelines to scale automated A/B testing systems. Experience developing multi-modal models that integrate multiple data types (e.g., text, images, audio). Experience applying transformer-based models or large language models (LLMs) to recommendation or personalization tasks. Experience with distributed training, including data parallelism and model parallelism. Experience with distributed data processing and big data technologies (e.g., Spark, Hadoop, Flink, Kafka, Hive, Presto, Databricks). #J-*****-Ljbffr ## Description About UsARRISE sets the benchmark for service delivery and excellence in the iGaming industry. Playing a key role in the success of its clients, which include Pragmatic Play, a brand relied upon by the world's biggest online casinos for its cutting-edge products, ARRISE helps to deliver exceptional gaming experiences to millions of players worldwide.About UsARRISE sets the benchmark for service delivery and excellence in the iGaming industry. Playing a key role in the success of its clients, which include Pragmatic Play, a brand relied upon by the world's biggest online casinos for its cutting-edge products, ARRISE helps to deliver exceptional gaming experiences to millions of players worldwide.Our global team of over 12,000 talented and driven professionals are shaping the future of iGaming. Headquartered in Gibraltar, we have offices spanning Canada, India, the Isle of Man, Latvia, Malta, Romania, Serbia, Bulgaria, and the UAE, and more exciting destinations on the horizon.At ARRISE, we take pride in creating growth opportunities at all levels, constantly investing in our people while welcoming new colleagues and forging strategic partnerships that open new opportunities for success.To achieve this, we bet on ourselves. We know that success is a collective effort, and our team is driven by ambition, collaboration, and a shared commitment to grow and succeed - while embracing every step of the journey.Be part of the future of iGaming with 12,000 ARRISERS! See a job that excites you? Apply now, and our friendly recruitment team will connect with you soon. Your journey starts here.What You'll Be DoingDesign, implement, and optimize end-to-end recommendation pipelines, from data ingestion to model inference.Build and maintain scalable ETL pipelines to support reliable and efficient data flows.Develop, evaluate, and continuously improve ML models for recommendation systems.Research, prototype, and implement state-of-the-art (SOTA) approaches to improve recommendation quality and drive key business metrics.Scale and optimize data and model pipelines to handle large volumes of data and real-time or batch processing needs.Integrate multi-modal data (e.g., behavioral, transactional, and contextual signals) from various systems into recommendation models.Ensure robustness and stability of pipelines by implementing unit and integration tests across data, modeling, and deployment workflows.Monitor and maintain end-to-end system performance, including data pipelines, model quality, and downstream impact.Design and analyze A/B tests to evaluate model performance and support data-driven product decisions.Build dashboards and observability tools to track model metrics, system health, and business KPIs.Collaborate closely with Data Engineers, Software Engineers, and stakeholders to deliver scalable, production-ready solutions.What We Ask Of YouBachelor's or Master's degree in Computer Science, Engineering, or a related fieldStrong Python experience with recent production use, including hands-on work with data science and machine learning libraries and frameworks (e.g., Pandas, Polars, NumPy, scikit-learn, PyTorch, TensorFlow, JAX, Hugging Face, ...)Experience building and deploying end-to-end machine learning systems on cloud AI platforms (Azure, GCP, or AWS), from ETL pipelines to deployment and monitoring, including model versioning and experiment tracking, supporting either batch or real-time workflows.Strong understanding of deep learning-based recommender systems for next-item prediction, and analogous NLP architectures that model sequential patterns and contextDemonstrated experience building efficient data transformation pipelines for both transactional (OLTP) and analytical (OLAP) workloads, with strong knowledge of SQL and NoSQL databases (e.g., PostgreSQL, MySQL, Redshift, Snowflake, BigQuery, MongoDB, Cassandra)Experience with unit and integration testing (e.g., Pytest), CI/CD pipelines, and Docker-based containerizationWhat Will Set You Up ApartExperience building large-scale recommender systems (e.g., candidate generation, ranking, retrieval, personalization).Track record of publications in deep learning at relevant conferences or journals.Experience with Azure Data Factory / AWS Glue / Google Cloud Dataflow.Experience designing and analyzing A/B tests, with a solid understanding of relevant evaluation metrics.Experience designing and implementing metadata-driven pipelines to scale automated A/B testing systems.Experience developing multi-modal models that integrate multiple data types (e.g., text, images, audio).Experience applying transformer-based models or large language models (LLMs) to recommendation or personalization tasks.Experience with distributed training, including data parallelism and model parallelism.Experience with distributed data processing and big data technologies (e.g., Spark, Hadoop, Flink, Kafka, Hive, Presto, Databricks).#J-*****-Ljbffr ## Related Videos - [MySQL Protocol Features You Should Be Aware Of](https://www.wearedevelopers.com/videos/100267-mysql-protocol-features-you-should-be-aware-of) - [Leveraging Real time data in FSIs](https://www.wearedevelopers.com/videos/806-leveraging-real-time-data-in-fsis) - [Vectorize all the things! Using linear algebra and NumPy to make your Python code lightning fast.](https://www.wearedevelopers.com/videos/562-vectorize-all-the-things-using-linear-algebra-and-numpy-to-make-your-python-code-lightning-fast) - [Data Science in Retail](https://www.wearedevelopers.com/videos/586-data-science-in-retail) - [Coding for Good: Achieving social change with an app](https://www.wearedevelopers.com/videos/1645-coding-for-good-achieving-social-change-with-an-app) - [How We Built a Machine Learning-Based Recommendation System (And Survived to Tell the Tale)](https://www.wearedevelopers.com/videos/752-how-we-built-a-machine-learning-based-recommendation-system-and-survived-to-tell-the-tale) ## Related Articles - [Coffee with Developers - Maria Apazoglou - Making AI understandable for all in production](https://www.wearedevelopers.com/magazine/475-coffee-with-developers-maria-apazoglou-making-ai-understandable-for-all-in-production) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Top Big Data Technologies That You Need to Know](https://www.wearedevelopers.com/magazine/108-top-big-data-technologies-that-you-need-to-know) - [Résumé-Driven Development: How IT trends affect the job market for software developers](https://www.wearedevelopers.com/magazine/59-resume-driven-development-how-it-trends-affect-the-job-market-for-software-developers) - [Making Data Warehouses Fast: A Developer’s Story](https://www.wearedevelopers.com/magazine/107-making-data-warehouses-fast-a-developer-s-story) - [The 13 Best Python Libraries for Developers in 2025](https://www.wearedevelopers.com/magazine/371-the-13-best-python-libraries-for-developers-in-2025)