> Markdown version of [/jobs/ext/1928083-machine-learning-engineer-madrid](https://www.wearedevelopers.com/jobs/ext/1928083-machine-learning-engineer-madrid). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Machine Learning Engineer, Madrid - **Company:** Atos SE - **Location:** Madrid, Spain - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Confluence, JIRA, Linux, Distributed Systems, Github, Hardware Design, Monitoring of Systems, InfiniBand, Python (Programming Language), Linux System Administration, Machine Learning, Scrum Methodology, Tensorflow, Prometheus, Software Engineering, Supercomputing, Graphics Processing Unit (GPU), High Performance Computing, Pytorch, System Availability, Model Validation, Reliability of Systems, Gitlab, Git, Containerization, Scikit Learn, Kubernetes, Information Technology, Data Analytics, Machine Learning Operations, Software Version Control, Workday - **Published:** August 5, 2026 - **Apply:** https://www.buscojobs.com.es/machine-learning-engineer-madrid-en-madrid-ID-366051649 ## About the Role Education: Masters or PhD in Computer Science, Artificial Intelligence, Data Science, Telecommunications or a related field. Skills Competencies: Strong experience with Machine Learning and Deep Learning frameworks (e.G., TensorFlow, PyTorch, Scikit-learn). Experience working with time-series data and anomaly detection models. Proficiency in Python and data science ecosystems. Experience with Prometheus or similar monitoring/telemetry systems. Familiarity with containerization and orchestration technologies, especially Kubernetes. Experience building production-grade ML pipelines. Experience handling large-scale monitoring and operational datasets. Understanding of distributed systems and infrastructure monitoring. Knowledge of HPC environments, GPUs, and high-speed interconnects (e.G., Infiniband) is highly desirable. Proficiency with Git-based version control systems (GitHub, GitLab). Solid experience working in Linux environments. Good understanding of Scrum methodology and experience with Jira and Confluence. Location: Spain Benefits: - Flexible Work Schedule: Half day Fridays and an intensive summer workday supporting work life balance. - Learning and Growth: Opportunities to work with advanced AI technologies in an innovative and supportive R D environment. Join us! Here, your ideas, your curiosity and your technical excellence directly shape the next era of advanced computing - unlocking enterprise value, accelerating scientific progress and driving positive impact for society. Python, Machine Learning, TensorFlow, PyTorch, Scikit ## Description Machine Learning EngineerAbout Bull Bull is a story.One with a century of European innovation and a working environment where experts design powerful, sustainable, and sovereign digital solutions, enabling states and industries to retain full control over their data and their AI.Bull is also thousands of engineers, researchers and passionate tech people shaping the future of high-performance computing, AI, and quantum technologies.Every day, our teams push the boundaries of what is technologically possible - from next-generation HPC architectures to exascale supercomputers - supported by world-class R D, more than 1,600 patents, and unique end-to-end capabilities spanning hardware design, software engineering, data science and quantum research.We are a people-centric, innovation-driven company, where collaboration spans Europe, the Americas and India.We share a common vision of a responsible and sustainable innovation that delivers concrete impact for our customers.Role description: We are searching for a Machine Learning Engineer to join Bulls innovative R D team and contribute to the development of AI-driven solutions for infrastructure monitoring, reliability, and cybersecurity in High-Performance Computing (HPC) environments.The role focuses on leveraging large-scale operational telemetry, metrics, and logs to build predictive capabilities that improve system availability, detect anomalies, and support proactive operations.The selected candidate will be responsible not only for model development but also for rigorous validation, operationalization, and integration within a Kubernetes-based platform.Responsibilities: Design and develop ML/DL models for predicting hardware failures and detecting software or behavioral anomalies in HPC systems.Apply advanced analytics techniques such as time-series forecasting, anomaly detection, classification, and predictive maintenance using large-scale monitoring data.Build and maintain data pipelines and features from infrastructure telemetry and logs.Perform rigorous model validation to ensure robustness, reliability, and production readiness.Deploy and operationalize models within a Kubernetes-based environment, including scalable inference services and lifecycle management.Contribute to AI-driven cybersecurity use cases, such as detecting abnormal behaviors, potential intrusions, or security-related anomalies in infrastructure and system activity.Work within an Agile/Scrum environment, participating in sprint planning, stand-ups, and retrospectives.Collaborate with system administrators, support teams, and data engineers to translate operational challenges into data-driven solutions that enhance system reliability and automation.Education: Masters or PhD in Computer Science, Artificial Intelligence, Data Science, Telecommunications or a related field.Skills Competencies: Strong experience with Machine Learning and Deep Learning frameworks (e.G., TensorFlow, PyTorch, Scikit-learn).Experience working with time-series data and anomaly detection models.Proficiency in Python and data science ecosystems.Experience with Prometheus or similar monitoring/telemetry systems.Familiarity with containerization and orchestration technologies, especially Kubernetes.Experience building production-grade ML pipelines.Experience handling large-scale monitoring and operational datasets.Understanding of distributed systems and infrastructure monitoring.Knowledge of HPC environments, GPUs, and high-speed interconnects (e.G., Infiniband) is highly desirable.Proficiency with Git-based version control systems (GitHub, GitLab).Solid experience working in Linux environments.Good understanding of Scrum methodology and experience with Jira and Confluence.Location: Spain Benefits: - Flexible Work Schedule: Half day Fridays and an intensive summer workday supporting work life balance.- Learning and Growth: Opportunities to work with advanced AI technologies in an innovative and supportive R D environment.Join us!Here, your ideas, your curiosity and your technical excellence directly shape the next era of advanced computing - unlocking enterprise value, accelerating scientific progress and driving positive impact for society.Python, Machine Learning, TensorFlow, PyTorch, Scikit ## Related Videos - [How a Small Team Shrank a Microsoft Monorepo by 94%](https://www.wearedevelopers.com/videos/1236-how-a-small-team-shrank-a-microsoft-monorepo-by-94) - [Improving quality with Agentic AI with Rovo Dev and Xray](https://www.wearedevelopers.com/videos/2005-improving-quality-with-agentic-ai-with-rovo-dev-and-xray) - [Docker network without Docker](https://www.wearedevelopers.com/videos/1418-docker-network-without-docker) - [Introduction to Azure Machine Learning](https://www.wearedevelopers.com/videos/368-introduction-to-azure-machine-learning) - [Git for Code Reviews](https://www.wearedevelopers.com/videos/429-git-for-code-reviews) - [Developer Experience, Platform Engineering and AI powered Apps](https://www.wearedevelopers.com/videos/990-developer-experience-platform-engineering-and-ai-powered-apps) ## Related Articles - [Coffee with Developers - Maria Apazoglou - Making AI understandable for all in production](https://www.wearedevelopers.com/magazine/475-coffee-with-developers-maria-apazoglou-making-ai-understandable-for-all-in-production) - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [MLops – Deploying, Maintaining And Evolving Machine Learning Models in Production](https://www.wearedevelopers.com/magazine/115-mlops-deploying-maintaining-and-evolving-machine-learning-models-in-production) - [MLOps – What’s the deal behind it?](https://www.wearedevelopers.com/magazine/125-mlops-what-s-the-deal-behind-it) - [MLOps And AI Driven Development](https://www.wearedevelopers.com/magazine/82-mlops-and-ai-driven-development) - [Top-Paying Tech Jobs (with Salaries)](https://www.wearedevelopers.com/magazine/372-top-paying-tech-jobs-with-salaries)