> Markdown version of [/jobs/ext/2969895-site-reliability-engineer-data-platform](https://www.wearedevelopers.com/jobs/ext/2969895-site-reliability-engineer-data-platform). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Site Reliability Engineer - Data Platform - **Company:** IMC - **Location:** Amsterdam, Netherlands - **Contract:** Permanent contract - **Skills:** Java (Programming Language), Airflow, Build Automation, Bash Shell, Computer Programming, System Configuration, Data as a Services, Data Infrastructure, Software Debugging, Linux, Distributed Data Store, Apache Hadoop, Hadoop Distributed File System, Python (Programming Language), Pcap, Performance Tuning, Reliability Engineering, Ansible, Prometheus, Standard Sql, SQL Databases, Grafana, Apache Spark, Data Lakes, Kubernetes, Infrastructure Automation Frameworks, Apache Flink, Bare Metal, Integration Frameworks, Apache Kafka, Data Management, Presto, Vertica, Puppet, JVMs - **Published:** September 18, 2026 - **Apply:** https://www.adzuna.nl/details/5887955085 ## About the Role * Strong experience managing distributed data platforms (e.g. Kafka, Hadoop, Spark, Dremio); including full installation, debugging and performance tuning * Hands on experience deploying, configuring and orchestrating software on Linux and Kubernetes, with proven ability to troubleshoot issues in both environments * Strong experience with infrastructure as code (Ansible preferred) and best practices * Proficient programming experience in Python * Ability to read, write and tune SQL queries * Comfortable reading Java source code, tuning and debugging running JVMs. * Familarity with data lakehouse technologies (e.g. Iceberg or Delta Lake) as well as query engine technologies (e.g. Dremio, Presto or Trino) * Exposure to workflow orchestration tools like Airflow or Dagster * A proactive mindset: you're not just fixing issues but preventing them. * Comfortable working across teams, with minimal oversight. Our Tech Stack: * Data Tools: Hadoop (HDFS), Kafka, Dremio, Iceberg, Clickhouse, Spark, Airflow, Flink * Infrastructure Automation: Ansible, Puppet, Kubernetes (ArgoCD, Helm, Kustomize) * Observability: Prometheus, Grafana, AlertManager * Scripting: Python, Bash, SQL * Others: PCAP infrastructure ## Description IMC operates on the cutting-edge use of technology to create a competitive edge over the competition. We also grow quick and have plenty of complex technical challenges. We're looking for an experienced SRE with strong background in managing distributed data systems on both bare metal Linux and Kubernetes. We want someone who can help us standardize deployments, elevate observability, and improve automation as we scale our data platform and other critical data services. You will join our Data Platform team, part of our local data team that builds and runs the systems that are used by traders, quant researchers and engineering teams, for all their data needs. The Data Platform team is responsible for the foundational platform that our data frameworks and tooling is built on top of. This includes observability, scalability and supporting standardised deployments. Your Core Responsibilities: As an SRE within IMC you will join a sub-team that takes a central role in all the data needs and you'll be working to: * Design, implement and operate our data platforms. * Improve observability so we catch issues before our users do. * Build automation to reduce toil and allow our systems to scale * Support and own reliability of critical services (e.g. HDFS, Kafka and Dremio) * Drive long-term architectural improvements, not just fixing issues, but preventing them. ## Related Videos - [From Messy Queries to Scalable Systems - How Data Engineering actually works](https://www.wearedevelopers.com/videos/100203-from-messy-queries-to-scalable-systems-how-data-engineering-actually-works) - [Like Spring, but Faster: The new Java Jedi](https://www.wearedevelopers.com/videos/1595-like-spring-but-faster-the-new-java-jedi) - [Docker network without Docker](https://www.wearedevelopers.com/videos/1418-docker-network-without-docker) - [Modern Data Architectures need Software Engineering](https://www.wearedevelopers.com/videos/1030-modern-data-architectures-need-software-engineering) - [Practical performance tuning for Serverless Java on AWS](https://www.wearedevelopers.com/videos/2075-practical-performance-tuning-for-serverless-java-on-aws) - [Docker exec without Docker](https://www.wearedevelopers.com/videos/1094-docker-exec-without-docker) ## Related Articles - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Find a Developer Job: 12 Best Job Sites For Developers](https://www.wearedevelopers.com/magazine/165-find-a-developer-job-12-best-job-sites-for-developers) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Where To Find Software Engineering Jobs](https://www.wearedevelopers.com/magazine/396-where-to-find-software-engineering-jobs) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers)