Market Data Engineer (Domain Trading Expertise Required)

Bhft
Barcelona, Spain
14 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience required
5 years minimum
Working hours
Regular working hours
Languages
English

Tech stack

Airflow Algorithmic Trading Amazon S3 Data Analysis Apache HTTP Server Continuous Integration Information Engineering Data Systems Linux DevOps Distributed Computing Environment Network Packet
+22 more
Python (Programming Language) Pcap Multicasting Query Optimization Reference Data Backtesting Prometheus DataOps Software Engineering SQL Databases Sql Optimization Grafana Apache Spark Backend Data Layers Containerization Data Lakes Pyspark Kubernetes Decoding Terraform Docker

Job description

Market Data Engineer (Domain Trading Expertise Required)Full timeBHFT is a proprietary algorithmic trading firm.Our team manages the full trading cycle, from software development to creating and coding strategies and algorithms.Our trading operations cover key exchanges.The firm trades across a broad range of asset classes, including equities, equity derivatives, options, commodity futures, rates futures, etc.We employ a diverse and growing array of algorithmic trading strategies, utilizing both High- and Medium-Frequency Trading approaches.We’re a team of 200+ professionals, with a strong emphasis on technology-70% are technical specialists in development, infrastructure, testing, and analytics spheres.The remaining part of the team supports our business operations, such as Risks, Compliance, Legal, Operations and more.With a strong focus on innovation and performance, BHFT is actively expanding its presence in traditional financial markets.We value a results-driven culture, emphasizing collaboration, transparency, and constant improvement, all while offering the flexibility of remote work and a globally distributed team.The Data Engineering team is responsible for designing, building, and maintaining the Market Data Platform - a lakehouse infrastructure spanning the full path from raw exchange feeds to reliable, petabyte-scale data for research, backtesting, and real-time trading.Key ResponsibilitiesCapture & Ingestion.Own the full capture path from wire to lake: decode and normalize raw exchange feeds (pcap, multicast UDP / ITCH / FIX) and vendor sources (OneTick, Refinitiv, Bloomberg, ICE) into a unified canonical model with nanosecond timestamps.Build batch + stream pipelines (Airflow, Spark, dbt) for tick and reference data.Own L2/L3 order-book reconstruction with gap handling.Provide Python and Rust producer SDKs for internal feed handlers.Storage & Modeling - Apache Iceberg.Own the Iceberg-over-S3 lakehouse: design partitioning, sort orders, and row-group layout for fast scans; manage schema evolution, snapshots, time travel, compaction, and TTL.Maintain reference data as slowly?changing tables with point?in?time correctness for backtests.Drive storage cost optimisation via compaction, tiering, and snapshot expiry.Tooling & Libraries.Build libraries for schema management, data contracts, validation, and lineage on top of the Iceberg catalog.Develop shared access services (Spark + Polars) so Research, backtesting, and trading share one normalized data layer, including gap detection and pcap?vs?lake reconciliation.Reliability & Observability.Embed monitoring, alerting, SLAs/SLOs, and CI/CD across capture and pipeline layers on Kubernetes (EKS).Own data?quality dashboards and incident runbooks for the capture fleet.Collaboration.Partner with Quant Research, Data Science, Backend, and DevOps to translate requirements into platform capabilities and champion market?data engineering best practices.Qualifications5+ years building production?grade data systems, with proven expertise architecting and launching data lakes / lakehouses from scratch.Hands?on experience with Apache Iceberg (or comparable table formats - Delta / Hudi): partitioning, schema evolution, snapshots, compaction, and catalog operations; familiarity with Apache Arrow for zero?copy, columnar in?memory interchange.Experience with market data and/or network packet capture - decoding pcap, exchange feed protocols (ITCH, FIX/FAST, multicast UDP), order?book reconstruction, and time?series at scale (strong plus; willingness to learn required).Experience normalizing market data from multiple vendors - e.g. OneTick, Refinitiv/Reuters, Bloomberg, ICE - into a unified schema and symbology (strong plus).Expert-level Python (incl. Polars and/or PySpark); Rust a strong plus (relevant for high?performance capture/decoding).Modern orchestration (Airflow) and distributed processing (Apache Spark).Advanced SQL: complex aggregations, window functions, query optimization, partition pruning.Solid fundamentals in Linux, containerization (Docker, Kubernetes / EKS), and cloud object storage (AWS S3).DevOps & observability: CI/CD, infrastructure?as?code (Terraform), GitOps (ArgoCD), and metrics/dashboards/alerting (Grafana, Prometheus).Strong grasp of structured + unstructured / binary data, and storage optimization - partitioning, compression, cost management.English fluency for documentation and collaboration in an international team.We OfferWork in a modern IT company - no bureaucracy or legacy systems.Real opportunities for professional growth and to make your mark.Fully remote work from anywhere in the world, on a flexible schedule.Compensation for health insurance, sports, professional development, and more.#J-*****-Ljbffr

Requirements

5+ years building production?grade data systems, with proven expertise architecting and launching data lakes / lakehouses from scratch. Hands?on experience with Apache Iceberg (or comparable table formats - Delta / Hudi): partitioning, schema evolution, snapshots, compaction, and catalog operations; familiarity with Apache Arrow for zero?copy, columnar in?memory interchange. Experience with market data and/or network packet capture - decoding pcap, exchange feed protocols (ITCH, FIX/FAST, multicast UDP), order?book reconstruction, and time?series at scale (strong plus; willingness to learn required). Experience normalizing market data from multiple vendors - e.g. OneTick, Refinitiv/Reuters, Bloomberg, ICE - into a unified schema and symbology (strong plus). Expert-level Python (incl. Polars and/or PySpark); Rust a strong plus (relevant for high?performance capture/decoding). Modern orchestration (Airflow) and distributed processing (Apache Spark). Advanced SQL: complex aggregations, window functions, query optimization, partition pruning. Solid fundamentals in Linux, containerization (Docker, Kubernetes / EKS), and cloud object storage (AWS S3). DevOps & observability: CI/CD, infrastructure?as?code (Terraform), GitOps (ArgoCD), and metrics/dashboards/alerting (Grafana, Prometheus). Strong grasp of structured + unstructured / binary data, and storage optimization - partitioning, compression, cost management. English fluency for documentation and collaboration in an international team.

Benefits & conditions

Work in a modern IT company - no bureaucracy or legacy systems. Real opportunities for professional growth and to make your mark. Fully remote work from anywhere in the world, on a flexible schedule. Compensation for health insurance, sports, professional development, and more. #J-*****-Ljbffr

About the company

BHFT is a proprietary algorithmic trading firm. Our team manages the full trading cycle, from software development to creating and coding strategies and algorithms. Our trading operations cover key exchanges. The firm trades across a broad range of asset classes, including equities, equity derivatives, options, commodity futures, rates futures, etc. We employ a diverse and growing array of algorithmic trading strategies, utilizing both High- and Medium-Frequency Trading approaches. We’re a team of 200+ professionals, with a strong emphasis on technology-70% are technical specialists in development, infrastructure, testing, and analytics spheres. The remaining part of the team supports our business operations, such as Risks, Compliance, Legal, Operations and more. With a strong focus on innovation and performance, BHFT is actively expanding its presence in traditional financial markets. We value a results-driven culture, emphasizing collaboration, transparency, and constant improvement, all while offering the flexibility of remote work and a globally distributed team.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.buscojobs.com.es

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:07 min

Inspecting default bridge architectures and custom Docker networks

Oliver Seitz Oliver Seitz · WWC 2025

52 sec

Running persistent Linux environments directly on Windows

Ben Breard Ben Breard · WWC 2025

3:09 min

Balancing data science skillings alongside systems engineering rigor

Nico Schmidt · LIVE

2:34 min

Docker sandbox architecture and microVM environment integration

Manuel de la Peña Manuel de la Peña · WWC Europe 2026

3:55 min

Demonstrating .NET installation on Debian and Azure Linux

Silvano Coriani Silvano Coriani · Europe 2026 Virtual

2:57 min

Core technical practices for robust data engineering

Sandhya Menon Sandhya Menon · WWC Europe 2026

Videos

See all

Related articles

See all