Data Engineer - II (Biometrics)

Jumio
United States
11 days ago
Apply on job-boards.greenhouse.io
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Experienced
Experience required
3 years minimum
Working hours
Regular working hours

Tech stack

Airflow Amazon Web Services Amazon Elastic Compute Cloud Amazon S3 Computer Vision Big Data Biometrics Software as a Service Data Architecture Information Engineering Data Governance Extract Transform Load (ETL)
+10 more
Distributed Data Store Python (Programming Language) Machine Learning SQL Databases Data Processing Grafana Model Validation Data Management Machine Learning Operations Data Pipelines

Job description

At Jumio, we’re building trusted identity solutions powered by machine learning and biometrics. As a Data Engineer II, you will take ownership of designing, building, and scaling data pipelines that power our ML systems. You’ll work closely with ML engineers, platform teams, and product stakeholders to enable robust dataset creation, benchmarking, and production-grade data workflows on AWS., * Design, build, and maintain scalable and reliable data pipelines for dataset creation, transformation, and benchmarking

  • Own and optimize Airflow pipelines on AWS for data processing, orchestration, and evaluation workflows
  • Write efficient, production-grade SQL and Python code for large-scale data processing and analysis
  • Partner closely with ML engineers to enable model training, evaluation, and benchmarking pipelines
  • Improve pipeline performance, reliability, and observability, ensuring high data quality in production
  • Build and maintain systems to support model performance tracking and data drift monitoring
  • Troubleshoot and resolve data issues across pipelines, ensuring minimal impact on ML workflows
  • Contribute to data architecture decisions and best practices across the platform
  • Collaborate cross-functionally with ML, platform, and data teams to support scalable ML infrastructure

Requirements

  • 3-5 years of experience in Data Engineering, Data Platforms, or related roles
  • Strong proficiency in Python and SQL with experience in production systems
  • Hands-on experience with AWS services (S3, EC2, SageMaker or similar)
  • Solid experience building and managing Airflow (or similar orchestration tools)
  • Strong understanding of data engineering fundamentals (ETL/ELT, data modeling, pipeline design)
  • Experience working with large-scale datasets and distributed data systems
  • Experience supporting ML workflows, datasets, or evaluation pipelines
  • Strong problem-solving skills and ability to work independently in a fast-paced environment, * Experience with ML infrastructure, MLOps, or model evaluation workflows
  • Exposure to biometric systems or computer vision datasets
  • Familiarity with data quality frameworks, monitoring, and observability tools
  • Experience working in SaaS or high-scale production environments

About the company

Jumio is a B2B technology company dedicated to eradicating online identity fraud, money laundering and other financial crimes to help make the internet safer. We leverage AI, biometrics, machine learning, liveness detection and automation to create solutions that are trusted by leading brands worldwide and respected by industry thought leaders.

Jumio is the leading provider of online identity verification, eKYC and AML solutions. With a global footprint, we’re expanding the team to meet strong client demand across a range of industries including Financial Services, Travel, Sharing Economy, Fintech, Gaming, and others.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on job-boards.greenhouse.io
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:15 min

Empowering domain teams with an open data platform

Sandhya Menon Sandhya Menon · World Congress 2026 Europe

10:40 min

Visualizing Prometheus open metrics using custom Grafana dashboards

Stijn Polfliet · LIVE

3:28 min

Defining big data and machine learning fundamentals

Ayon Roy · LIVE

3:09 min

Balancing data science skillings alongside systems engineering rigor

Nico Schmidt · LIVE

2:10 min

Why organizations combine big data and machine learning

Ayon Roy · LIVE

1:04 min

Visualizing Keycloak performance via standard Grafana troubleshooting dashboards

Alexander Schwartz Alexander Schwartz · World Congress 2025

Videos

See all

Related articles

See all