Senior Data Quality Engineer - INTL - India

Insight Global
United States
3 days ago
Apply on www.juju.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
6 years minimum
Working hours
Regular working hours
Job source

Tech stack

Microsoft Azure Big Data Data Architecture Data Validation Information Engineering Data Governance Extract Transform Load (ETL) Database Testing Distributed Systems Github Python (Programming Language) Metadata
+14 more
DataOps SQL Databases Systems Integration Technical Data Management Systems Datadog Enterprise Software Applications Azure Data Factory Apache Spark Data Lakes Pyspark Apache Kafka Video Streaming SDET Databricks

Job description

This role is a highly technical Data Quality Engineer supporting enterprise data validation and quality engineer efforts across the Azure/Databricks platform. As a senior level developer, this role requires expertise in data testing, automation, distributed systems, and modern data architectures. Other key responsibilities include:

  • Design and implement end-to-end data validation strategies for pipelines built on Azure Data Factory, Databricks (Delta Lake), and ADLS

  • Build and perform large-scale data validation using Python and Spark-based validation frameworks

  • Validate data movement, orchestration workflows, and failure handling

  • Define and implement data quality SLAs and KPIs along with tracking data/pipeline health

Drive best practices for data quality and testing standards

Requirements

  • 6-8+ years’ experience in data quality engineer, SDET, data engineering testing, etc.

  • Strong hands-on experience with Azure Data Factory, Databricks (PySpark), and SQL for data quality validation

  • Proven experience building automated data quality frameworks from scratch

  • Strong understanding of data governance, lineage, and metadata

Experience integrating with enterprise systems (ERP, CRM, MDM, etc. * Data observability platforms (Monte Carlo, Soda, Datadog)

  • Experience building CI/CD data testing pipelines (Azure DevOps / GitHub Actions)

  • Streaming Technologies (Kafka, Event Hub)

  • Hands-on experience with dbt (Data Built Tool)

  • Azure Data Engineer (DP-203) Certification or Databricks Certified Professional

Experience within supply chain or large enterprise data environments

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.juju.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

6:36 min

Funding open source through GitHub Accelerator and Sponsors

Stormy Peters · World Congress 2023

4:32 min

Harnessing Spark with Python using PySpark and Py4J

Ayon Roy · LIVE

3:09 min

Balancing data science skillings alongside systems engineering rigor

Nico Schmidt · LIVE

2:40 min

Using GitHub primitives for internal documentation and corporate operations

Kyle Daigle · Coffee With Developers

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

3:37 min

Scaling machine learning pipelines from prototypes to petabytes

Julian Joseph · LIVE

Videos

See all

Related articles

See all