Data Engineer

Highbrow LLC
United States
1 day ago
Apply on www.thejobnetwork.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
10 years minimum
Working hours
Regular working hours

Tech stack

Java (Programming Language) Artificial Intelligence Airflow Apache HTTP Server Microsoft Azure Continuous Delivery Continuous Integration Customer Data Management Data Governance Data Security Dimensional Modeling Apache Hive
+20 more
Identity and Access Management Python (Programming Language) Operational Databases Performance Tuning Power BI SQL Databases Data Streaming Snowflake Change Data Capture Data Lakes Debezium Kubernetes Information Technology Apache Flink Apache Kafka Spark Streaming Data Management Presto Looker Analytics Data Pipelines

Job description

Build and operate production-grade batch and streaming data pipelines across SQL Server, cloud applications, device/event streams, Customer Relationship Management systems, and third-party sources.

Requirements

10+ years of hands-on experience in data engineering, data warehousing, or data platform development.

\n

Strong production experience building and supporting end-to-end data pipelines.

\n

Expert SQL, including complex queries, query-plan analysis, and performance tuning.

\n

Strong development skills in Python, Java, or Scala.

\n

Hands-on experience with Apache Iceberg, Delta Lake, or Apache Hudi.

\n

Experience with distributed query technologies such as Trino, Presto, or Spark SQL.

\n

Hands-on Snowflake experience, including data modeling, performance optimization, external/Iceberg tables, secure data sharing, or access management.

\n

Experience with streaming and Change Data Capture technologies such as Apache Flink, Kafka, Spark Structured Streaming, or Debezium.

\n

Strong understanding of dimensional modeling, data quality, schema evolution, and data contracts.

\n

Experience with Kubernetes, containers, Continuous Integration and Continuous Deployment, and infrastructure-as-code.

\n

Experience implementing monitoring, alerting, lineage, reconciliation, and operational support for production data platforms.

\n

Ability to independently troubleshoot complex production issues and drive them through resolution., Experience with multi-tenant Software as a Service data platforms and customer data isolation.

\n

Experience integrating open lakehouse platforms with Snowflake.

\n

Azure cloud experience.

\n

Experience with Airflow, Dagster, dbt, Power BI, Looker, or Superset.

\n

Experience with data governance, cataloging, masking, and access-control frameworks.

\n

Exposure to Internet of Things/device telemetry, physical security, commercial real estate, or property technology.

\n

Experience using Artificial Intelligence-assisted engineering tools to accelerate data development., Bachelor’s or Master’s Degree in Computer Science, Computer or Electrical Engineering, Mathematics, or a related field.

Benefits & conditions

Develop ingestion, transformation, and enrichment pipelines using Apache Flink, Kafka, Change Data Capture, SQL, and Python/Java/Scala.

\n

Implement lakehouse data layers across landing, raw, conformed, and consumption zones using Apache Iceberg and Trino.

\n

Build dimensional and analytical models for entities such as buildings, tenants, credentials, devices, and people.

\n

Implement robust handling for Change Data Capture, late-arriving data, retries, backfills, replay, schema evolution, and data reconciliation.

\n

Build and maintain integration between the lakehouse and Snowflake, including Iceberg-backed/external tables, secure sharing, and downstream data contracts.

\n

Develop automated data-quality checks, pipeline monitoring, lineage, freshness alerts, and operational dashboards.

\n

Troubleshoot and resolve production data issues, pipeline failures, performance bottlenecks, and data-quality problems.

\n

Optimize SQL queries, storage layouts, partitioning, clustering, compute utilization, and overall platform cost.

\n

Implement security controls including role-based access, row/column-level security, masking, and protection of sensitive data.

\n

Build and maintain Continuous Integration and Continuous Deployment pipelines, automated testing, infrastructure-as-code, and environment promotion processes.

\n

Participate in production support, incident resolution, Root Cause Analysis, deployment, and on-call activities.

\n

Work closely with Product, Application Engineering, Quality Assurance, DevOps/Site Reliability Engineering, Analytics, and globally distributed engineering teams.

\n

Contribute to technical design decisions and documentation while remaining primarily hands-on in implementation and production delivery.

\n

\n

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.thejobnetwork.com
Prepare application

Good distractions

Loading talks and stories from around this role…