ETL Developer

Corporate Brokers, LLC
Austin, TX, United States
21 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Working hours
Regular working hours

Tech stack

Airflow Amazon Web Services Microsoft Azure BigQuery Cloud Computing Software Quality Continuous Integration Data Validation Extract Transform Load (ETL) Data Warehousing Relational Databases Python (Programming Language)
+13 more
SQL Databases SQLAlchemy Scripting Snowflake Pandas Containerization Git Flow Kubernetes Star Schema Data Pipelines Docker Amazon Redshift Databricks

Job description

We are looking for an ETL Developer with strong Python and Airflow expertise to build, manage, and optimize data pipelines. This role focuses on developing reliable workflows that transform and move data across systems efficiently. You’ll be responsible for writing production-ready code, ensuring pipeline performance, and maintaining clean, reusable solutions. What You’ll Do

  • Develop, test, and deploy Python-based ETL pipelines using Apache Airflow.
  • Write efficient, reusable Python scripts for transformations, validations, and data quality checks.
  • Manage scheduling, orchestration, and monitoring of workflows within Airflow.
  • Collaborate with data engineers and analysts to design pipelines aligned to business needs.
  • Troubleshoot and optimize existing ETL jobs for performance, scalability, and reliability.
  • Implement best practices for code quality, testing, and CI/CD integration.
  • Contribute to documentation, pipeline observability, and knowledge sharing.

Requirements

  • Strong experience with Python (pandas, SQLAlchemy, or similar libraries for ETL).
  • Proficiency with Apache Airflow DAG design, task orchestration, and Airflow operators.
  • Experience with dependency and environment management tools such as pipenv, poetry, or conda.
  • Solid understanding of SQL and relational databases.
  • Knowledge of data modeling and transformation patterns (star schema, slowly changing dimensions, etc.).
  • Familiarity with Git-based workflows and CI/CD pipelines.
  • Ability to work independently and collaboratively in a fast-moving environment.

Nice to Have

  • Cloud platform experience (AWS, GCP, or Azure) for data pipelines and storage.
  • Familiarity with containerization (Docker, Kubernetes).
  • Exposure to data warehouses (Snowflake, BigQuery, Redshift).
  • Extensive Databricks experience is ideal.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on public-rest40.bullhornstaffing.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:15 min

Empowering domain teams with an open data platform

Sandhya Menon Sandhya Menon · WWC Europe 2026

2:57 min

Core technical practices for robust data engineering

Sandhya Menon Sandhya Menon · WWC Europe 2026

2:03 min

Accelerating pandas dataframes using cudf module plugins

Ankit Patel Ankit Patel · WWC 2024

2:07 min

Inspecting default bridge architectures and custom Docker networks

Oliver Seitz Oliver Seitz · WWC 2025

3:33 min

Refactoring data science workflows using Rapids QDF and Pandas

Paul Graham Paul Graham · LIVE

3:05 min

Audience questions on AI agents and pipeline vectorization

Joy Joy · WWC 2024

Videos

See all

Related articles

See all