Data Engineer w/ AI
NTH VENTURE, INC
United States
13 days ago
- Discuss this with your agent
- Open in Claude
- Open in ChatGPT
Role details
Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Compensation
$197,600.0
Working hours
Regular working hours
Job source
Tech stack
Artificial Intelligence
Airflow
Confluence
Microsoft Azure
Continuous Integration
Data Validation
Information Engineering
Data Infrastructure
Extract Transform Load (ETL)
Data Systems
Distributed Data Store
Document-Oriented Databases
+20 more
Github
Python (Programming Language)
SQL Azure
Cloud Services
Systems Architecture
Azure Service Bus
Data Processing
Cloud Platform System
Azure Data Factory
Git
Kubernetes
Data Analytics
Data Management
Interactive Whiteboards
Terraform
Azure Synapse Analytics
Software Version Control
Data Pipelines
Docker
Databricks
Job description
The ideal candidate has strong Python skills, experience orchestrating workflows with Apache Airflow, and hands-on expertise working with modern data libraries such as Polars. You will work extensively with Docker and deploy data solutions in the Microsoft Azure cloud environment. This role will also focus on modernizing the data platform by migrating existing data pipelines from Azure Synapse Analytics to Apache Airflow-based orchestration frameworks., * Use Claude to assist in your delivery
- Design, develop, and maintain reliable, scalable data pipelines using Python, Polars, and Apache Airflow
- Maintain Airflow deployments in Docker to ensure portability and consistency across environments
- Deploy, manage, and monitor data solutions in Azure (e.g. Azure Container Instances, Azure Storage, Azure SQL)
- Optimize data performance, reliability, and cost in cloud-based architectures
- Collaborate with analytics, BI, and data science teams to ensure high-quality, well-modeled data
- Implement data quality checks, monitoring, and alerting
- Follow best practices for version control, testing, CI/CD, and infrastructure-as-code (e.g., GitHub, Azure DevOps, Azure Artefacts, Azure Container Registry, Terraform)
- Document data pipelines, system architecture, and operational procedures (Confluence and Miro)
Requirements
- Claude usage
- Strong proficiency in Python for data engineering use cases
- Hands-on experience with Apache Airflow for workflow orchestration
- Practical knowledge of Polars for data processing and transformation
- Solid experience building and running Docker containers
- Experience working in Microsoft Azure cloud environments
- Understanding of data modeling and schemas, ETL/ELT patterns, and distributed data systems
- Experience with Git and collaborative development workflows
- Strong problem-solving skills and attention to detail, * Experience with Azure data services such as Azure Data Factory, Synapse Analytics, Databricks, or Event Hubs
- Experience deploying containerized workloads using AKS or similar orchestration platforms
- Familiarity with CI/CD pipelines for data platforms
Benefits & conditions
- Opportunity to work on modern, cloud-native data platforms
- Collaborative, data-driven engineering culture
- Competitive compensation and benefits
- Professional growth and learning opportunities
Pay: From $95.00 per hour
Apply for this position
This job is hosted externally. Click below to view the full posting and apply.
Apply on www.indeed.comGood distractions
Talks and stories from around this role — technically off-topic, practically not.
Moments
Explore playlistsVideos
See allRelated articles
See all
CH
Chris Heilmann
almost 2 years ago
BR
Benjamin Ruschin
Navigating the AI Shift
11 months ago
BB
Benedikt Bischof
MLOps And AI Driven Development
over 4 years ago
CH
Chris Heilmann
Dev Digest 120 - Apple and peers
about 2 years ago
LM
Luis Minvielle
How to Become an AI Engineer
over 2 years ago
EM
Eli McGarvie
Highest Paying Tech Companies for Developers
over 3 years ago