> Markdown version of [/jobs/ext/2864326-data-engineer](https://www.wearedevelopers.com/jobs/ext/2864326-data-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Data Engineer - **Company:** Princeton University - **Location:** Princeton, NJ, United States - **Experience:** Expert - **Salary:** $120,000.0 - $135,000.0 - **Contract:** Permanent contract - **Skills:** Active Directory, Airflow, Amazon Web Services, Microsoft Azure, Bash Shell, Business Intelligence Development, Cloud Computing, Databases, Continuous Integration, Data Validation, Data Dictionary, Data Governance, Data Integration, Extract Transform Load (ETL), Data Transformation, Data Mining, Data Warehousing, IBM InfoSphere DataStage, Linux, DevOps, Document-Oriented Databases, Domain Name System (DNS), Identity and Access Management, Python (Programming Language), Lightweight Directory Access Protocols (LDAP), PostgreSQL, Microsoft SQL Server, Networking Basics, OAuth, Oracle (Applications), Cloud Services, Shell Script, SQL Databases, TCP/IP, Workflow Management Systems, Azure Service Bus, SSL Certificate Management, Scripting, Cloud Platform System, Azure Data Factory, Sql Optimization, Delivery Pipeline, Snowflake, Data Build Tool (dbt), Microsoft Fabric, Information Technology, Data Lineage, Apache Kafka, Restful APIs, Data Pipelines, Databricks - **Published:** September 12, 2026 - **Apply:** https://www.jofdav.com/jobs/59678470-data-engineer ## About the Role * CORE SKILLS + Linux / Shell scripting + Python programming + SQL (query authoring and data validation) + Basic networking (DNS, HTTP/S, TCP/IP, proxies, firewalls) + Cloud infrastructure fundamentals (any major provider) + IAM: LDAP, Active Directory, OAuth 2.0, certificate management * 5+ years of proven experience with enterprise ETL/ELT tooling * Advanced SQL skills across multiple platforms (Oracle, SQL Server, PostgreSQL) * Strong Python programming skills for data transformation, scripting, and pipeline automation * Shell scripting proficiency for batch job automation on Linux/Unix servers * Hands-on experience with at least one cloud data platform: Microsoft Fabric, Snowflake (with dbt and/or Fivetran), or Databricks * Experience with dbt (data build tool) for transformation layer development * Linux/Unix systems fluency including file management, cron scheduling, and process monitoring * Understanding of data warehousing concepts: dimensional modelling, star/snowflake schemas, slowly changing dimensions * Familiarity with basic networking, storage, and cloud infrastructure concepts * Experience with IAM and access control: LDAP, Active Directory, and database-level permission management * Working knowledge of REST APIs for source system data extraction and pipeline orchestration * Ability to document data flows, pipeline architecture, and transformation logic clearly * Education: Bachelor's degree in computer science, * 7+ years proven experience with enterprise ETL/ELT tooling * Familiarity with orchestration tools such as Apache Airflow, Azure Data Factory, or Prefect * Exposure to streaming or near-real-time ingestion patterns (Kafka, Kinesis, Event Hubs) * Experience with data quality frameworks (Great Expectations, Soda, or equivalent) * Cloud platform certifications (Azure, AWS, or GCP data engineering tracks) ## Description The Princeton DMIA Integration team is looking for a Data Engineer to own and evolve our data integration practice. You will be responsible for ingesting data from enterprise source systems into our data warehouse platform - what the CIO office refers to as system-to-data-repository integrations. You will play a key role in our migration to a cloud-native data platform, with candidates expected to bring expertise in modern tooling such as Microsoft Fabric, Snowflake with dbt and Fivetran, or Databricks. The existing data warehouse technologies include IBM DataStage, SQL, Oracle, and Shell-based pipelines running on on-premises Linux infrastructure., Architect, Design, Develop: * Design, build, and maintain ETL/ELT pipelines to ingest, transform, and load data from source systems into the enterprise data warehouse. * Lead the migration of on-premises data pipelines to the organization's future cloud-native data platform (one of Fabric, Snowflake + dbt + Fivetran, or Databricks). * Implement and enforce data quality checks, data lineage tracking, and pipeline observability across all integration workflows. * Ensure data security and compliance requirements are met, including encryption at rest and in transit, and access controls aligned with IAM policies. * Optimize pipeline performance, scheduling, and resource utilization across batch and incremental load patterns. Collaborate and Coordinate: * Partner with data analysts, BI developers, and source system owners to understand data requirements and translate them into robust ingestion pipelines. Production Support: * Operate and support ingestion and transformation pipelines. * Develop and maintain data pipeline documentation, data dictionaries, and SLA agreements for ingestion jobs. * Participate in on-call production support rotation and respond to integration incidents per SLA. * Contribute to CI/CD pipeline setup and DevOps practices for data integration deployments. ## Related Videos - [An Applied Introduction to eBPF with Go](https://www.wearedevelopers.com/videos/1075-an-applied-introduction-to-ebpf-with-go) - [Enjoying SQL data pipelines with dbt](https://www.wearedevelopers.com/videos/823-enjoying-sql-data-pipelines-with-dbt) - [Docker network without Docker](https://www.wearedevelopers.com/videos/1418-docker-network-without-docker) - [Keeping applications secure by evolving OAuth 2.0 and OpenID Connect](https://www.wearedevelopers.com/videos/100152-keeping-applications-secure-by-evolving-oauth-2-0-and-openid-connect) - [Turning Container security up to 11 with Capabilities](https://www.wearedevelopers.com/videos/718-turning-container-security-up-to-11-with-capabilities) - [From Messy Queries to Scalable Systems - How Data Engineering actually works](https://www.wearedevelopers.com/videos/100203-from-messy-queries-to-scalable-systems-how-data-engineering-actually-works) ## Related Articles - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [Making Data Warehouses Fast: A Developer’s Story](https://www.wearedevelopers.com/magazine/107-making-data-warehouses-fast-a-developer-s-story) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Top Big Data Technologies That You Need to Know](https://www.wearedevelopers.com/magazine/108-top-big-data-technologies-that-you-need-to-know) - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [What Are The Top Skills Required For Azure Developers?](https://www.wearedevelopers.com/magazine/77-what-are-the-top-skills-required-for-azure-developers)