> Markdown version of [/jobs/ext/2017334-data-platform-engineer](https://www.wearedevelopers.com/jobs/ext/2017334-data-platform-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Data Platform Engineer - **Company:** Treeswift Inc - **Location:** New York, NY, United States - **Experience:** Experienced - **Salary:** $160,000.0 - $200,000.0 - **Contract:** Permanent contract - **Skills:** Geographic Information Systems, Artificial Intelligence, Airflow, Amazon Web Services, Amazon S3, Apache HTTP Server, Cloud Computing, Cloud Storage, Computer Engineering, Directed Acyclic Graph (Directed Graphs), Information Engineering, Data Files, Data Infrastructure, Software Debugging, Programming Tools, Python (Programming Language), Machine Learning, MongoDB, Operational Databases, Oracle (Applications), Software Engineering, Scripting, Enterprise Software Applications, Data Storage Technologies, Kubernetes, Information Technology, Luigi, Data Management, Lidar, Data Pipelines - **Published:** August 10, 2026 - **Apply:** https://www.careerbuilder.com/job-details/data-platform-engineer-new-york-ny--8954268d-1e28-4167-aac4-bf26eb487d5e ## About the Role * Bachelor's degree in Computer Science, Computer Engineering, Math, or a related field (or equivalent experience). * 4+ years of data engineering or backend engineering experience with a focus on pipelines, orchestration, or platform. * Hands-on experience building and maintaining production data pipelines (e.g. Airflow, Prefect, Luigi, or similar). We use Python for our pipeline environment, machine learning, and developer tooling; we don't require Python expertise and are happy for you to learn on the job. * Experience with cloud object storage and data-at-scale (we use AWS and S3; cloud experience is required, but prior AWS experience is not). * Comfort with Kubernetes and container-based deployments in practice: running workloads on K8s, resource and volume configuration, and debugging pod/worker issues. * Ability to own work end-to-end: design, implement, test, and operate pipelines and related tooling. You are comfortable picking up new parts of the stack when needed. * Strong collaboration and communication; you work well with ML, hardware, and product stakeholders and can explain tradeoffs clearly. Nice-to-haves * Experience in early-stage or fast-moving environments where scope and ownership evolve. * Experience with Apache Airflow (especially 3.x) and/or Astronomer. * Experience with geospatial data, imagery, lidar, or point clouds. * Interest in utilities, forestry, or field operations and how data pipelines support those domains., Amazon Simple Storage Service (S3), Amazon Web Services (AWS), Apache, Artificial Intelligence (AI), Cloud Computing, Cloud Storage, Communication Skills, Computer Engineering, Computer Science, Construction, Construction Planning, Cross-Functional, Data Management, Data Sets, Data Storage, Debugging Skills, Enterprise Applications, Forestry, Fundraising, Light Detection and Ranging (LiDAR)\Laser Detection and Ranging (LADAR), Machine Learning, Machine Tool, Mathematics, MongoDB, Oracle, Programming Tools, Project Planning, Python Programming/Scripting Language, Regulations, Risk, Risk Management, Robotics, Software Design, Software Development, Software Engineering, Team Player, Testing ## Description You are a skilled and motivated Data Platform Engineer. You will: * Design, build, and maintain data pipelines at scale. We run Apache Airflow 3 on Astronomer with pipelines that process terabytes of real-world physical data across many file types-imagery, audio, point clouds, and more. You will develop and evolve DAGs that orchestrate complex, multi-step workflows: dozens of tasks, fan out/in in pipelines, Python and Kubernetes operators split across generalized and specialized node pools, and dynamic DAG generation. You will work closely with our in-house ML team (feature pipelines and model deployment live in these DAGs) and coordinate with our hardware team on ingestion and formats. Scope is a mix of pipeline development and platform ownership and we are happy to adjust the scope and balance of responsibilities based on your interests and strengths. * Help us scale and harden our data platform. We have one dedicated data engineer today; you will be the second. The broader engineering team is highly collaborative and you will work with members of the full-stack and machine learning teams. We are looking for someone to improve DAG design and execution, resource and cost tuning, reliability and observability, and contribute to how we run Airflow and Kubernetes in the cloud. If you enjoy writing pipelines and improving the platform that runs them, this role has room for both. * Stay curious, collaborative, and cross-functional. We are a small team where many people wear multiple hats. You will work alongside ML engineers, hardware engineers, and software engineers. Turning a technically complex set of requirements from a critical industry into a rich data set is at the center of what we do. We take pride in managing complexity and providing high-fidelity data that our customers can use to make better-informed decisions. * Be an owner at Treeswift; make the company better in whatever form that takes. We value the full picture you bring-whether that's deep expertise in orchestration, a knack for debugging at scale, or hidden talents outside work. We launched our platform last fall and have only scratched the surface of what's possible in terms of finding ways to add value for our customer. You will partner closely with some of the largest utilities in the country and contribute to efforts to develop new workflows in work planning, construction and disaster response. This is a full-time, hybrid role based out of our Lower Manhattan, NYC office (2 days per week in person, currently pinned to Tuesdays and Wednesdays). ## Related Videos - [From Messy Queries to Scalable Systems - How Data Engineering actually works](https://www.wearedevelopers.com/videos/100203-from-messy-queries-to-scalable-systems-how-data-engineering-actually-works) - [How to develop an autonomous car end-to-end: Robotic Drive and the mobility revolution](https://www.wearedevelopers.com/videos/22-how-to-develop-an-autonomous-car-end-to-end-robotic-drive-and-the-mobility-revolution) - [Remote Driving on Plant Grounds with State-of-the-Art Cloud Technologies](https://www.wearedevelopers.com/videos/251-remote-driving-on-plant-grounds-with-state-of-the-art-cloud-technologies) - [Enjoying SQL data pipelines with dbt](https://www.wearedevelopers.com/videos/823-enjoying-sql-data-pipelines-with-dbt) - [AI Model Management Life Circles: ML Ops For Generative AI Models From Research to Deployment](https://www.wearedevelopers.com/videos/1152-ai-model-management-life-circles-ml-ops-for-generative-ai-models-from-research-to-deployment) - [Modern Data Architectures need Software Engineering](https://www.wearedevelopers.com/videos/1030-modern-data-architectures-need-software-engineering) ## Related Articles - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [Top Big Data Technologies That You Need to Know](https://www.wearedevelopers.com/magazine/108-top-big-data-technologies-that-you-need-to-know) - [Making Data Warehouses Fast: A Developer’s Story](https://www.wearedevelopers.com/magazine/107-making-data-warehouses-fast-a-developer-s-story) - [A Guide to Green Tech and Green IT Careers](https://www.wearedevelopers.com/magazine/374-a-guide-to-green-tech-and-green-it-careers) - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk)