> Markdown version of [/jobs/ext/1178814-data-architect](https://www.wearedevelopers.com/jobs/ext/1178814-data-architect). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Data Architect - **Company:** Climb Group, Inc. - **Location:** United States (Remote available) - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Airflow, Amazon Web Services, Microsoft Azure, Code Review, Continuous Integration, Data Architecture, Information Engineering, Data Infrastructure, Machine Learning, Operational Databases, Performance Tuning, Standard Sql, Software Deployment, SQL Databases, Enterprise Data Management, Apache Spark, Build Management, Data Lakes, Pyspark, Data Management, Data Pipelines, Databricks - **Published:** July 4, 2026 - **Apply:** https://www.indeed.com/viewjob?jk=e41e55af397dabdd ## About the Role * 6+ years in data engineering or analytics engineering. * Advanced experience building production data pipelines with Apache Spark (PySpark and SQL), including performance optimization. * Hands-on production experience with Databricks (Delta Lake, Jobs, Workflows, Unity Catalog). * Deep experience with at least one major cloud (AWS, Azure, or GCP). * Strong SQL and data modeling fundamentals. * A track record of building and operating production-grade data pipelines, not just prototypes. * Comfortable working directly with client stakeholders to translate business requirements into technical solutions. * Demonstrated ability to independently own a data engineering workstream from design through production., * Experience with Delta Live Tables, structured streaming, or real-time pipelines. * Familiarity with dbt, Airflow, or other modern data stack tooling. * A demonstrated performance- and cost-optimization mindset (cluster sizing, Photon, file layout, partitioning). * Exposure to governed or regulated environments (e.g., financial services, healthcare). * Prior consulting, systems integrator, or professional services experience. Note on certification: Existing Databricks certifications are a plus. Where not already held, Databricks certification (e.g., Data Engineer Associate/Professional) is expected to be obtained post-hire. ## Description Climb is a Data and AI consultancy that partners with enterprises to design, build, and operationalize modern data platforms and production AI systems. As a Databricks partner, we go deep on lakehouse architecture, machine learning, and applied AI, with a bias toward production over proof of concept. Our team brings deep technical expertise and a builder's mindset to every engagement, and we measure our work not just by what ships, but by the business impact it drives., Senior Data Engineers own complete data engineering workstreams within client engagements. You design, build, and operationalize the data platforms that power analytics and AI, taking your work from discovery and design through implementation, production, and handoff. This is a hands-on engineering role built around ownership. You work directly with client stakeholders to translate business requirements into scalable data solutions, making the technical decisions necessary to deliver reliable, secure, and cost-efficient platforms. While the Data Architect owns the overall platform architecture, you own the successful delivery of your workstreams and the quality of the systems you build. Our work focuses on modernizing enterprise data platforms using Databricks and the broader cloud ecosystem, enabling organizations to build trustworthy data products and AI-ready foundations that last beyond the engagement., * Own one or more data engineering workstreams, from technical design through production deployment and handoff. * Design and implement scalable data models and lakehouse architectures, including medallion patterns where appropriate. * Optimize performance and cost across Databricks and the underlying cloud: you treat compute spend as your problem, not someone else's. * Orchestrate workflows using Databricks Workflows, Delta Live Tables, or equivalent tooling. * Implement governance, security, observability, and lineage with Unity Catalog, and stand up CI/CD for data. * Design and build AI-ready data platforms that enable reliable analytics, machine learning, and agentic applications. * Work directly with client stakeholders to gather requirements and translate them into technical solutions. * Review code, mentor junior engineers, and maintain high engineering standards across your workstreams. * Contribute reusable accelerators, frameworks, and best practices back to the practice., * Outcomes, not hours. We sell and deliver against business results. Advancement is tied to delivery performance and account impact, not utilization targets. * Senior team, no body-shop drag. Small pods of A-players, heavy internal AI leverage, and no bloated middle layers between you and the work. * IP that compounds. Every engagement feeds reusable accelerators, patterns, and points of view back into the practice. ## Related Videos - [Enjoying SQL data pipelines with dbt](https://www.wearedevelopers.com/videos/823-enjoying-sql-data-pipelines-with-dbt) - [From Messy Queries to Scalable Systems - How Data Engineering actually works](https://www.wearedevelopers.com/videos/100203-from-messy-queries-to-scalable-systems-how-data-engineering-actually-works) - [PySpark - Combining Machine Learning & Big Data](https://www.wearedevelopers.com/videos/44-pyspark-combining-machine-learning-big-data) - [Data Science in Retail](https://www.wearedevelopers.com/videos/586-data-science-in-retail) - [Cutting LLM Costs Without Cutting Quality: How to Beat Proprietary LLMs with Fine-Tuned Open Source](https://www.wearedevelopers.com/videos/100151-cutting-llm-costs-without-cutting-quality-how-to-beat-proprietary-llms-with-fine-tuned-open-source) - [OLTP in the Lakehouse: Redefining Data for AI Workloads](https://www.wearedevelopers.com/videos/2038-oltp-in-the-lakehouse-redefining-data-for-ai-workloads) ## Related Articles - [Making Data Warehouses Fast: A Developer’s Story](https://www.wearedevelopers.com/magazine/107-making-data-warehouses-fast-a-developer-s-story) - [Graph and AI Trends 2026: Why Is AI Running but Not Yet Delivering?](https://www.wearedevelopers.com/magazine/680-graph-and-ai-trends-2026-why-is-ai-running-but-not-yet-delivering) - [What Are The Top Skills Required For Azure Developers?](https://www.wearedevelopers.com/magazine/77-what-are-the-top-skills-required-for-azure-developers) - [Best US AI Conferences for CTOs in 2026: Build vs. Buy, Vendor Evaluation, and Peer Intelligence](https://www.wearedevelopers.com/magazine/736-best-us-ai-conferences-for-ctos-in-2026-build-vs-buy-vendor-evaluation-and-peer-intelligence) - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer)