> Markdown version of [/jobs/ext/1419536-sr-databricks-migration-engineer](https://www.wearedevelopers.com/jobs/ext/1419536-sr-databricks-migration-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Sr. Databricks Migration Engineer - **Company:** LinTech Global, Inc. - **Location:** Washington, DC, United States (Remote available) - **Experience:** Expert - **Salary:** $145,000.0 - $165,000.0 - **Contract:** Permanent contract - **Skills:** Query Performance, Amazon Web Services, Microsoft Azure, Big Data, Cloud Engineering, Code Review, Cyber Security, Databases, Information Engineering, Data Governance, Extract Transform Load (ETL), Data Security, Data Systems, Data Warehousing, Relational Databases, Software Design Patterns, Dimensional Modeling, Python (Programming Language), Microsoft SQL Server, Raw Data, Role-Based Access Control, Power BI, Azure Active Directory, Azure Data Lake, SQL Stored Procedures, SQL Databases, Enterprise Data Management, Data Processing, Google Cloud, Data Storage Technologies, Autoscaling, Apache Spark, Data Lakes, Pyspark, Information Technology, Optimization Algorithms, Performance Monitor, Data Management, Data Pipelines, Serverless Computing, Databricks - **Published:** July 24, 2026 - **Apply:** https://www.indeed.com/viewjob?jk=62c73bc887d2281c ## About the Role * Mastery of data engineering principles, including data modeling, ETL (Extract, Transform, Load) processes, and data pipelines * Proficiency with Azure Data Lake data storage and processing services * Skilled at designing, building, and optimizing data pipelines for ingesting, transforming, and loading data * Proficiency in languages such as SQL and Python/PySpark for data manipulation and pipeline development * Skilled at identifying and resolving data-related challenges * Skilled at creating efficient data models that meet business requirements. * Skilled at optimizing query performance and system scalability, * Bachelor's degree or higher from an accredited college or university in Computer Science, Engineering, or a related technical field * 5+ years' experience in data engineering, data system development or related roles * 5+ years' experience with cloud platforms (e.g. Azure, AWS, GCP) * 1+ year leading complex, cross-functional data projects and technical teams * Experience with Databricks Lakehouse, Apache Spark, Delta Lake, cloud-native databases, storage solutions, and distributed compute platforms * Experience with data warehousing, dimensional modeling, enterprise data lakes, incremental data loads, and metadata-driven ingestion and data quality frameworks using PySpark * Must be able to pass a background investigation required to obtain a Public Trust. ## Description * The individual serves as the authoritative resource for the agency who specializes in preparing big data infrastructure for analytical or operational uses. * They are responsible for designing and creating systems that collect, manage, and convert raw data into usable information for data scientists and business analysts to interpret and enables the agency to make smarter decisions and optimize operations. Position Requirements: * Lead the technical migration from legacy SQL Server stored procedures and ADF pipelines to Databricks Lakehouse (Delta Lake), ensuring best practice Lakehouse design. * Translate traditional relational data warehousing paradigms into scalable, distributed Lakehouse frameworks (Bronze, Silver, Gold). * Design robust, reusable ETL/ELT frameworks using PySpark, Delta Live Tables (DLT), and Databricks Workflows. * Architect and refine the Gold Layer (dimensional models, star schemas) specifically to maximize Power BI performance. * Optimize Databricks SQL Warehouses to support high-concurrency, low-latency Power BI queries (DirectQuery and Import modes). * Implement advanced optimization techniques, including Z-Ordering, data skipping, liquid clustering, and materialized views. * Define and enforce governance standards for cluster sizing, auto-scaling policies, and serverless SQL compute to balance performance with cost. * Implement proactive monitoring dashboards to track Databricks Unit (DBU) consumption and identify cost-saving opportunities. * Establish best practices for partition strategies and file size management within Delta Lake. * Design and implement a robust data security model using Unity Catalog for centralized governance. * Enforce row-level and column-level security policies to ensure compliant data access for Power BI consumers and internal analysts. * Align the Lakehouse security architecture with existing enterprise Azure Active Directory (Microsoft Entra ID) and RBAC standards. * Act as the primary technical lead, conducting dedicated pair-programming sessions, workshops, and code reviews to transition the team from SQL-centric to Spark-centric thinking. * Create comprehensive technical documentation, including architecture diagrams, design patterns, and optimization playbooks. * Build a foundational knowledge transfer framework to ensure the internal team is fully self-sufficient post-migration. * Communicate effectively verbally and in written form to both technical and non-technical audience * Work in an organized fashion, completing tasks timely while paying close attention to details ## Related Videos - [PySpark - Combining Machine Learning & Big Data](https://www.wearedevelopers.com/videos/44-pyspark-combining-machine-learning-big-data) - [Beyond Dashboards: Fixing Text-to-SQL with Semantic RAG](https://www.wearedevelopers.com/videos/2036-beyond-dashboards-fixing-text-to-sql-with-semantic-rag) - [Data Governance in the Era of AI](https://www.wearedevelopers.com/videos/1622-data-governance-in-the-era-of-ai) - [From Messy Queries to Scalable Systems - How Data Engineering actually works](https://www.wearedevelopers.com/videos/100203-from-messy-queries-to-scalable-systems-how-data-engineering-actually-works) - [Data Science in Retail](https://www.wearedevelopers.com/videos/586-data-science-in-retail) - [Enjoying SQL data pipelines with dbt](https://www.wearedevelopers.com/videos/823-enjoying-sql-data-pipelines-with-dbt) ## Related Articles - [Making Data Warehouses Fast: A Developer’s Story](https://www.wearedevelopers.com/magazine/107-making-data-warehouses-fast-a-developer-s-story) - [Top Big Data Technologies That You Need to Know](https://www.wearedevelopers.com/magazine/108-top-big-data-technologies-that-you-need-to-know) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Coffee with Developers - Maria Apazoglou - Making AI understandable for all in production](https://www.wearedevelopers.com/magazine/475-coffee-with-developers-maria-apazoglou-making-ai-understandable-for-all-in-production)