> Markdown version of [/jobs/ext/1388699-software-engineer-lake-warehouse-infrastructure](https://www.wearedevelopers.com/jobs/ext/1388699-software-engineer-lake-warehouse-infrastructure). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Software Engineer, Lake & Warehouse Infrastructure - **Company:** Datavant - **Location:** San Francisco, CA, United States (Remote available) - **Contract:** Permanent contract - **Skills:** Sql Data Warehouse, Artificial Intelligence, Amazon Web Services, Amazon S3, Microsoft Azure, Big Data, BigQuery, Code Review, Continuous Integration, Information Engineering, Data Governance, Data Infrastructure, Data Integration, Data Migration, Github, Identity and Access Management, Subnetting, Python (Programming Language), Routing, Operational Databases, Role-Based Access Control, Standard Sql, Software Engineering, Tableau (Software), Snowflake, Apache Spark, Amazon Virtual Private Cloud (VPC), Data Lakes, Kubernetes, Terraform, Looker Analytics, Amazon Redshift, Databricks - **Published:** July 22, 2026 - **Apply:** https://arc.dev/remote-jobs/j/redirect/p6lblgd8m6 ## About the Role * 2+ years in data engineering, platform engineering, or data-focused software engineering * 2+ years of hands-on AWS with knowledge of networking (VPC, subnets, routing, PrivateLink, security groups) and IAM (roles, policies, permission boundaries) * 1+ years writing production Terraform or equivalent IaC: owning modules, reasoning about state and blast radius, shipping changes safely * 1+ years building self-service tooling, internal platforms, or paved path frameworks consumed by other engineers * Strong SQL and production experience with Snowflake (or an equivalent cloud data warehouse) RBAC, warehouse sizing, cost awareness, and reasoning about how data physically lives in a warehouse or lake * Solid CI/CD discipline in GitHub or equivalent: branching, code review, automated checks, repeatable deployment * An AI-native working style: capable of achieving great leverage through the smart application of AI to every aspect of work * Working knowledge of Python * Strong written and verbal communication skills, * Direct production experience with Iceberg or another open table format, especially bridging Snowflake and Databricks * Building platforms at large scale: Datavant has over 10,000 employees * Running a BI platform close to the warehouse (Sigma, Looker, Tableau, or similar) * Hands-on Databricks or Spark * Kubernetes experience * Snowflake or Databricks certification(s) * Azure or GCP experience -- we're primarily AWS, but our customers and acquisitions aren't always * Integrating data systems with managed identity platforms, particularly via SCIM * Prior experience in healthcare or another highly regulated industry like Finance * Prior DBA, SRE, or DRE work operating production data systems under pressure ## Description * Architect, deploy, and maintain multiple big data platforms like Snowflake, Databricks, Redshift, BigQuery, etc. * Design and implement account and warehouse strategies, RBAC and data governance, and cost discipline * Own and evolve the Terraform that governs these platforms * Keep our 3rd party ecosystem like our data catalog and BI tools tightly integrated with all of our big data platforms * Build and operate the distributed S3 Iceberg lakehouse that bridges Snowflake, Databricks, and others without copying or moving data * Build selfservice tooling so analytics and product engineers can ship models and grants without becoming warehouse experts * Work crossfunctionally with BI&A, data science, and product teams to ensure that the data platform is exceeding their needs and supporting their ambitions * Own troubleshooting and incident response across our warehouses, data governance systems, and our Iceberg data lake * Set the bar on engineering quality, observability, cost discipline, and security in everything we ship ## Related Videos - [Alibaba Big Data and Machine Learning Technology](https://www.wearedevelopers.com/videos/37-alibaba-big-data-and-machine-learning-technology) - [Innovating Developer Tools with AI: Insights from GitHub Next](https://www.wearedevelopers.com/videos/1268-innovating-developer-tools-with-ai-insights-from-github-next) - [Creating a routing app with Google Maps API from scratch](https://www.wearedevelopers.com/videos/831-creating-a-routing-app-with-google-maps-api-from-scratch) - [From Messy Queries to Scalable Systems - How Data Engineering actually works](https://www.wearedevelopers.com/videos/100203-from-messy-queries-to-scalable-systems-how-data-engineering-actually-works) - [Bringing AI Model Testing and Prompt Management to Your Codebase with GitHub Models](https://www.wearedevelopers.com/videos/1536-bringing-ai-model-testing-and-prompt-management-to-your-codebase-with-github-models) - [Modern Data Architectures need Software Engineering](https://www.wearedevelopers.com/videos/1030-modern-data-architectures-need-software-engineering) ## Related Articles - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Making Data Warehouses Fast: A Developer’s Story](https://www.wearedevelopers.com/magazine/107-making-data-warehouses-fast-a-developer-s-story) - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer)