> Markdown version of [/jobs/ext/3038442-lead-data-engineer-ai-ml](https://www.wearedevelopers.com/jobs/ext/3038442-lead-data-engineer-ai-ml). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Lead Data Engineer - AI/ML - **Company:** Stanford Health Care - **Location:** Palo Alto, CA, United States - **Experience:** Expert - **Salary:** $202,134.0 - $267,862.0 - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Big Data, Hospital Information Systems, Software as a Service, Cloud Computing, Software Quality, Continuous Integration, Data Infrastructure, Software Debugging, Python (Programming Language), Machine Learning, Software Requirements Analysis, SQL Databases, Data Streaming, Data Processing, Usage Tracking, Information Technology, Deployment Automation, Machine Learning Operations, Data Pipelines, Programming Languages - **Published:** September 23, 2026 - **Apply:** https://dejobs.org/x/x/C31C3A5945994604B12C75FEEEFB4A27/job/ ## About the Role * Bachelor's or Master's degree in Computer Science, Engineering, or related, or equivalent working experience, * 5+ years experience in building data infrastructure for analytics teams, including ability to write code in SQL, R, or Python for processing large datasets in distributed cloud environments Required Knowledge, Skills and Abilities * Experience with cloud deployment strategies and CI/CD. * Experience building and working with data infrastructure in a SaaS environment. * Experience overseeing, developing or implementing machine learning operations (MLOps) processes. * Experience mentoring junior engineers and enforcing best practices around code quality. * Knowledge of multiple programming languages, commitment to choosing languages based on project-specific requirements, and willingness to learn new programming languages as necessary. * Knowledge of resource management and automation approaches such as workflow runners. * Collaborative mentality and excitement for iterative design working closely with the Data Science team. ## Description The Lead Data Engineer will be part of a team building Stanford Health Care's (SHC) solutions incorporating Artificial Intelligence including providing health care solutions in the areas of patient care, medical research and administrative services. This group is designed to bring Artificial Intelligence (AI) and other emerging machine learning (ML) based innovations in data science into healthcare and will partner closely with individuals across clinical specialties and operations areas to deploy algorithms that can lead to better patient outcomes. Reporting to the Data Science Director and working closely with Stanford Medicine's inaugural Chief Data Scientist, this role will be responsible for building, scaling and maintaining the compute frameworks, analysis tooling, model implementations and agentic solutions that form our core AI platform., The Lead Data Engineer will be part of a team building Stanford Health Care's (SHC) solutions incorporating Artificial Intelligence including providing health care solutions in the areas of patient care, medical research and administrative services. This group is designed to bring Artificial Intelligence (AI), predictive algorithms and other emerging machine learning (ML) based innovations in data science into healthcare and will partner closely with individuals across clinical specialties and operations areas to deploy algorithms that can lead to better patient outcomes. Reporting to the Data Science Director and working closely with Stanford Medicine's inaugural Chief Data Scientist, this role will be responsible for maintaining compute frameworks, analysis tooling, and/or model implementations used or created by the team. The individual will design, implement, and support data processing software and infrastructure. Locations Stanford Health Care What you will do * Build end-to-end data pipelines and infrastructure for ML models used by the Data Science team and others at SHC. * Understand the requirements of data processing and analysis pipelines and make appropriate technical design and interface decisions. Elucidating these requirements will require training, developing, and validating researcher-built or vendor provided machine learning algorithms on hospital data as well as working with other members of the data science team. * Understand data flows among the SHC applications and use this knowledge to make recommendations and design decisions for languages, tools, and platforms used in software and data projects. * Troubleshoot and debug environment and infrastructure problems found in production and non-production environments for projects by the Data Science Team. * Work with other groups at SHC and the Technology and Digital Solutions (TDS) group to ensure servers and system maintenance based on updates, system requirements, data usage, and security requirements., * Category II - Tasks that involve NO exposure to blood, body fluids or tissues, but employment may require performing unplanned Category I tasks ## Related Videos - [Alibaba Big Data and Machine Learning Technology](https://www.wearedevelopers.com/videos/37-alibaba-big-data-and-machine-learning-technology) - [Why and when should we consider Stream Processing frameworks in our solutions](https://www.wearedevelopers.com/videos/1085-why-and-when-should-we-consider-stream-processing-frameworks-in-our-solutions) - [Fault Tolerance and Consistency at Scale: Harnessing the Power of Distributed SQL Databases](https://www.wearedevelopers.com/videos/1146-fault-tolerance-and-consistency-at-scale-harnessing-the-power-of-distributed-sql-databases) - [Data Science in Retail](https://www.wearedevelopers.com/videos/586-data-science-in-retail) - [PySpark - Combining Machine Learning & Big Data](https://www.wearedevelopers.com/videos/44-pyspark-combining-machine-learning-big-data) - [Python-Based Data Streaming Pipelines Within Minutes](https://www.wearedevelopers.com/videos/1233-python-based-data-streaming-pipelines-within-minutes) ## Related Articles - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [MLOps – What’s the deal behind it?](https://www.wearedevelopers.com/magazine/125-mlops-what-s-the-deal-behind-it) - [What Industries Outside of AI Are Hiring The Most AI Experts?](https://www.wearedevelopers.com/magazine/98-what-industries-outside-of-ai-are-hiring-the-most-ai-experts) - [Top Big Data Technologies That You Need to Know](https://www.wearedevelopers.com/magazine/108-top-big-data-technologies-that-you-need-to-know) - [Got AI ideas but no money? Here are 10 free ways to level up your AI skills with Google Cloud](https://www.wearedevelopers.com/magazine/600-got-ai-ideas-but-no-money-here-are-10-free-ways-to-level-up-your-ai-skills-with-google-cloud)