> Markdown version of [/jobs/ext/456163-staff-engineer-data-platform](https://www.wearedevelopers.com/jobs/ext/456163-staff-engineer-data-platform). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Staff Engineer, Data Platform - **Company:** Lila Sciences, Inc. - **Location:** San Francisco, CA, United States - **Experience:** Expert - **Salary:** $192,000.0 - $272,000.0 - **Contract:** Permanent contract - **Skills:** Airflow, Amazon Web Services, Systems Engineering, Architectural Patterns, Cloud Computing, Data Governance, Data Infrastructure, Data Security, Data Systems, Cursor (Graphical User Interface Elements), Fault Tolerance, Python (Programming Language), Machine Learning, NoSQL, Query Optimization, Scientific Computating, SQL Databases, Unstructured Data, Large Language Models, Data Lakes, Core Data, Kubernetes, Storage Technologies, Information Technology, Production Code, Machine Learning Operations - **Published:** June 4, 2026 - **Apply:** https://www.indeed.com/viewjob?jk=9062684de1fa7f87 ## About the Role Do you have experience in Systems engineering?, * Bachelor's or Master's degree in Computer Science, Engineering, or a related field, and 8+ years as a software or data engineer with a focus on building and operating data infrastructure. * Designed and shipped data platform components from the ground up, including ingestion frameworks, storage abstractions, and orchestration systems. Fluent in Python and SQL and writes production-quality code. * Production experience with relational and NoSQL databases, schema design, query optimization, and operational concerns at scale. Comfortable working across structured, semi-structured, and unstructured data. * Proven track record of working cross-functionally with scientists, ML researchers, and engineers. Able to translate domain requirements into platform decisions and explain technical trade-offs to diverse audiences. * Experience with cloud infrastructure and containerized deployment (AWS, Kubernetes). * Hands-on experience with modern table formats and open lakehouse patterns (Iceberg, Delta Lake, Hudi). Bonus Points For * Experience with workflow orchestration systems (Flyte, Airflow, Dagster, or similar). * Experience building data infrastructure that serves agentic and LLM-driven workflows, including vector databases, RAG infrastructure, and retrieval-optimized data access patterns. * Background in scientific computing, life sciences, or research software. * Proficiency with AI-assisted development tools (Cursor, Claude Code, or similar) and ability to incorporate them effectively into day-to-day engineering work. ## Description We are looking for a Staff Engineer to set the technical direction for our core data infrastructure: ingestion frameworks, storage architecture, orchestration patterns, and the interfaces that let scientists and ML researchers work with data reliably at scale. You will work closely with software engineers, machine learning researchers, and lab scientists to understand requirements and translate them into durable platform capabilities. This is a role for engineers who care deeply about how data systems are designed. You will establish the architectural patterns and engineering standards the broader team builds on, mentor engineers across the data platform group, and make technical decisions that compound over time. What You'll Be Building * Data Platform Architecture: Design and evolve the core data infrastructure that ingests, stores, and serves data across scientific and ML workflows. Make principled build-vs-buy decisions and establish architectural patterns adopted by the broader engineering organization. * Ingestion and Integration: Build reliable pipelines that bring in data from diverse sources: laboratory instruments, public scientific datasets, and external research literature. Own the interfaces between upstream producers and downstream consumers. * Orchestration and Reliability: Operate and extend workflow orchestration systems that run complex, multi-step scientific pipelines. Ensure observability, fault tolerance, and reproducibility across the data stack. * Data Modeling and Schema Strategy: Define and maintain data models, schema evolution practices, and data contracts that ensure consistency, discoverability, and long-term durability of scientific and platform data assets. * Cross-Functional Technical Leadership: Partner with ML researchers, lab scientists, and product engineers to translate scientific and research requirements into platform capabilities. Drive alignment on data standards and integration patterns across teams. * Engineering Standards and Mentorship: Establish coding, review, and design standards for the data platform team. Mentor engineers, lead design reviews, and raise the technical bar across the group. ## Related Videos - [From Messy Queries to Scalable Systems - How Data Engineering actually works](https://www.wearedevelopers.com/videos/100203-from-messy-queries-to-scalable-systems-how-data-engineering-actually-works) - [Leveraging Real time data in FSIs](https://www.wearedevelopers.com/videos/806-leveraging-real-time-data-in-fsis) - [Enjoying SQL data pipelines with dbt](https://www.wearedevelopers.com/videos/823-enjoying-sql-data-pipelines-with-dbt) - [Modern Data Architectures need Software Engineering](https://www.wearedevelopers.com/videos/1030-modern-data-architectures-need-software-engineering) - [Building Multi-Tenant ASP.NET Core Applications: Best Practices and Real-World Solutions](https://www.wearedevelopers.com/videos/1552-building-multi-tenant-asp-net-core-applications-best-practices-and-real-world-solutions) - [AI Model Management Life Circles: ML Ops For Generative AI Models From Research to Deployment](https://www.wearedevelopers.com/videos/1152-ai-model-management-life-circles-ml-ops-for-generative-ai-models-from-research-to-deployment) ## Related Articles - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [What Are Large Language Models?](https://www.wearedevelopers.com/magazine/304-what-are-large-language-models) - [Making Data Warehouses Fast: A Developer’s Story](https://www.wearedevelopers.com/magazine/107-making-data-warehouses-fast-a-developer-s-story) - [MLops – Deploying, Maintaining And Evolving Machine Learning Models in Production](https://www.wearedevelopers.com/magazine/115-mlops-deploying-maintaining-and-evolving-machine-learning-models-in-production) - [Software Engineer Salary London](https://www.wearedevelopers.com/magazine/252-software-engineer-salary-london)