> Markdown version of [/jobs/ext/542217-senior-data-engineer](https://www.wearedevelopers.com/jobs/ext/542217-senior-data-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Senior Data Engineer - **Company:** Daybreak AI, Inc. - **Location:** San Francisco, CA, United States (Remote available) - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Airflow, Amazon Web Services, Amazon S3, Cloud Computing, Databases, Information Engineering, Software Debugging, DevOps, Distributed Data Store, Distributed Systems, Hadoop Distributed File System, Apache Hive, Python (Programming Language), PostgreSQL, Machine Learning, Operational Databases, Sql Optimization, Apache Spark, Containerization, Kubernetes, Information Technology, Deployment Automation, Machine Learning Operations, Data Pipelines, Docker - **Published:** June 12, 2026 - **Apply:** https://www.indeed.com/viewjob?jk=dc750ce13952bfd8 ## About the Role * 5+ years of experience in data engineering, with demonstrated ownership of production pipelines * Strong proficiency in Python for data engineering workloads * Advanced SQL skills across multiple databases (PostgreSQL and others) * Hands-on experience with pipeline orchestration tools such as Apache Airflow or Dagster * Experience working with distributed data systems - Spark, Hive, or HDFS * Solid understanding of containers and Docker in a development/deployment context * Strong debugging skills and systematic approach to resolving data quality issues * Bachelor's or advanced degree in Computer Science, Engineering, or a related field * Ability to thrive in a fast-paced, evolving environment and pick up new technologies quickly Helpful to Have * Cloud experience, preferably AWS (S3, EMR, or equivalent services) * Experience processing large-scale time series data * Familiarity with Kubernetes for containerized workloads * Understanding of ML model lifecycle and MLOps pipelines * Experience working in interdisciplinary teams across engineering and data science ## Description As a Senior Data Engineer on our India engineering team, you will own the design and delivery of scalable, reliable data pipelines that process real-world supply chain data at scale. You will work closely with Data Science, DevOps, and Customer Success teams to deploy and evolve our core platform - and your hands-on experience with production data will directly shape product decisions. This is a high-ownership role. You will be expected to solve ambiguous problems, mentor junior engineers, and contribute to the technical direction of the team. What You'll Do * Design, build, and maintain production-grade data pipelines handling large-scale supply chain datasets * Own end-to-end data modeling, warehousing architecture, and pipeline reliability across multiple customer environments * Collaborate with Data Science, DevOps, Infrastructure, and Customer Success teams to deliver and iterate on product deployments * Debug complex pipeline failures and performance bottlenecks across distributed systems * Drive improvements to engineering practices, tooling, and deployment automation * Mentor and support junior data engineers, fostering technical growth within the team * Contribute to Daybreak's culture of continuous learning and operational excellence ## Related Videos - [From Messy Queries to Scalable Systems - How Data Engineering actually works](https://www.wearedevelopers.com/videos/100203-from-messy-queries-to-scalable-systems-how-data-engineering-actually-works) - [From DevOps to Scaled DevOps: How We’re Rebuilding Continuous Delivery as a Platform](https://www.wearedevelopers.com/videos/100018-from-devops-to-scaled-devops-how-we-re-rebuilding-continuous-delivery-as-a-platform) - [Docker Compose: Rediscovered](https://www.wearedevelopers.com/videos/1978-docker-compose-rediscovered) - [Alibaba Big Data and Machine Learning Technology](https://www.wearedevelopers.com/videos/37-alibaba-big-data-and-machine-learning-technology) - [Enjoying SQL data pipelines with dbt](https://www.wearedevelopers.com/videos/823-enjoying-sql-data-pipelines-with-dbt) - [AI Model Management Life Circles: ML Ops For Generative AI Models From Research to Deployment](https://www.wearedevelopers.com/videos/1152-ai-model-management-life-circles-ml-ops-for-generative-ai-models-from-research-to-deployment) ## Related Articles - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Coffee with Developers - Maria Apazoglou - Making AI understandable for all in production](https://www.wearedevelopers.com/magazine/475-coffee-with-developers-maria-apazoglou-making-ai-understandable-for-all-in-production) - [Navigating the AI Shift](https://www.wearedevelopers.com/magazine/629-navigating-the-ai-shift) - [Top Big Data Technologies That You Need to Know](https://www.wearedevelopers.com/magazine/108-top-big-data-technologies-that-you-need-to-know)