> Markdown version of [/jobs/ext/1317627-senior-software-engineer-data-engineering](https://www.wearedevelopers.com/jobs/ext/1317627-senior-software-engineer-data-engineering). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Senior Software Engineer, Data Engineering - **Company:** Omada Health, Inc. - **Location:** United States (Remote available) - **Experience:** Expert - **Salary:** $171,600.0 - $224,300.0 - **Contract:** Permanent contract - **Skills:** Clean Code Principles, Java (Programming Language), Third Normal Form, Application Programming Interfaces (APIs), Artificial Intelligence, Airflow, Amazon Web Services, Amazon S3, Data Analysis, Automation of Tests, Big Data, BigQuery, Cloud Computing, Cloud Database, Computer Programming, Databases, Data Architecture, Data Validation, Information Engineering, Data Governance, Data Infrastructure, Data Integration, Extract Transform Load (ETL), Data Systems, Relational Databases, Cursor (Graphical User Interface Elements), Decision Support Systems, Distributed Data Store, Distributed Systems, EHealth, Fault Tolerance, Graph Database, Python (Programming Language), PostgreSQL, MySQL, Online Analytical Processing, NoSQL, Online Transaction Processing, Performance Tuning, Productivity Software, Ruby on Rails, Raw Data, Amazon Simple Notification Service (SNS), Software Engineering, SQL Databases, Data Streaming, Tableau (Software), Workflow Management Systems, Datadog, Data Processing, Snowflake, Apache Spark, Parallel Computation, Generative AI, Data Lakes, Gitlab-ci, Kubernetes, Information Technology, Low Latency, Apache Kafka, Functional Programming, Amazon Simple Queue Service (SQS), Data Pipelines, Serverless Computing, Docker, Amazon Redshift, Databricks - **Published:** July 17, 2026 - **Apply:** https://www.indeed.com/viewjob?jk=0663114dfd06c9eb ## About the Role * You are passionate about building data-driven systems to enable Data Scientist, Data Analysts and AI/ML Engineers. * You want to make a difference to empower digital healthcare through Data-driven decision making. * You would like to learn how to build scalable, performant and reliable data pipelines. About you: Experience: * 5+ years of experience building, maintaining, and orchestrating scalable data pipelines. * 3+ years of experience as a data engineer developing or maintaining integration with software such as Airflow or any Python-based data pipeline codebase. * Experience applying a variety of integration patterns for different use cases. * Experience in backend software development to contribute to distributed computing development and data technologies, with broad experience across systems, contexts, and ideas. * Experience implementing data pipelines and improving the performance of ETL processes and related SQL queries. * Experience in data modeling for OLTP and OLAP applications * Experience with Cloud platforms such as Amazon AWS * Familiarity with workflow management tools (Airflow preferred) * Familiarity with cloud-based data warehouses (Amazon Redshift preferred) * Exceptional Problem solving and analytical skills * Experience working with sensitive data i.e. PHI / PII & security best practices * Familiarity with data governance practices and principles. Technical Skills: * Proficiency in SQL and experience with relational databases (e.g., MySQL, PostgreSQL). * Proficiency in Analytical SQL (e.g. Analytics Queries, Distributed database queries) and experience working with massive parallel processing (MPP) databases (e.g., Redshift, BigQuery, Snowflake). * Proficiency in programming languages such as Python, Java, or Scala. * Knowledge of data modeling techniques (3NF) and tools (e.g., ER/Studio, ERwin). * Software Engineering Mindset: Apply best practices to write elegant, maintainable code and understand automated testing concepts. * Familiarity with business intelligence tools and environments. * Familiarity with big data technologies (e.g., Lambda, Iceberg or Delta Lake, Spark, or Kafka) * Software engineering mindset and an ability to write elegant, maintainable code while following engineering best practices Communication Skills: Excellent communication and collaboration skills, both written and verbal, with the ability to convey complex technical concepts to non-technical stakeholders. Self-Directed: Leading projects and tasks effectively with cross functional stakeholders minimal guidance. You care about writing quality software and recognize that there are often many right answers. Ability to lead tasks and projects effectively. Education: Bachelor's degree in Computer Science or a similar discipline preferred. Technologies we use: Ruby on Rails, Redshift, Athena, Postgres, SQL, Python, Apache Airflow, Appflow, S3, SNS, SQS, Kafka, Docker, Kubernetes, AWS infrastructure, Lambda, Serverless, Tableau, Bugsnag, Datadog, GitLabCI, Cursor, OpenMetadata, Databricks Bonus Points for: * Experience with NoSQL databases (e.g., document and graph databases) Experience building internal frameworks or development productivity tools would be a big plus. * Experience building data infrastructure, frameworks and automation is a big plus. * Experience using Generative AI tools or technologies to improve engineering workflows, automation, analytics, or data platform capabilities. ## Description We are seeking a highly skilled and motivated Data Engineer to join our team. The ideal candidate will be responsible for designing, building, and maintaining robust data architectures and engineering data models and pipelines. This role will play a critical part in ensuring the integrity, scalability, and performance of our data processing and products., * Data Architecture: Design, develop, and implement scalable, secure, and efficient data solutions that meet the needs of the organization. * Data Modeling: Create and maintain logical and physical data models to support business intelligence, analytics, and reporting requirements. * Pipeline Engineering: Design, build, and optimize ETL (Extract, Transform, Load) processes and data pipelines to ensure smooth and efficient data flow from various sources. * Data Integration: Integrate diverse data sources, including APIs, databases, and third-party data, into a unified data platform. * Performance Optimization: Monitor and optimize the performance of data systems and pipelines to ensure low latency and high throughput. * Data Quality and Governance: Implement data quality checks, validation processes, and governance frameworks to ensure the accuracy and reliability of data. * Collaboration: Partner closely with data scientists, analysts, and other stakeholders to understand data requirements and deliver solutions that meet their needs. * Documentation: Maintain comprehensive documentation of data architectures, models, and pipelines for ongoing maintenance and knowledge sharing. * Training: You'll train and collaborate with teammates effectively in data engineering best practices * Technical Influence/Leadership: Recommends policy changes and establishes department-wide procedures. Uses extensive experience and knowledge to resolve complex problems. * Monitor and manage production environment to deliver data within defined SLAs How you can make an impact: * You'll evaluate, benchmark, and improve the scalability, robustness, and performance of our data platform and applications * You'll make significant contributions to the architecture and design of our data processing platform * You'll implement scalable, fault tolerant, and accurate ETLs frameworks. * You'll gather and process raw data at scale from diverse sources * You'll collaborate with product management, data scientists, analysts, and other engineers on technical vision, design, and planning * You'll implement and maintain a high level of data quality monitoring in our analytics & ML ecosystem * You'll train and collaborate with teammates effectively in data engineering best practices * You will be responsible for leading, documenting, and collaborating across teams for technical projects. ## Related Videos - [From Messy Queries to Scalable Systems - How Data Engineering actually works](https://www.wearedevelopers.com/videos/100203-from-messy-queries-to-scalable-systems-how-data-engineering-actually-works) - [Docker Compose: Rediscovered](https://www.wearedevelopers.com/videos/1978-docker-compose-rediscovered) - [Leveraging Real time data in FSIs](https://www.wearedevelopers.com/videos/806-leveraging-real-time-data-in-fsis) - [MySQL Protocol Features You Should Be Aware Of](https://www.wearedevelopers.com/videos/100267-mysql-protocol-features-you-should-be-aware-of) - [Modern Data Architectures need Software Engineering](https://www.wearedevelopers.com/videos/1030-modern-data-architectures-need-software-engineering) - [Docker build without Docker](https://www.wearedevelopers.com/videos/100114-docker-build-without-docker) ## Related Articles - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [Making Data Warehouses Fast: A Developer’s Story](https://www.wearedevelopers.com/magazine/107-making-data-warehouses-fast-a-developer-s-story) - [Coffee with Developers - Maria Apazoglou - Making AI understandable for all in production](https://www.wearedevelopers.com/magazine/475-coffee-with-developers-maria-apazoglou-making-ai-understandable-for-all-in-production) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Navigating the AI Shift](https://www.wearedevelopers.com/magazine/629-navigating-the-ai-shift)