> Markdown version of [/jobs/ext/1907514-data-infrastructure-quality-engineering](https://www.wearedevelopers.com/jobs/ext/1907514-data-infrastructure-quality-engineering). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Data Infrastructure / Quality Engineering - **Company:** Ouster, Inc. - **Location:** San Francisco, CA, United States - **Experience:** Expert - **Salary:** $140,000.0 - $200,000.0 - **Contract:** Permanent contract - **Skills:** Cloud Computing, Cloud Engineering, Data Validation, Data Governance, Data Infrastructure, Extract Transform Load (ETL), Data Systems, Regression Testing, Unstructured Data, Core Data, Integration Tests, Data Lineage, Machine Learning Operations, Software Version Control, Data Pipelines - **Published:** August 3, 2026 - **Apply:** https://www.dice.com/job-detail/02652154-2fdd-4392-be4c-addcc4b1bc0f ## About the Role The Industrial Autonomy team is looking for a self-starter who can independently drive complex data systems from conception to completion with a high degree of autonomy, transforming raw, multi-sensor streams into robust, reproducible training datasets for our AI pipelines., * 8+ years of experience designing, building, and validating scalable cloud infrastructure and data pipelines * Experience building and testing data systems to enterprise-grade standards capable of processing massive, unstructured datasets at a production scale * Extensive knowledge of modern data infrastructure, cloud platforms, and data quality validation frameworks * Experience ramping at least one core data platform from initial prototype to production release + supporting its long-term stability Preferred experience * Direct experience with autonomy, robotics, industrial equipment, or automotive data loops, specifically handling massive streams of multimodal vehicle telemetry and sensor data * Experience building and validating active learning pipelines, continuous training infrastructure, and automated data curation systems * Experience with data governance, safety-critical data validation frameworks, or compliance standards for autonomous systems * Experience deploying and optimizing high-performance GPU cloud inference services, with specific expertise utilizing the NVIDIA architecture (e.g., Triton) * Experience collaborating with data labeling services, including internal labeling, third-party labeling vendors, and integrating external annotation services The base pay will be dependent on your skills, work experience, location, and qualifications. This role may also be eligible for equity & benefits. ($140,000 - $ 200,000) ## Description * Design and develop robust cloud infrastructure, storage systems, and automated testing frameworks for AI training datasets and machine learning pipelines * Own data infrastructure from concept through prototype architecture, data quality validation, and production-scale release * Experience architecting and validating data lakehouse/warehouse systems, feature stores, and automated data governance frameworks to ensure data lineage, security, and reproducible training datasets * Partner with SW and ML engineers to build and optimize sensor data ingestion, model/data/label versioning systems, cloud orchestration, and high-throughput pipeline architectures * Develop automated data validation scripts, core ETL pipelines, infrastructure-as-code (IaC), and comprehensive regression testing suites * Support pipeline deployments, continuous architectural iteration, and root cause analysis for data corruption, pipeline bottlenecks, or infrastructure failures * Support data-tooling integration, automated data labeling workflows, and third-party vendor integration testing * Contribute to pilot data deployments and field telemetry loops, incorporating learnings into future architectural designs ## Related Videos - [From Messy Queries to Scalable Systems - How Data Engineering actually works](https://www.wearedevelopers.com/videos/100203-from-messy-queries-to-scalable-systems-how-data-engineering-actually-works) - [Enjoying SQL data pipelines with dbt](https://www.wearedevelopers.com/videos/823-enjoying-sql-data-pipelines-with-dbt) - [Industrializing your Data Science capabilities](https://www.wearedevelopers.com/videos/178-industrializing-your-data-science-capabilities) - [Modern Data Architectures need Software Engineering](https://www.wearedevelopers.com/videos/1030-modern-data-architectures-need-software-engineering) - [Building Multi-Tenant ASP.NET Core Applications: Best Practices and Real-World Solutions](https://www.wearedevelopers.com/videos/1552-building-multi-tenant-asp-net-core-applications-best-practices-and-real-world-solutions) - [How to develop an autonomous car end-to-end: Robotic Drive and the mobility revolution](https://www.wearedevelopers.com/videos/22-how-to-develop-an-autonomous-car-end-to-end-robotic-drive-and-the-mobility-revolution) ## Related Articles - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Best US AI Conferences for CTOs in 2026: Build vs. Buy, Vendor Evaluation, and Peer Intelligence](https://www.wearedevelopers.com/magazine/736-best-us-ai-conferences-for-ctos-in-2026-build-vs-buy-vendor-evaluation-and-peer-intelligence) - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [Highest Paying Tech Companies in Europe](https://www.wearedevelopers.com/magazine/162-highest-paying-tech-companies-in-europe) - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [Got AI ideas but no money? Here are 10 free ways to level up your AI skills with Google Cloud](https://www.wearedevelopers.com/magazine/600-got-ai-ideas-but-no-money-here-are-10-free-ways-to-level-up-your-ai-skills-with-google-cloud)