> Markdown version of [/jobs/ext/1668817-data-engineer-fleet-monitoring-analysis](https://www.wearedevelopers.com/jobs/ext/1668817-data-engineer-fleet-monitoring-analysis). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Data Engineer, Fleet Monitoring & Analysis - **Company:** Coreweave, Inc. - **Location:** United States - **Experience:** Expert - **Salary:** $153,000.0 - $204,000.0 - **Contract:** Permanent contract - **Skills:** Java (Programming Language), Application Programming Interfaces (APIs), Airflow, Amazon Web Services, Business Analytics Applications, Data Analysis, Apache HTTP Server, Microsoft Azure, Big Data, Databases, Data as a Services, Information Engineering, Data Governance, Data Infrastructure, Extract Transform Load (ETL), Data Security, Data Stores, Data Visualization, Data Warehousing, Database Queries, Python (Programming Language), Meta-Data Management, NoSQL, Performance Tuning, SQL Databases, Data Streaming, Data Processing, Google Cloud, Apache Spark, Indexer, Data Lakes, Information Technology, Data Pipelines, Programming Languages - **Published:** July 17, 2026 - **Apply:** https://www.dice.com/job-detail/f1520067-4e4f-4d1b-97c3-995ccc0e274b ## About the Role * Bachelor's degree in Computer Science, Engineering, or a related field. * 4 - 7 years of experience as a Data Engineer or in a similar data-focused role in a fast-paced environment. * Strong SQL skills for data manipulation, modeling, and querying large datasets. * Proficiency in at least one programming language commonly used for data engineering such as Python, Java, or Scala. * Hands-on experience with data pipeline orchestration tools (e.g., Apache Airflow) and big data technologies (e.g., Apache Spark). * Experience designing, operating, and optimizing data lake and/or data warehouse solutions, with a solid understanding of data modeling and performance tuning. * Knowledge of cloud platforms (e.g., AWS, Google Cloud Platform, Azure) and related data services (e.g., object storage, managed databases, analytics services). * Familiarity with database systems (e.g., SQL and NoSQL) and data warehousing concepts, including partitioning, indexing, and schema design. * Experience building, maintaining, and monitoring ETL/ELT pipelines in production environments, including alerting and observability. * Experience creating and maintaining reporting and analytics solutions (dashboards, reports, and metrics) for technical and non-technical audiences. Preferred: * Experience with modern data lakehouse technologies and table formats such as Apache Iceberg (or similar technologies like Delta Lake or Apache Hudi). * Experience with Trino or other distributed SQL query engines at scale. * Experience with Apache Superset or other BI/visualization tools for building self-service analytics. * Experience with data quality frameworks, data observability tooling, and/or metadata management. * Experience supporting executive-level reporting and KPI design in partnership with business and finance stakeholders. Wondering if you're a good fit? We believe in investing in our people, and value candidates who can bring their own diversified experiences to our teams - even if you aren't a 100% skill or experience match. Here are a few qualities we've found compatible with our team. If some of this describes you, we'd love to talk. * You love owning data infrastructure end-to-end-from ingestion and modeling to analytics and visualization. * You're curious about modern data lake and lakehouse architectures and enjoy working with open-source data tooling. * You're an expert at turning loosely defined business questions into concrete data products, metrics, and dashboards that drive decisions. ## Description As a Senior Data Engineer you will own and evolve the data lake and analytics stack that powers observability and decision-making for CoreWeave's global hardware fleet. You'll maintain, monitor, and upgrade our data lake infrastructure (Apache Iceberg, Trino, Apache Airflow, Apache Spark, Apache Superset) and ETL pipelines, while delivering ad-hoc analysis and executive-ready reporting for team, director, and leadership stakeholders. You'll also create visualizations, documentation, and integrations that make fleet monitoring data reliable, discoverable, and actionable across the organization. In this role, you will: * Design, develop, and maintain robust and scalable data pipelines to collect, process, and store data from various sources, including APIs, databases, and third-party services. * Maintain, monitor, and upgrade CoreWeave's data lake infrastructure, including Apache Iceberg, the Trino query layer, Apache Airflow, Apache Spark, and Apache Superset. * Maintain, monitor, and upgrade ETL/ELT pipelines to ensure reliable, performant, and observable data flows across batch and (where applicable) streaming workloads. * Create and optimize data models and data products to support analytics and reporting, ensuring data accuracy, consistency, and performance. * Provide ad-hoc analysis and reporting for team, director, and executive-level stakeholders, translating business questions into data-driven insights and clear narratives. * Create visualizations and dashboards (e.g., in Apache Superset or similar tools) that surface key metrics, trends, and operational KPIs for a variety of internal audiences. * Develop and maintain documentation and runbooks for data pipelines, data lake infrastructure, data models, and usage patterns to support knowledge sharing and troubleshooting. * Implement data security and governance best practices to protect sensitive information and comply with data privacy regulations. * Collaborate with cross-functional teams to integrate data into applications and analytics platforms, helping to visualize performance metrics and identify opportunities for improvement. ## Related Videos - [From Messy Queries to Scalable Systems - How Data Engineering actually works](https://www.wearedevelopers.com/videos/100203-from-messy-queries-to-scalable-systems-how-data-engineering-actually-works) - [Optimizing Discovery: PostgreSQL's Role in Transforming GetYourGuide's Search](https://www.wearedevelopers.com/videos/1647-optimizing-discovery-postgresql-s-role-in-transforming-getyourguide-s-search) - [Leveraging Real time data in FSIs](https://www.wearedevelopers.com/videos/806-leveraging-real-time-data-in-fsis) - [Modern Data Architectures need Software Engineering](https://www.wearedevelopers.com/videos/1030-modern-data-architectures-need-software-engineering) - [Enjoying SQL data pipelines with dbt](https://www.wearedevelopers.com/videos/823-enjoying-sql-data-pipelines-with-dbt) - [AI Model Management Life Circles: ML Ops For Generative AI Models From Research to Deployment](https://www.wearedevelopers.com/videos/1152-ai-model-management-life-circles-ml-ops-for-generative-ai-models-from-research-to-deployment) ## Related Articles - [Making Data Warehouses Fast: A Developer’s Story](https://www.wearedevelopers.com/magazine/107-making-data-warehouses-fast-a-developer-s-story) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [Where To Find Software Engineering Jobs](https://www.wearedevelopers.com/magazine/396-where-to-find-software-engineering-jobs) - [How to Answer the Interview Question: “Why Do You Want to Be a Software Engineer?”](https://www.wearedevelopers.com/magazine/392-how-to-answer-the-interview-question-why-do-you-want-to-be-a-software-engineer)