> Markdown version of [/jobs/ext/2589464-senior-big-data-engineer](https://www.wearedevelopers.com/jobs/ext/2589464-senior-big-data-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Senior Big Data Engineer - **Company:** Pi-Square Technologies LLC - **Location:** Farmington Hills, MI, United States - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** Java (Programming Language), Application Programming Interfaces (APIs), Airflow, Big Data, Cyber Security, Continuous Integration, Protocol Buffers, JSON, Python (Programming Language), Software Engineering, Product Software Implementation Methods, SQL Databases, Data Streaming, Management of Software Versions, Apache Spark, Change Data Capture, Data Lakes, Integration Tests, Debezium, Data Lineage, Apache Flink, Avro, Qlikview, Apache Kafka, Spark Streaming, Data Pipelines, Apache Beam - **Published:** August 1, 2026 - **Apply:** https://www.careerjet.com/jobad/us011f221076647bd034e2f536057ed100 ## About the Role Job Category: PD Operations and Quality Degree Level: Bachelor's Degree or equivalent Job Description: We made history and now we work to transform the future - for our custo… + 2 months ago ## Description Data Engineer Owns the data pipelines that move data from operational source systems onto the data platform extraction, ingestion, transformation, orchestration, and the day-to-day operational health of those pipelines. Also plays a role as a hands-on builder of Foundational Data Products (raw record-of-truth) and potentially Derivative Data Products (composed from upstream products). Core skills SQL & Python (table stakes); Scala/Java for high-throughput streaming Pipeline build & operations: design, develop, deploy, monitor, and remediate batch and streaming pipelines that land source data on the platform; manage backfills, replays, late-arriving data, schema drift, and SLA breaches Ingestion patterns: full-load, incremental, change-data-capture (Debezium, Fivetran, Qlik Replicate, GoldenGate), event-driven ingest, API and file-based intake Pipeline frameworks: dbt, Apache Spark, Apache Beam, Airflow, Dagster, Prefect Streaming: Kafka, Kinesis, Flink, Spark Structured Streaming Data product packaging: schema contracts (Avro/Protobuf/JSON Schema), versioning, SLAS/SLOs, data contracts, output-port design (SQL, file, API, event) Storage formats: loeberg, Delta Lake, Hudi (open table formats are now the mesh default) Quality & observability: Great Expectations, Soda, Monte Carlo, dbt tests, data lineage (OpenLineage) CI/CD for data: GitOps pipelines, unit + integration tests, environment promotion Al-adjacent vector store ingestion (pgvector, Pinecone), feature stores (Feast, Tecton), RAG-ready chunking and embedding pipelines Mesh-specific: knowing when to build a derivative product vs. extending an existing one; consuming upstream products through governed input ports rather than reaching into source systems Pi-square technologies is a Michigan (USA) Headquartered Automotive Embedded Engineering Services company, Synergy Partner for major OEMs and Tier 1s and their implementation partners in Automotive Embedded Product Development, Projects, Requirements Analysis, Software Design, Software Implementation, Efficient Build, Release Process, and turnkey software V & V Services. We have more than 20+ years of industry expertise with specialization in the latest cutting-edge automotive technologies such as Infotainment, connected vehicles, Cyber security, OTA, and Advanced Safety/ Body electronics. ## Related Videos - [From Messy Queries to Scalable Systems - How Data Engineering actually works](https://www.wearedevelopers.com/videos/100203-from-messy-queries-to-scalable-systems-how-data-engineering-actually-works) - [Alibaba Big Data and Machine Learning Technology](https://www.wearedevelopers.com/videos/37-alibaba-big-data-and-machine-learning-technology) - [Tips and Tricks for Working with JSON](https://www.wearedevelopers.com/videos/1229-tips-and-tricks-for-working-with-json) - [From event streaming to event sourcing 101](https://www.wearedevelopers.com/videos/91-from-event-streaming-to-event-sourcing-101) - [Introducing JSON Structure](https://www.wearedevelopers.com/videos/100219-introducing-json-structure) - [Enjoying SQL data pipelines with dbt](https://www.wearedevelopers.com/videos/823-enjoying-sql-data-pipelines-with-dbt) ## Related Articles - [Top Big Data Technologies That You Need to Know](https://www.wearedevelopers.com/magazine/108-top-big-data-technologies-that-you-need-to-know) - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Making Data Warehouses Fast: A Developer’s Story](https://www.wearedevelopers.com/magazine/107-making-data-warehouses-fast-a-developer-s-story) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [Dev Digest 121 - AI goes offline](https://www.wearedevelopers.com/magazine/456-dev-digest-121-ai-goes-offline)