Senior Data Platform Engineer

Stack AV
Pittsburgh, PA, United States
about 2 months ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
7 years minimum
Working hours
Regular working hours
Job source

Tech stack

Artificial Intelligence Computing Platforms Cloud Computing Continuous Integration Data Infrastructure Data Systems Relational Databases Software Debugging Distributed Data Store Distributed Systems Fault Tolerance Python (Programming Language)
+12 more
PostgreSQL Machine Learning MySQL Online Analytical Processing Online Transaction Processing Open Source Technology SQL Databases Data Streaming Feature Engineering Apache Spark Data Lakes Apache Kafka

Job description

Stack is developing revolutionary AI and advanced autonomous systems designed to enhance safety, reliability, and efficiency of modern operations. Stack’s autonomous technology incorporates cutting-edge advancements in artificial intelligence, robotics, machine learning, and cloud technologies, empowering us to create innovative solutions that address the needs and challenges of the dynamic trucking transportation industry. With decades of experience creating and deploying real world systems for demanding environments, the Stack team is dedicated to developing an autonomous solution ecosystem tailored to the trucking industry’s unique demands.

About the Role:

In the Compute Platform team, our mission is to provide the foundational compute platform that powers large-scale autonomous systems development. The team is responsible for enabling engineers and researchers to efficiently run compute and data intensive workloads on Stack AV infrastructure.

The Data Platform team is responsible for designing, implementing and maintaining the Stack AV on-premises data platform. The team supports large scale OLAP/OLTP and feature engineering workloads for multiple Product Development groups across the company. You will work at the intersection of infrastructure, distributed systems, and developer experience-ensuring that our critical services and pipelines are reliable, efficient, and easy to run.

As a Senior Data Platform Engineer, you will design and operate high scale data systems that power engineers across the company.

Responsibilities:

  • Design and operate distributed storage systems for scheduling and executing large-scale batch workloads.
  • Build and maintain an open source, modern data platform.
  • Optimize utilization of storage resourcesImprove reliability and fault tolerance of large-scale storage systems and data platform components.
  • Collaborate with teams across the company to understand workload requirements and improve platform capabilities.
  • Contribute to platform tooling, automation, and CI/CD workflows.

Requirements

Do you have experience in System troubleshooting?, * 7+ years of experience building and operating distributed storage systems or modern data platforms.

  • Experience operating streaming platforms such as Kafka or Pulsar.
  • Fluent in Python, and SQL, with experience writing and maintaining highly available data applications using Trino and Apache Spark.
  • Knowledge of table formats (Iceberg, Delta Lake, Hudi, Xtable).
  • Experience operating and optimizing at least one RDBMS (Postgres, MySQL).
  • Strong debugging and problem-solving skills in complex distributed systems.
  • Ability to collaborate across teams and communicate technical concepts clearly.

About the company

We are proud to be an equal opportunity workplace. We believe that diverse teams produce the best ideas and outcomes. We are committed to building a culture of inclusion, entrepreneurship, and innovation across gender, race, age, sexual orientation, religion, disability, and identity.

Check out our Privacy Policy.

Please Note: Pursuant to its business activities and use of technology, Stack AV complies with all applicable U.S. national security laws, regulations, and administrative requirements, which can restrict Stack AV’s ability to employ certain persons in certain positions pursuant to a range of national security-related requirements. As such, this position may be contingent upon Stack AV verifying a candidate’s residence, U.S. person status, and/or citizenship status. This position may also involve working with software and technologies subject to U.S. export control regulations. Under these regulations, it may be necessary for Stack AV to obtain a U.S. government export license prior to releasing its technologies to certain persons. If Stack AV determines that a candidate’s residence, U.S. person status, and/or citizenship status will require a license, prohibit the candidate from working in this position, or otherwise be subject to national security-related restrictions, Stack AV expressly reserves the right to either consider the candidate for a different position that is not subject to such restrictions, on whatever terms and conditions Stack AV shall establish in its sole discretion, or, in the alternative, decline to move forward with the candidate’s application.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on indeed.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:18 min

Scaling MySQL databases for massive user growth

Johannes Nicolai Johannes Nicolai +1 · LIVE

3:24 min

The governance failures of centralized data lakes

Mario Meir-Huber · LIVE

5:37 min

Extensibility and programmability features of the PostgreSQL database

Silvano Coriani Silvano Coriani · Europe 2026 Virtual

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

6:24 min

Distributed data lakes and containerized computing clusters

Ulrich Wurstbauer +1 · LIVE

1:48 min

Analyzing network packets with database protocol tools

Daniël van Eeden Daniël van Eeden · WWC Europe 2026

Videos

See all

Related articles

See all