Software Engineer, Data Engineering

Poshmark, Inc.
Redwood City, CA, United States
2 months ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
2 years minimum
Working hours
Regular working hours
Job source

Tech stack

Java (Programming Language) Airflow Amazon Web Services Amazon S3 Big Data C++ (Programming Language) Software Quality Information Engineering Data Structures Data Warehousing Fault Tolerance Apache Hadoop
+19 more
MapReduce Apache HBase Apache Hive Python (Programming Language) PostgreSQL Object-Oriented Software Development Software Engineering SQL Databases Data Processing Apache Spark Electronic Medical Records Data Lakes Druid Apache Flink Apache Kafka Stream Processing Data Pipelines Jenkins Amazon Redshift

Job description

  • Build highly scalable, available, fault-tolerant data processing systems using AWS technologies, MapReduce, Hive, Kafka, Spark, Flink and other big data technologies to handle batch and real-time data processing over 10s of terabytes of data ingested every day and a petabyte-sized data warehouse.
  • Responsible for taking care of some of the most critical data pipelines at Poshmark.
  • Participate in architecture discussions, influence product roadmap, build best practices, take ownership and responsibility over new projects.
  • Create WBS (Work breakdown structure) for projects and execute them, taking full ownership, with minimal guidance.

Requirements

  • 3+ years of relevant software engineering experience.
  • 2+ year of Data Engineering experience.
  • Excellent technical problem solving using data structures and algorithms, with emphasis on optimization and code quality.
  • Expertise in architecting and building large-scale data processing systems using Big Data technologies like Hadoop, Hbase, Spark, Kafka, Druid, Flink, DataLake, Redshift etc.
  • Eagerness to try out newer technologies and adopt them as and when needed.
  • Object Oriented Patterns, Scala / Java / Python / C++ / SQL / Apache Spark, Flink, Hadoop, AWS S3, Redshift / Postgresql, Kinesis, EMR, Apache Airflow, Jenkins.

About the company

Poshmark is the leading fashion marketplace where style comes alive through discovery, self-expression, and human connection. Powered by a vibrant community of 165 million members, Poshmark brings real people and taste to shopping through a social experience shaped by shared discovery. Buying and selling fashion feels simple, joyful, and personal, while every item tells its own story. Poshmark empowers sellers to grow meaningful businesses, keeps fashion in circulation longer, and gives shoppers access to unique and trusted finds, from everyday pieces to one-of-a-kind vintage and luxury.

The Big Data team is a central player in the Poshmark organization. Our mission is to build a world-class big data platform to bring value out of data for us and for our customers. Our goal is to democratize data, support exploding business, provide reporting and analytics self-service tools, and fuel existing and new business critical initiatives.

The Data Engineering team at Poshmark is looking for an experienced software engineer to scale Datalake, ensuring real-time access to quality data for all the stakeholders. The role provides opportunity to showcase in-depth software development skills to build and maintain real-time and batch data pipelines with a focus on scalability and optimizations while collaborating with Data Science, Analytics, and Platform Engineering teams to build tools to access petabyte scale data.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on dice.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:15 min

Empowering domain teams with an open data platform

Sandhya Menon Sandhya Menon · WWC Europe 2026

1:02 min

Applying an ETL methodology to infrastructure configuration management

Axel Barbier · WWC 2023

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

3:05 min

Audience questions on AI agents and pipeline vectorization

Joy Joy · WWC 2024

1:34 min

Pivoting careers into specialized platform engineering roles

Xavier Portilla Edo · LIVE

57 sec

Extracting API schemas automatically during continuous integration builds

Axel Barbier · WWC 2023

Videos

See all

Related articles

See all