Data Engineer

DATA ENGINEERING, LLC
San Jose, CA, United States
10 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
4 years minimum
Compensation
$210,800.0 - $272,800.0
Working hours
Regular working hours
Job source

Tech stack

Artificial Intelligence Airflow Data Analysis BigQuery Code Generation Databases Information Engineering Data Files Data Infrastructure Extract Transform Load (ETL) Data Mart Data Warehousing
+7 more
Python (Programming Language) Machine Learning Systems Development Life Cycle Cloud Services Software Engineering SQL Databases Core Data

Job description

Thumbtack is the one app you need to take care of and improve your home - from personalized guidance to AI tools and a best-in-class hiring experience. Every day in every county of the U.S., people turn to Thumbtack to complete urgent repairs, seasonal maintenance and bigger improvements. We help homeowners know which projects to do, when to do them and who to hire from our growing community of 300,000 local service businesses. If making an impact inspires you, join us. Imagine what we’ll build together. About the Data Engineering Team

Thumbtack’s Data Engineering org is split into 3 groups: Core Data Engineering, Data Platform and Embedded Data Engineering. This role is on the Embedded Data Engineering team, which breaks down into multiple business units that data engineers are embedded into. This role would be on the commercial operations side of the house and will work closely with engineers, analysts, data scientists and machine learning engineers to help design and curate data sets originating from internal and third-party sources to meet current and future needs. Over the next year, it will continue to build on its prior successes in building a more cohesive data warehouse while starting to work more deeply upstream to build data best practices into the full software development lifecycle (SDLC).

The challenge

There are several teams all over Thumbtack with Terabytes of data and unique challenges trying to clean and organize this data to measure their performance. In this role, you will work with Engineers, Data Scientists, Managers, and others to understand their needs, and actively work to build datasets to tackle these challenges.

What you’ll do

  • Collaboratively refine and evangelize a comprehensive framework for integrating data-thinking into the software development lifecycle for product teams.
  • Design, architect, and maintain core operations datasets, data marts, and feature stores that support a blend of mature products and features with a rapidly evolving product line, in partnership with analytics, data science, and machine learning.
  • Integrate deeply with our cross functional partners to understand their data needs, and help design datasets with the same engineering rigor as any other software we design.
  • Drive data quality and best practices across different business areas.
  • Help build the next generation data products at Thumbtack, leveraging AI models for code generation and incorporating agents into our workflows.

Requirements

  • 4+ years of experience designing and building data sets and warehouses.
  • Excellent ability to understand the needs of and collaborate with stakeholders in other functions, especially Analytics, and identify opportunities for process improvements across teams.
  • Expertise in SQL for analytics/reporting/business intelligence and also for building SQL- and Python-based transforms inside an ETL pipeline, or similar.
  • Experience designing, architecting, and maintaining a data warehouse and data marts that seamlessly stitches together data from production databases, clickstream data, and external APIs to serve multiple stakeholders.
  • Expertise building the above with a modern data stack based on a cloud-native data warehouse, in our case we use BigQuery, dbt, and Apache Airflow, but a similar stack is fine.
  • Experience using AI to generate design plans, code and documentation as well as applying AI-enabled workflows to accelerate development velocity and improve data engineering practices.
  • Strong sense of ownership and pride in your work, from ideation and requirements-gathering to project completion and maintenance.

About the company

  • For candidates living in San Francisco / Bay Area, San Jose, New York City, or Seattle metros, the expected salary range for the role is currently $210,800.00 - $272,800.00.
  • For candidates living in Austin, TX or Washington DC metros or in California, Massachusetts, New Jersey, or Washington states, the expected salary range for the role is currently $189,600.00 - $245,300.00.
  • For candidates living in all other US locations, the expected salary range for this role is currently $179,400.00 - $232,100.00.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.dice.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:15 min

Empowering domain teams with an open data platform

Sandhya Menon Sandhya Menon · WWC Europe 2026

4:21 min

Challenges of traditional mobile data synchronization

Timothy Marland · WWC 2023

3:27 min

Explaining query execution overhead and caching limitations in BigQuery

Adnan Rahic · JS Congress

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

2:57 min

Core data types and collection structures in Clojure

Tobias Schröder · WWC 2021

3:05 min

Audience questions on AI agents and pipeline vectorization

Joy Joy · WWC 2024

Videos

See all

Related articles

See all