Data Engineer

Vagaro, Inc.
Pleasanton, CA, United States
1 day ago
Apply on www.wayup.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Experienced
Experience required
3 years minimum
Compensation
$135,000.0 - $150,000.0
Working hours
Regular working hours
Job source

Tech stack

Adaptable Database Systems Amazon Web Services Microsoft Azure Code Review Information Systems Data Architecture Information Engineering Data Governance Data Infrastructure Data Integration Extract Transform Load (ETL) Data Migration
+29 more
Data Stores Data Systems DevOps Dimensional Modeling Elasticsearch Github Python (Programming Language) Machine Learning Meta-Data Management Microsoft SQL Server MongoDB NoSQL Azure Data Lake SQL Databases Enterprise Data Management Parquet Data Logging Data Processing Apache Spark Pandas Microsoft Fabric Data Lakes Pyspark Information Technology Deployment Automation Data Management Restful APIs Data Pipelines Docker

Job description

**Why Vagaro? **At Vagaro, we believe in fostering a collaborative and inclusive work environment where every team member can thrive. Our culture is built on innovation, continuous learning, and a passion for making a positive impact. We support our employees’ growth and vision for themselves, offering opportunities for professional development and career advancement. Join us and be part of a team that values creativity, teamwork, and a commitment to excellence. Plus, we know how to have fun while getting the job done! What you’ll Be Doing: Vagaro is seeking a skilled Data Engineer to join our dynamic team and partner with our Data Science, Machine Learning, and Business Intelligence teams to build enterprise-grade data solutions. As a Data Engineer at Vagaro, you will play a critical role in designing and maintaining scalable data architectures on the Microsoft Fabric platform. You will lead the design and optimization of ETL/ELT pipelines, establish data governance standards, and ensure seamless data integration across multiple platforms and data sources. Your expertise will directly enable our teams to deliver reliable, high-performance data infrastructure that supports real-time analytics, machine learning model training, and business intelligence initiatives. ** *** This role is based onsite in Pleasanton, CA Monday through Friday*, + Design, build, and maintain scalable ETL/ELT pipelines to ingest, transform, and store data from a variety of internal and external data sources.

  • Optimize data pipelines for performance, scalability, reliability, and cost efficiency.
  • Monitor, troubleshoot, and resolve data pipeline failures and performance issues.
  • Develop reusable Python libraries and frameworks to standardize data engineering solutions.
  • Build and maintain CI/CD pipelines to automate deployment, testing, and release processes.
  • Experience integrating with REST APIs to ingest, transform, and process data from external systems. Microsoft Fabric Platform

  • **Design, implement, and maintain enterprise data solutions using the Microsoft Fabric platform.
  • Build and optimize Fabric Warehouses, Lakehouses, Semantic Models, and Data Pipelines.
  • Define enterprise Fabric architecture, workspace strategy, security model, and governance standards.
  • Manage Fabric capacities, optimize compute utilization, and monitor platform health and performance.
  • Lead capacity planning, workload isolation, reservation management, and cost optimization across Business Intelligence, Data Engineering, and Machine Learning workloads.
  • Implement monitoring, alerting, logging, and operational best practices for Fabric environments. Data Architecture & Management

  • Design and maintain scalable data models and enterprise data architectures.
  • Implement data governance, security, lineage, and metadata management best practices.
  • Ensure data quality, integrity, availability, and compliance across all data platforms.
  • Maintain comprehensive documentation for data pipelines, architectures, standards, and operational procedures. Collaboration

  • Partner with Data Scientists and Machine Learning engineers to provide reliable data infrastructure for model training and deployment.
  • Collaborate with Business Intelligence teams to ensure timely and reliable data availability for reporting, dashboards, and analytics.
  • Work closely with business stakeholders, analysts, and software engineers to understand requirements and deliver scalable data solutions. Continuous Improvement

  • Evaluate emerging technologies and recommend improvements to the data platform.
  • Continuously improve data engineering processes, automation, and operational efficiency.
  • Participate in architecture reviews, code reviews, and mentoring to promote engineering best practices. **What you Bring

Requirements

  • Bachelor’s or Master’s degree in Computer Science, Information Systems, Engineering, or a related technical field. Experience

  • **3-5 years of experience as a Data Engineer or similar role.
  • Experience designing and maintaining enterprise-scale ETL/ELT solutions.
  • Hands-on experience with Microsoft Fabric, including data migration projects.
  • Strong experience with relational and NoSQL databases.
  • Experience building modern cloud-based data platforms in Azure (AWS or GCP experience is also beneficial).
  • Experience with DevOps practices, CI/CD pipelines, and Infrastructure as Code. Technical Skills

  • Expert-level SQL.
  • Strong Python development skills.
  • Experience with: o PySpark o Pandas o Spark

  • Strong understanding of data modeling, dimensional modeling, and schema design.
  • Experience working with: o SQL Server o NoSQL (MongoDB, Elasticsearch, etc.) o Azure Data Lake Storage o Delta Lake o Parquet

  • Familiarity with REST APIs and data integration patterns.
  • Experience working with large-scale structured and semi-structured datasets. Soft Skills

  • Excellent analytical and problem-solving abilities.
  • Strong communication and collaboration skills.
  • Ability to work independently while collaborating across multiple teams.
  • Strong attention to detail and commitment to data quality. Preferred Qualifications

  • Experience with Azure DevOps or GitHub Actions.
  • Experience with Docker and Kubernetes.
  • Experience implementing monitoring, observability, and alerting for enterprise data platforms.

Benefits & conditions

  • Base Annual Salary: $135,000 to $150,000, + **Attractive Compensation & Performance Bonuses: **Enjoy a competitive salary paired with performance-based bonuses
  • Generous PTO: 15 accrued days, plus 10 company holidays annually.
  • Health & Wellness: Comprehensive healthcare, dental, and vision plans for you and your family.
  • Exclusive Perks: Discounts on attractions, theme parks, shows, sports events, movies, hotels, and more through TicketsAtWork.
  • Beauty Perks: $30/month reimbursement for any Vagaro service, including health, beauty, or wellness treatments.
  • Food Perks: $50 monthly stipend for our onsite microkitchen and a complimentary DoorDash DashPass subscription.
  • Growth Opportunities: College Assistance Reimbursement, access to EAP & Work/Life Programs, and a LinkedIn Learning account.
  • Financial Security: 401k program with 4% matching and optional life/supplemental insurance.
  • Stay Active: Access to our on-site gym, flavored water dispenser, and basketball court to keep you fit and energized! Equal Opportunity Employer: Vagaro is proud to be an Equal Employment Opportunity and affirmative action employer. We foster an inclusive environment where individuals are evaluated without discrimination based on gender, race, ethnicity, age, disability, religion, sexual orientation, gender identity, veteran status, or any other characteristics protected by law. Privacy Policy

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.wayup.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:37 min

Comparing traditional SQL tables versus NoSQL non-tabular databases

Stanimira Vlaeva · JS Congress

2:07 min

Inspecting default bridge architectures and custom Docker networks

Oliver Seitz Oliver Seitz · World Congress 2025

2:03 min

Accelerating pandas dataframes using cudf module plugins

Ankit Patel Ankit Patel · World Congress 2024

3:09 min

Balancing data science skillings alongside systems engineering rigor

Nico Schmidt · LIVE

3:16 min

Terminology differences between relational and NoSQL databases

Tim Faulkes · LIVE

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

Videos

See all

Related articles

See all