Sr. Data Engineer

BambooHR LLC
United States
10 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Working hours
Regular working hours
Job source

Tech stack

Artificial Intelligence Amazon Web Services Software Applications Big Data Business Systems Cloud Database Information Systems Information Engineering Data Governance Data Mart Data Systems Data Warehousing
+31 more
Cursor (Graphical User Interface Elements) Software Design Documents Document-Oriented Databases Python (Programming Language) Machine Learning Netsuite Query Optimization Salesforce.Com SQL Databases Data Streaming Unstructured Data Management of Software Versions Feature Engineering Data Ingestion Retrieval-Augmented Generation Large Language Models Multi-Agent Systems Generative AI Change Data Capture Git Data Lakes AI Platforms Pyspark Core Data Information Technology Marketo Data Management Machine Learning Operations Data Pipelines Databricks Zuora

Job description

As a Senior Data Engineer, you will play a key role in designing, building, and operating scalable data platforms, analytics systems, and AI/ML infrastructure. We’ll rely on your expertise across data, analytics, ML, and AI engineering to develop, automate, and maintain pipelines and intelligent systems.

Your ability to use AI in building reliable, performant, and scalable data, ML, and AI systems-effectively building and leveraging AI agents and agentic workflows-will be critical to your success.

You will:

  • Collaborate with data analysts, data scientists, ML engineers, software engineers, and business stakeholders to enable effective use of core data assets
  • Design, develop, and maintain scalable data ingestion and transformation pipelines using Python, SQL, and modern data tooling
  • Build and optimize data lake, lakehouse, warehouse, and data mart architectures
  • Develop and maintain data models including facts, dimensions, feature datasets, and domain-specific data products
  • Translate business requirements into design documents (e.g., ERDs, data flow diagrams) data models and ML feature pipelines
  • Design and manage cloud-based data and ML infrastructure (Databricks preferred), including development, staging, and production environments
  • Design, build, and operationalize machine learning pipelines for training, validation, deployment, and observability (e.g., performance, drift, reliability)
  • Support ML model lifecycle management, including versioning, reproducibility, and lineage
  • Develop and maintain ML feature stores and reusable feature pipelines for ML models
  • Build and integrate AI-powered applications and agentic workflows (e.g., LLM-based agents, retrieval-augmented generation systems, workflow automation agents)
  • Design and implement data pipelines for AI systems, including unstructured data (text, logs, embeddings, vector stores)
  • Develop and maintain unit, integration, and data quality tests
  • Participate in peer code reviews, pull requests, and team coding standards
  • Document data pipelines, ML pipelines, models, infrastructure, and standard operating procedures
  • Define infrastructure as code and support CI/CD pipelines for data and ML systems
  • Ensure data privacy, security, and access control best practices (including AI data governance considerations)
  • Identify and implement improvements in efficiency, scalability, resilience, and performance
  • Contribute to evolving data, ML, and AI platform architecture, tools, and best practices

You’ll help power analytics, machine learning, and intelligent decision-making across domains such as finance, marketing, sales, product, and customer experience., Our process utilizes AI as an assistant to efficiently process and analyze candidate data. Recruiters and hiring managers maintain full oversight and accountability, ensuring that all final selection and rejection decisions are human-made and based solely on objective job qualifications. Please see our General Privacy Notice and California Privacy Notice for more details.

Requirements

  • Ability to gather requirements and translate business processes into data, ML, and AI solutions
  • Comfortable working cross-functionally with both technical and non-technical stakeholders
  • Ability to quickly learn new domains and technologies

Core Technical Skills

  • Strong Python development experience
  • Advanced SQL development and query optimization skills
  • Understanding of Databricks and large-scale data processing
  • Experience building and scaling data pipelines using Databricks and PySpark
  • Deep understanding of data lake, lakehouse, data warehouse, and data mart architectures
  • Experience with data modeling across a variety of business domains
  • Experience with modern data tooling (e.g., dbt or similar transformation frameworks)
  • Knowledge of data formats, data patterns, and modeling best practices
  • Experience with cloud platforms (AWS preferred)
  • Experience with CI/CD pipelines in a data engineering environment
  • Git-based development workflows

AI, ML & MLOps Skills

  • Hands-on experience with AI prompt and agent frameworks (e.g., Claude Code, Cursor, Windsurfer, or similar)
  • Experience building AI agents and agentic workflows
  • Exposure to LLMs, embeddings, vector databases, or generative AI systems
  • Familiarity with handling structured and unstructured data (e.g., text, logs, embeddings)
  • Experience building or supporting machine learning pipelines in production
  • Familiarity with AI and MLOps in Databricks
  • Experience with ML feature engineering and feature stores
  • Understanding of ML model lifecycle management, monitoring, and evaluation

Beyond technical skills, we’re looking for someone who is:

  • A clear and effective communicator
  • Analytical and pattern-oriented
  • Creative in both engineering and AI/data modeling approaches
  • Detail-oriented and persistent in solving problems
  • Comfortable working in a fast-paced, dynamic environment
  • Passionate about learning and continuously improving-especially in the evolving AI/ML landscape
  • Bachelor’s degree in computer science, information systems, a quantitative field, or equivalent practical experience

What Will Make Us REALLY Love You

  • Familiarity with common business metrics across multiple domains
  • Exposure to business systems like Netsuite, Salesforce, Marketo, Zuora, Gainsight, or Pendo
  • Experience building KPI frameworks or domain-specific models (e.g., attribution, funnel, retention, financial metrics)
  • Experience with streaming data and change data capture (CDC)
  • Experience with real-time ML inference systems

Benefits & conditions

  • A Great Company Culture that has been recognized by multiple organizations like Inc, and Salt Lake Tribune
  • Comprehensive health, life, and disability insurance
  • Generous leave policies that include 4 weeks of vacation, 12 company holidays, parental leave, and volunteer time off so you can enjoy quality of life
  • 401k plans with up to 6% company match
  • $2000 Paid-Paid Vacation bonus
  • EAP through Headspace
  • Check out all our benefits that benefit you

About the company

At BambooHR, we’re all about setting people free to do great work, and we believe AI is a powerful partner in that mission. We’re leaning into intelligent tools to streamline our workflows, giving us more time for high-impact innovation. We look for curious, forward-thinking people who are ready to explore how AI can elevate their work and help us reimagine the future of HR., At BambooHR, we’re building something different: we’re building a people intelligence platform that transforms HR and sets people free to do great work! We’re a proven market leader driving innovation while building lasting success through thoughtful, sustainable growth. Here, you’ll find a place that champions growth: both professional and personal, both individual and collective.

We invest in potential, giving you the space to stretch your capabilities and turn good ideas into reality while providing the safety net of a supportive, values-driven culture. Our approach combines meaningful work with meaningful lives, offering competitive benefits, professional development, and the flexibility to thrive both in and outside the office.

What sets us apart isn’t just what we do, but how we do it: with openness, integrity, and a shared commitment to doing the right thing. Join us in creating HR software that makes work better for everyone, while we make work better for you.

BambooHR is committed to the full inclusion of all qualified individuals and will ensure that persons with disabilities are provided reasonable accommodations throughout the hiring process. If you would like to request accommodations, please let your recruiter know.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.dice.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

6:21 min

Investigating push inefficiencies with upstream Git experts

Jonathan Creamer · Coffee With Developers

4:32 min

Harnessing Spark with Python using PySpark and Py4J

Ayon Roy · LIVE

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

56 sec

Favorite git commands and the importance of patch commits

Eileen Uchitelle Eileen Uchitelle +1 · Coffee With Developers

3:09 min

Balancing data science skillings alongside systems engineering rigor

Nico Schmidt · LIVE

3:37 min

Scaling machine learning pipelines from prototypes to petabytes

Julian Joseph · LIVE

Videos

See all

Related articles

See all