Databricks Engineer

REVSTAR MEDIA INC.
United States
3 months ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Experienced
Experience required
2 years minimum
Working hours
Regular working hours
Job source

Tech stack

Agile Methodology Artificial Intelligence Airflow Amazon S3 Big Data Cloud Engineering Cloud Storage Data Architecture Data Validation Information Engineering Data Governance Extract Transform Load (ETL)
+26 more
Data Security Data Systems DevOps Python (Programming Language) Machine Learning Meta-Data Management Performance Tuning Azure Data Lake SQL Databases Data Streaming Workflow Management Systems Feature Engineering Data Ingestion Apache Spark Data Lakes Infrastructure Automation Frameworks Data Lineage Low Latency Integration Frameworks Apache Kafka Data Management Machine Learning Operations Terraform Software Version Control Data Pipelines Databricks

Job description

We are seeking a highly skilled Databricks Engineer to join our team. This role will be hands-on, working closely with architects, data scientists, and customers to build, optimize, and deploy high-performance data and AI solutions., As a Databricks Engineer, you will be responsible for building and optimizing data pipelines, implementing data processing frameworks, and enabling AI/ML solutions within Databricks. You will work across data ingestion, transformation, and orchestration while ensuring scalability, performance, and security., Data Engineering & Pipeline Development

  • Develop and optimize data pipelines using Apache Spark and Delta Lake within Databricks.
  • Implement ETL/ELT workflows, ensuring efficient data ingestion, transformation, and storage.
  • Design Lakehouse architecture-based solutions that scale across structured and unstructured data sources.
  • Integrate Databricks with cloud storage solutions (Azure Data Lake, AWS S3, Google Cloud Storage) for seamless data management.

Performance Optimization & Automation

  • Optimize Spark jobs for scalability, cost efficiency, and low latency.
  • Implement monitoring and alerting solutions to track job performance and detect failures.
  • Develop automated data validation, testing, and quality assurance processes.

Management AI/ML Integration & MLOps Support

  • Support ML model training and deployment within Databricks, integrating with MLflow for experiment tracking and model versioning.
  • Collaborate with data scientists and ML engineers to enable scalable AI solutions.
  • Implement feature engineering pipelines and integrate models into production environments.

Security, Governance & Best Practices

  • Ensure data security, access control, and compliance with industry standards (GDPR, HIPAA, SOC 2, etc.).
  • Follow Databricks best practices for data lineage, governance, and metadata management.
  • Document processes, configurations, and best practices for internal and client use.

Requirements

Do you have experience in Spark?, This is a technical hands-on role, requiring expertise in Apache Spark, Delta Lake, and MLOps, as well as experience working with large-scale data architectures. You will collaborate with architects and business stakeholders to ensure solutions align with customer needs and best practices., Must-Have:

  • 3+ years of hands-on experience in data engineering, with a focus on big data processing and cloud-native architectures.
  • 2+ years of hands-on experience with Databricks, including Apache Spark, Delta Lake, and MLflow.
  • Databricks Certifications (Mandatory): *

  • Databricks Certified Data Engineer Associate (or higher)
  • Proficiency in Python, SQL, and Spark-based frameworks.
  • Experience in developing and optimizing large-scale ETL/ELT pipelines.
  • Strong understanding of Lakehouse architecture and cloud-agnostic data solutions.
  • Familiarity with CI/CD pipelines and Infrastructure-as-Code (IaC) for Databricks (e.g., Terraform, Databricks CLI).
  • Knowledge of data governance, security, and compliance best practices.
  • Experience working in Agile development environments, following DevOps/MLOps best practices.

Nice-to-Have:

  • Additional Databricks Certifications (e.g., Databricks Certified Machine Learning Associate).
  • Experience with real-time streaming solutions (e.g., Kafka, Kinesis, Event Hub).
  • Familiarity with cloud storage and orchestration tools (e.g., Apache Airflow, Prefect).
  • Background in AI/ML integration within Databricks, assisting in feature engineering and model deployment.
  • Experience working in client-facing roles or consulting environments.

Benefits & conditions

Pulled from the full job description

  • 401(k)
  • Health insurance
  • Paid time off
  • Vision insurance
  • Dental insurance, * Paid Time Off - Take the time you need to recharge and stay productive.
  • Remote-First Working Environment - Collaborate from anywhere while staying connected with our global team.
  • Comprehensive Health Coverage - Medical, Dental, Vision
  • 401(k) Retirement Plan - Plan for your future with access to a company-sponsored 401(k) program.
  • Annual Learning & Development Stipend - Invest in your skills with conferences, certifications, or courses.
  • Peer Mentorship & Coaching - Learn from experienced engineers, product managers, and architects to accelerate your growth.
  • Professional Growth Opportunities - Exposure to cutting-edge AWS GenAI, data, and cloud technologies across diverse industries.
  • Company Outings & Volunteer Opportunities - Build relationships and give back to the community.
  • Collaborative, Innovative Culture - Work alongside top talent in a fast-paced, supportive environment that values curiosity and initiative.

About the company

RevStar is a Databricks Partner launching a cloud-agnostic practice focused on Data, Machine Learning (ML), and Artificial Intelligence (AI) services. Our mission is to help businesses modernize their data platforms, optimize analytics workflows, and implement scalable AI-driven solutions using Databricks.

We are passionate about what we build and how we build it. From architecture and design to coding and delivery, we approach each project with an agile mindset, continuously analyzing goals and business needs to ensure optimal outcomes.

At RevStar, we foster a collaborative, remote-first culture where teams freely share ideas, innovate together, and grow both individually and collectively. By joining us, you’ll have the opportunity to work with cutting-edge technologies across diverse industries, delivering value-driven products for clients who prioritize quality and performance. We believe in the pursuit of better, not just in cloud-native app development, but in creating meaningful experiences and outcomes that matter.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on indeed.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:15 min

Empowering domain teams with an open data platform

Sandhya Menon Sandhya Menon · WWC Europe 2026

3:09 min

Balancing data science skillings alongside systems engineering rigor

Nico Schmidt · LIVE

2:17 min

Mapping the maturity roadmap for scaled devops adoption

Dominik Krichbaum Dominik Krichbaum · WWC Europe 2026

3:28 min

Defining big data and machine learning fundamentals

Ayon Roy · LIVE

3:18 min

Scaling global network engineering through DevOps culture

Stuart Clark · LIVE

3:05 min

Audience questions on AI agents and pipeline vectorization

Joy Joy · WWC 2024

Videos

See all

Related articles

See all