Data Engineer - Databricks

Data Inc
United States
2 months ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
2 years minimum
Working hours
Regular working hours
Job source

Tech stack

Agile Methodology Airflow Amazon Web Services Microsoft Azure Big Data Cloud Database Continuous Integration Data as a Services Data Architecture Information Engineering Data Governance Data Infrastructure
+21 more
Data Security Software Debugging Distributed Computing Environment PostgreSQL MySQL NoSQL Data Processing Google Cloud Apache Spark Backend Data Lakes Pyspark Kubernetes Infrastructure Automation Frameworks Data Lakehouse Data Delivery Terraform Software Version Control Data Pipelines Docker Databricks

Job description

We are seeking a skilled and motivated Data Engineer with deep Databricks and cloud experience to support our Federal Services team. In this role, you will help design and implement scalable data solutions, build and optimize pipelines, and automate infrastructure within secure cloud environments.

This role is ideal for experienced developers who enjoy tackling complex data engineering challenges, optimizing performance, and contributing to impactful government programs. You’ll help shape robust, secure, and efficient pipelines that serve critical public-sector missions.

Day-to-Day Impact

  • Design, develop, and optimize robust data pipelines in Databricks to process and transform large-scale data sets
  • Collaborate with data scientists, engineers, and cloud architects to build reliable, secure data systems
  • Automate infrastructure provisioning using Terraform and manage resources on Azure (or AWS/GCP)
  • Monitor, troubleshoot, and support production-grade data pipelines and jobs
  • Develop CI/CD workflows to streamline deployment and version control of data processing code
  • Ensure security, performance, and reliability across all stages of data workflow
  • Utilize and manage Docker containers where needed to deploy data services
  • Participate in Agile team ceremonies and help shape iterative data delivery plans
  • Support continuous improvement in data architecture, automation, and governance practices
  • Other duties as reasonably required

Requirements

Do you have experience in Version control systems?, Do you have a Bachelor’s degree?, * 5+ years of experience in data engineering or backend data infrastructure roles

  • 2+ years of hands-on experience with Databricks, Spark, or distributed data processing systems in a production environment
  • Deep proficiency with PySpark and Python for data engineering tasks
  • Experience with data modeling
  • Solid experience with Airflow, dbt, or other orchestration and transformation tools
  • Proven skills in data governance, lineage, or security best practices
  • Experience working with Azure (preferred), AWS, or Google Cloud
  • Strong working knowledge of CI/CD tools and version control workflows
  • Comfortable working with relational (e.g., PostgreSQL, MySQL) and NoSQL databases
  • Proven experience in monitoring, debugging, and scaling cloud-based data systems
  • Strong communication skills and ability to work independently in a remote team environment

Kudos If You Have:

  • Experience with Kubernetes and Helm for deploying data workloads (not required)
  • Solid experience with Terraform or other Infrastructure as Code tools
  • Familiarity with containerization using Docker
  • Knowledge of data lakehouse architectures and Delta Lake
  • Experience working on Federal or government projects
  • Existing security clearance (preferred but not required)
  • Databricks Certification, * Bachelor’s (Preferred)

Experience:

  • data engineering: 7 years (Preferred)
  • Databricks: 3 years (Preferred)

Benefits & conditions

Pulled from the full job description

  • 401(k) 4% Match
  • 401(k)
  • Health insurance
  • Retirement plan
  • 401(k) matching
  • Paid time off
  • Vision insurance, Data Surge offers a competitive compensation and total rewards package.
  • Comprehensive Benefits
  • Remote-first work environment
  • PTO - Including holidays + your birthday
  • 401(k) with 4% Match, immediately vested
  • Growth and development opportunities
  • Bonuses + more, * 401(k)
  • 401(k) matching
  • Dental insurance
  • Health insurance
  • Paid time off
  • Retirement plan
  • Vision insurance

Application Question(s):

  • As part of our hiring process, we conduct background checks on all potential candidates. Are you willing to undergo a background check if you are selected for further consideration?
  • What are your salary expectations for the role?
  • Due to the nature of the role, this position requires all employees to have full U.S. Citizenship at the time of hire. Are you able to meet that requirement?
  • Would you be willing to obtain a clearance?
  • Do you have an active clearance?

About the company

Data Surge is disrupting the services industry with cutting-edge technology that brings together the very best consultants from the big data, machine learning, and software modernization space.

At Data Surge, we value collaboration, curiosity, and being customer centric. Our team thrives on collective efforts, innovation, and a commitment to wellness, ensuring employees are empowered to thrive professionally and personally.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on indeed.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:09 min

Balancing data science skillings alongside systems engineering rigor

Nico Schmidt · LIVE

2:37 min

Comparing traditional SQL tables versus NoSQL non-tabular databases

Stanimira Vlaeva · JS Congress

2:18 min

Scaling MySQL databases for massive user growth

Johannes Nicolai Johannes Nicolai +1 · LIVE

2:07 min

Inspecting default bridge architectures and custom Docker networks

Oliver Seitz Oliver Seitz · WWC 2025

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

3:16 min

Terminology differences between relational and NoSQL databases

Tim Faulkes · LIVE

Videos

See all

Related articles

See all