Senior Data Engineer-Databricks

ExlService Holdings, Inc.
United States
about 1 month ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Compensation
$110,000.0 - $150,000.0
Working hours
Regular working hours
Job source

Tech stack

Java (Programming Language) Agile Methodology Amazon Web Services Bash Shell Big Data Cloud Database Information Engineering Extract Transform Load (ETL) Data Warehousing Database Design Dimensional Modeling Apache Hadoop
+27 more
Apache Hive Python (Programming Language) Linked Data Unix Shell Microsoft SQL Server Oracle Databases Oracle (Applications) Cloud Services Azure Data Lake Shell Script Software Engineering SQL Databases Talend Unstructured Data Cloud Platform System Data Ingestion Informatica Powercenter System Availability Apache Spark Data Strategy Vba Programming Language Data Management Dynamic Data Restful APIs Looker Analytics Data Pipelines Databricks

Job description

We are seeking a highly skilled and motivated Senior Data Engineer specializing in Databricks to join our dynamic data team. In this role, you will lead the design, development, and optimization of large-scale data systems, enabling advanced analytics and business intelligence solutions. Your expertise will drive innovative data management strategies across cloud platforms, ensuring seamless integration and high-performance processing of complex datasets. This position offers an exciting opportunity to work with cutting-edge big data technologies and contribute to impactful data-driven decision-making., * Design, develop, and maintain scalable ETL pipelines using Databricks, Spark, Python, and SQL to process vast amounts of structured and unstructured data.

  • Collaborate with cross-functional teams to translate business requirements into robust data models and architectures aligned with best practices in dimensional modeling and data warehousing design.
  • Optimize big data systems leveraging Hadoop, Azure Data Lake, AWS, and other cloud-based platforms to ensure high availability and performance.
  • Implement data management integration strategies utilizing Informatica, Talend, RESTful APIs, and other tools to streamline data ingestion and transformation processes.
  • Develop comprehensive documentation for data workflows, models, and pipelines while ensuring adherence to security standards and compliance regulations.
  • Support model training, analysis activities, and business intelligence reporting using Looker, SQL databases such as Oracle or Microsoft SQL Server, and cloud databases.
  • Participate in Agile development cycles to continuously improve data engineering solutions while maintaining flexibility for evolving project needs.

Requirements

  • Proven experience in software development with a strong background in data engineering within big data environments.
  • Extensive hands-on expertise with Databricks platform, Spark (including Apache Hive), Hadoop ecosystems, and cloud services such as Azure Data Lake or AWS Public Cloud.
  • Proficiency in programming languages including Python, Java, Bash (Unix shell), Shell Scripting, and VBA for automation and analysis tasks.
  • Deep understanding of ETL pipeline development, SQL query management, database design (including dimensional modeling), and data warehousing concepts.
  • Familiarity with business intelligence tools like Looker or similar platforms for visualization and reporting purposes.
  • Knowledge of data modeling techniques for linked data, Oracle databases, SQL databases, cloud databases, and enterprise data warehouse architecture.
  • Experience working with Big Data systems such as Hadoop or Spark for large-scale processing projects.
  • Strong analytical skills combined with excellent communication abilities to translate complex technical concepts into clear insights.
  • Ability to work effectively within an Agile environment while managing multiple priorities simultaneously.

Benefits & conditions

3.23.2 out of 5 stars Remote $110,000 - $150,000 a year - Full-time, Pulled from the full job description

  • Parental leave
  • 401(k)
  • Health insurance
  • Retirement plan
  • 401(k) matching
  • Paid time off
  • Vision insurance, * 401(k)
  • 401(k) matching
  • Dental insurance
  • Health insurance
  • Health savings account
  • Life insurance
  • Paid time off
  • Parental leave
  • Retirement plan
  • Vision insurance

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.indeed.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:28 min

Defining big data and machine learning fundamentals

Ayon Roy Ā· LIVE

2:27 min

Managing traffic and tracking costs with Databricks Unity Catalog

Viktoria Semaan Viktoria Semaan Ā· WWC Europe 2026

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou Ā· Coffee With Developers

1:59 min

Key takeaways and accessing the Databricks developer toolkit

Viktoria Semaan Viktoria Semaan Ā· WWC Europe 2026

2:10 min

Why organizations combine big data and machine learning

Ayon Roy Ā· LIVE

3:09 min

Balancing data science skillings alongside systems engineering rigor

Nico Schmidt Ā· LIVE

Videos

See all

Related articles

See all