Lead Scala Data Engineer

ESG
United States
4 days ago
Apply on www.dice.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
4 years minimum
Working hours
Regular working hours
Job source

Tech stack

Agile Methodology Amazon Web Services Application Frameworks Cloud Computing Data Infrastructure IBM InfoSphere DataStage Linux Distributed Data Store Drools Apache Hadoop Hadoop Distributed File System Apache Hive
+10 more
Cloudera Scala (Programming Language) Systems Integration Apache Yarn Cloudera Manager Apache Spark Git Atlassian Tools Software Version Control Service Stack

Job description

We are seeking an experienced Lead Scala Data Engineer to serve as the primary technical owner of an enterprise data-processing framework supporting Medicaid Encounter Processing and enterprise data ingestion. \n

\n This is a hands-on technical ownership position responsible for maintaining and enhancing a mission-critical production application built with Scala, Apache Spark, Hive, Drools, and Cloudera Data Platform (CDP). The ideal candidate will have extensive experience developing distributed data-processing applications and supporting them in production. \n

\n Key Responsibilities \n \n

  • Serve as the primary technical owner of the Scala/Spark application framework supporting Medicaid Encounter Processing.\n
  • Maintain and enhance production applications developed using Scala, Spark, Hive, and Drools.\n
  • Own Drools business-rule implementation and ongoing rule maintenance.\n
  • Support enterprise ingestion and processing of provider, member, reference, eligibility, and encounter data.\n
  • Own Spark and Hive batch-processing workflows.\n
  • Troubleshoot production issues, identify root causes, and resolve application defects.\n
  • Implement business, regulatory, and application changes.\n
  • Manage production releases, version control, and deployment coordination.\n
  • Perform Spark performance tuning and optimization.\n
  • Monitor and provide basic operational support for the Cloudera Data Platform (CDP).\n
  • Maintain technical documentation, operational procedures, and knowledge-transfer materials.\n
  • Coordinate with infrastructure, cloud operations, QA, business, and other technical teams.\n

Requirements

  • 7+ years of experience developing enterprise-scale distributed data-processing applications.\n
  • Strong hands-on Scala development experience, preferably 4-6+ years.\n
  • 4-6+ years of Apache Spark and Hive development experience.\n
  • Significant experience developing and maintaining applications on Cloudera Data Platform (CDP) 7.x or equivalent enterprise Hadoop environments.\n
  • Hands-on experience implementing business rules using the Drools Rules Engine, preferably 4-6 years.\n
  • Strong SQL development and query-optimization skills.\n
  • Experience supporting Linux-based production environments.\n
  • Strong experience troubleshooting distributed Spark applications in production.\n
  • Experience with Git and modern version-control practices.\n
  • Demonstrated experience taking technical ownership of production applications, including incidents, defects, enhancements, releases, deployments, and performance issues.\n, * Medicaid or healthcare industry experience.\n
  • Experience with Medicaid Encounter Processing.\n
  • Experience with Cloudera Manager, HDFS, and YARN.\n
  • Experience integrating with IBM DataStage.\n
  • Familiarity with AWS infrastructure supporting Cloudera.\n
  • Experience working in Agile environments using Jira and Confluence.\n, n We are specifically looking for a Scala/Spark Data Engineer with strong Cloudera/CDP experience, rather than a Cloudera Administrator. The successful candidate should be capable of independently owning a production application across the complete Scala + Spark + Hive + Drools + Cloudera technology stack. \n

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.dice.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

52 sec

Running persistent Linux environments directly on Windows

Ben Breard Ben Breard · World Congress 2025

6:21 min

Investigating push inefficiencies with upstream Git experts

Jonathan Creamer · Coffee With Developers

3:09 min

Balancing data science skillings alongside systems engineering rigor

Nico Schmidt · LIVE

3:55 min

Demonstrating .NET installation on Debian and Azure Linux

Silvano Coriani Silvano Coriani · Europe 2026 Virtual

56 sec

Favorite git commands and the importance of patch commits

Eileen Uchitelle Eileen Uchitelle +1 · Coffee With Developers

1:41 min

Visualizing the complex developer journey for JVM ecosystems

Bobur Umurzokov · LIVE

Videos

See all

Related articles

See all