Lead Data Engineer

Insight Global
Upper Dublin Township, PA, United States
14 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Compensation
$135,000.0 - $145,000.0
Working hours
Regular working hours

Tech stack

Microsoft Azure Big Data Continuous Integration Data Architecture Information Engineering Python (Programming Language) Enterprise Data Management Microsoft Fabric Pyspark Information Technology Epic Caboodle Data Pipelines

Job description

Day-to-Day:

  • Serve as a senior technical lead within the Data Engineering organization
  • Oversee data intake processes and determine how new data sources are integrated into the enterprise environment
  • Design, build, and maintain Microsoft Fabric Lakehouse solutions
  • Develop and support PySpark-based data engineering solutions
  • Lead operational support and ongoing optimization of enterprise data platforms
  • Build and maintain CI/CD pipelines and deployment processes
  • Partner with business and technical teams to develop scalable data architectures
  • Support the full data lifecycle from ingestion through consumption by BI teams and external data consumers
  • Manage multiple concurrent initiatives while ensuring platform performance and reliability
  • Provide technical guidance and leadership across the engineering organization
  • Act as a hands-on contributor while helping drive team success and operational excellence, Insight Global is seeking a Lead Data Engineer for a leading healthcare client in the Philadelphia area. This individual will serve as a lead member of the Data Engineering organization responsible for designing, building, and supporting enterprise data platforms within Microsoft Fabric and Azure environments. They will be one of three leads on the team and will balance leadership and hands-on development. The ideal candidate brings experience with PySpark, Fabric Lakehouse development, CI/CD pipeline support, and modern cloud-based data architectures. This role goes beyond traditional development work, it requires someone who can take ownership of data initiatives from intake through delivery, building the underlying infrastructure that enables data to move through its lifecycle and ultimately be consumed by BI teams, external partners, and enterprise stakeholders. The team is seeking a highly proactive, hands-on leader who can independently drive projects forward, support operations, mentor team members, and serve as a trusted technical resource. Candidates must possess strong communication skills, leadership potential, and the ability to step in immediately to support team operations.

Requirements

Must-Haves:

  • 15+ years of experience in Data Engineering and a combination of technical leadership and hands-on engineering experience
  • Bachelor’s degree in Computer Science, Information Technology, or related field
  • Some hands-on experience with Microsoft Fabric
  • Strong Python experience, specifically PySpark for Fabric Lakehouse development
  • Experience supporting and maintaining CI/CD pipelines
  • Hands-on experience with Azure technologies, Azure DevOps, and enterprise data platforms
  • Experience designing, building, and supporting large-scale data infrastructure and data pipelines

Nice to Have Skills & Experience

Plusses:

  • Healthcare or payer or academia industry experience
  • Epic experience or certifications (Clarity, Caboodle, Cogito)

Benefits & conditions

Benefit packages for this role will start on the 1st day of employment and include medical, dental, and vision insurance, as well as HSA, FSA, and DCFSA account options, and 401k retirement account access with employer matching. Employees in this role are also entitled to paid sick leave and/or other paid time off as provided by applicable law.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on jobs.insightglobal.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:28 min

Defining big data and machine learning fundamentals

Ayon Roy · LIVE

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

4:32 min

Harnessing Spark with Python using PySpark and Py4J

Ayon Roy · LIVE

2:10 min

Why organizations combine big data and machine learning

Ayon Roy · LIVE

3:37 min

Scaling machine learning pipelines from prototypes to petabytes

Julian Joseph · LIVE

3:09 min

Balancing data science skillings alongside systems engineering rigor

Nico Schmidt · LIVE

Videos

See all

Related articles

See all