Data Engineer - Enterprise Data Frameworks (Java Spark Focus)

Citizens Financial Group
United States
19 days ago

Role details

Contract type
Permanent contract
Employment type
Part-time (≤ 32 hours)
Experience level
Expert
Experience required
7 years minimum
Compensation
$105,000.0 - $158,000.0
Working hours
Regular working hours

Tech stack

Java (Programming Language) Adobe InDesign Application Programming Interfaces (APIs) Agile Methodology Amazon Web Services Amazon S3 Code Review Continuous Integration Data as a Services Data Architecture Data Governance Software Debugging
+29 more
Distributed Computing Environment Distributed Systems Event-Driven Programming IntelliJ IDEA Java Virtual Machine (JVM) Python (Programming Language) Metadata Object-Oriented Software Development Performance Tuning Web Application Frameworks Parquet Data Processing Scripting Cloud Platform System Data Ingestion Apache Spark Spring-boot Backend Git Containerization Data Lakes Kubernetes Information Technology Apache Kafka Data Management Front End Software Development Api Gateway Data Pipelines Docker

Job description

The Senior Data Engineer will design and implement data processing pipelines using Java and Spark, optimize performance, and mentor junior engineers while ensuring compliance with data governance standards., The Enterprise Data Frameworks team is seeking a Senior Java focused software engineer who builds and maintains large scale data processing systems using Java, Apache Spark, and Kafka. This role is intended for experienced backend engineers with strong core Java fundamentals who apply traditional software engineering practices to data intensive platforms and distributed systems.

The ideal candidate has hands on experience developing production grade Java applications using modern frameworks and IntelliJ based development workflows, paired with practical experience building Spark based processing pipelines and Kafka driven data ingestion services. In this role, you will contribute to the core components of enterprise data frameworks, working closely with senior engineers and architects while remaining deeply involved in design and implementation., * Design, build, and maintain Java based data processing pipelines using Apache Spark for batch and streaming workloads

  • Develop and support Java based backend services and APIs that orchestrate data workflows and framework components
  • Apply solid object oriented design principles to distributed data processing systems
  • Optimize Spark applications for performance, reliability, and cost efficiency across cloud and on premises environments
  • Collaborate with architects and senior engineers on technical design and framework evolution
  • Partner with frontend and platform teams to integrate backend data services where applicable
  • Translate business and technical requirements into scalable, well structured engineering solutions
  • Participate in code reviews, contribute to shared standards, and mentor junior engineers
  • Ensure solutions align with data governance, security, and change management standards
  • Participate in Agile ceremonies and support continuous improvement of engineering practices

Requirements

This role requires strong Java engineering skills, the ability to debug and optimize complex Spark applications, and comfort operating in regulated environments where stability, data quality, and reliability are critical., * 7+ years of experience as a software engineer or data engineer with strong emphasis on Java backend development

  • Strong hands on proficiency in Java, including object oriented design, debugging, and performance optimization
  • Experience building Spark applications in Java for large scale data ingestion and transformation
  • Practical experience with Apache Kafka and event driven data architectures
  • Experience developing Java based services using frameworks such as Spring Boot
  • Solid understanding of distributed systems concepts and data processing architectures
  • Familiarity with data lake technologies, columnar storage formats such as Parquet or Iceberg, and metadata driven frameworks
  • Experience using Git based workflows, CI CD pipelines, and modern development practices
  • Exposure to cloud environments, with preference for AWS based data platforms such as S3, EMR, Lambda, or API Gateway
  • Experience working in regulated or enterprise environments with strong data governance requirements

Preferred Experience

  • Experience collaborating with UI or platform teams to integrate backend data services
  • Experience with containerization and orchestration tools such as Docker or Kubernetes
  • Exposure to additional JVM or scripting languages such as Scala or Python in a data context
  • Experience with developer productivity or automation tooling

Education

  • Bachelor’s or Master’s degree in Computer Science, Engineering, or a related technical field

Benefits & conditions

Hours per Week 40 Location Hybrid, 4 days on site from a Citizens corporate office, 1 day remote Schedule Monday through Friday

Pay Transparency

The salary range for this position is $105,000-158,000 per year, plus an opportunity to earn an annual discretionary bonus. Actual pay is based on various factors including but not limited to the work location, and relevant skills and experience.

We offer competitive pay, comprehensive medical, dental and vision coverage, retirement benefits, maternity/paternity leave, flexible work arrangements, education reimbursement, wellness programs and more. Note, Citizens’ paid time off policy exceeds the mandatory, paid sick or paid time-away policy of very local and state jurisdiction in the United States. For an overview of our benefits, visit https://jobs.citizensbank.com/benefits.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on hcgn.fa.us2.oraclecloud.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

6:21 min

Investigating push inefficiencies with upstream Git experts

Jonathan Creamer · Coffee With Developers

2:07 min

Inspecting default bridge architectures and custom Docker networks

Oliver Seitz Oliver Seitz · WWC 2025

2:50 min

How Parquet metadata enables efficient data reading

Matthias Niehoff Matthias Niehoff · WWC Europe 2026

3:09 min

Balancing data science skillings alongside systems engineering rigor

Nico Schmidt · LIVE

56 sec

Favorite git commands and the importance of patch commits

Eileen Uchitelle Eileen Uchitelle +1 · Coffee With Developers

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

Videos

See all

Related articles

See all