Java Spark Engineer

InfiCare Inc
Berkeley Heights, NJ, United States
about 1 month ago
Apply on www.dice.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
5 years minimum
Working hours
Shift work
Job source

Tech stack

Java (Programming Language) Cloud Computing Continuous Integration Data Governance Distributed Systems Memory Management Fault Tolerance Performance Tuning Standard Sql Workflow Management Systems Parquet Apache Yarn
+12 more
Apache Spark Containerization Data Lakes Kubernetes Infrastructure Automation Frameworks Information Technology Apache Flink Avro Apache Kafka Data Management Stream Processing Data Pipelines

Job description

  • Architect and build scalable, fault-tolerant data pipelines using Apache Spark (Java)
  • Drive performance tuning: partitioning strategy, memory management, shuffle/skew optimization
  • Mentor mid-level and junior engineers; act as technical escalation point
  • Partner with product, analytics, and platform teams to translate requirements into scalable systems
  • Own production reliability - incident response and root-cause analysis for pipeline failures
  • Contribute to capacity planning and cost optimization for cluster infrastructure, Berkeley Heights, NJ - fully onsite, 5 days per week. Candidates must be flexible to support weekend operations when needed.

Requirements

Experience: 7+ years Java development; 5+ years Apache Spark in production, * 7+ years professional Java development experience

  • 5+ years hands-on Apache Spark in production environments
  • Expert-level distributed systems knowledge: fault tolerance, data locality, shuffle mechanics, resource management
  • Proven track record designing systems at terabyte+ scale
  • Strong SQL and deep familiarity with columnar storage formats: Parquet, ORC, Avro, Delta Lake/Iceberg
  • Experience with cluster managers: YARN, Kubernetes, cloud-managed Spark
  • Proficiency with Apache Kafka
  • Strong grasp of CI/CD, containerization, and infrastructure-as-code practices

Preferred Skills

  • Experience with Apache Flink or other stream-processing frameworks
  • Familiarity with data governance, lineage, and quality frameworks
  • Experience with workflow orchestration at scale
  • Background in system design for multi-tenant or multi-region data platforms, 5 openings available. Contract engagement. Bachelor’s or Master’s degree in Computer Science, Engineering, or related field required.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.dice.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:02 min

Audience Q&A on data formats and engine tradeoffs

Matthias Niehoff Matthias Niehoff · World Congress 2026 Europe

2:28 min

Understanding Kubernetes architecture and core cluster components

Marc Nimmerrichter · World Congress 2022

2:50 min

How Parquet metadata enables efficient data reading

Matthias Niehoff Matthias Niehoff · World Congress 2026 Europe

3:55 min

Infrastructure challenges when combining Kafka with Apache Flink

Bobur Umurzokov · LIVE

1:52 min

Customizing block storage tiers and formats

Ricardo Sueiras Sueiras · LIVE

4:04 min

Overview of Kubernetes operators and custom resource definitions

Philipp Krenn · World Congress 2022

Videos

See all

Related articles

See all