Java Spark Engineer
InfiCare Inc
Berkeley Heights, NJ, United States
about 1 month ago
- Discuss this with your agent
- Open in Claude
- Open in ChatGPT
Apply on www.dice.com
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Role details
Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
5 years minimum
Working hours
Shift work
Job source
Tech stack
Java (Programming Language)
Cloud Computing
Continuous Integration
Data Governance
Distributed Systems
Memory Management
Fault Tolerance
Performance Tuning
Standard Sql
Workflow Management Systems
Parquet
Apache Yarn
+12 more
Apache Spark
Containerization
Data Lakes
Kubernetes
Infrastructure Automation Frameworks
Information Technology
Apache Flink
Avro
Apache Kafka
Data Management
Stream Processing
Data Pipelines
Job description
- Architect and build scalable, fault-tolerant data pipelines using Apache Spark (Java)
- Drive performance tuning: partitioning strategy, memory management, shuffle/skew optimization
- Mentor mid-level and junior engineers; act as technical escalation point
- Partner with product, analytics, and platform teams to translate requirements into scalable systems
- Own production reliability - incident response and root-cause analysis for pipeline failures
- Contribute to capacity planning and cost optimization for cluster infrastructure, Berkeley Heights, NJ - fully onsite, 5 days per week. Candidates must be flexible to support weekend operations when needed.
Requirements
Experience: 7+ years Java development; 5+ years Apache Spark in production, * 7+ years professional Java development experience
- 5+ years hands-on Apache Spark in production environments
- Expert-level distributed systems knowledge: fault tolerance, data locality, shuffle mechanics, resource management
- Proven track record designing systems at terabyte+ scale
- Strong SQL and deep familiarity with columnar storage formats: Parquet, ORC, Avro, Delta Lake/Iceberg
- Experience with cluster managers: YARN, Kubernetes, cloud-managed Spark
- Proficiency with Apache Kafka
- Strong grasp of CI/CD, containerization, and infrastructure-as-code practices
Preferred Skills
- Experience with Apache Flink or other stream-processing frameworks
- Familiarity with data governance, lineage, and quality frameworks
- Experience with workflow orchestration at scale
- Background in system design for multi-tenant or multi-region data platforms, 5 openings available. Contract engagement. Bachelor’s or Master’s degree in Computer Science, Engineering, or related field required.
Apply for this position
This job is hosted externally. Click below to view the full posting and apply.
Apply on www.dice.com
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Good distractions
Talks and stories from around this role — technically off-topic, practically not.
Moments
Explore playlistsVideos
See allRelated articles
See all
EM
Eli McGarvie
over 3 years ago
IK
Igor Khokhriakov
How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again
24 days ago
DS
Dhannush Subramani
Top Big Data Technologies That You Need to Know
about 4 years ago
CH
Chris Heilmann
Dev Digest 121 - AI goes offline
over 2 years ago
LM
Luis Minvielle
7 Cloud Computing Trends Coming in 2025 for Developers
over 2 years ago
CH
Chris Heilmann
Dev Digest 120 - Apple and peers
about 2 years ago