Data Engineer - Data Platform (Spark/Kafka/Flink/Scala/Java)

VDart, Inc.
Atlanta, GA, United States
1 day ago
Apply on www.careerjet.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
2 years minimum
Working hours
Regular working hours

Tech stack

Java (Programming Language) Airflow Microsoft Azure Batch Processing Big Data Cloud Computing Cloud Engineering Databases Continuous Delivery Continuous Integration Information Engineering Data Infrastructure
+35 more
Data Integration Extract Transform Load (ETL) Data Transformation Data Systems Relational Databases File Systems Distributed Data Store Distributed Systems Github PostgreSQL Performance Tuning Cloud Services Software Engineering Data Streaming Freeform SQL Cloud Platform System Database Optimization Apache Spark Software Application Programming Gitlab Event Driven Architecture Containerization Kubernetes Infrastructure Automation Frameworks Information Technology Apache Flink Deployment Automation Cassandra Apache Kafka Data Management Stream Processing Stream Analytics Software Version Control Data Pipelines Devsecops

Job description

  • Join a team building next-generation, cloud-native data platforms powering real-time and batch data movement at enterprise scale.
  • You will design and develop streaming and batch frameworks, Kafka/Flink/Spark-based processing engines, and cloud-native solutions on Azure and Kubernetes.
  • Ideal candidates have strong distributed systems experience and a passion for platform engineering, automation, and data modernization.

Day to Day Job Duties: (What this person will do on a daily/weekly basis)

  • Design, develop, and support scalable real-time and batch data pipelines using Apache Spark, Apache Flink, Apache Kafka, and Airflow.
  • Build and enhance metadata-driven self-service data integration platforms and reusable connectors.
  • Develop and maintain source and target connectors for relational databases, file systems, Kafka, Cassandra, YugabyteDB, and other enterprise data stores.
  • Design, deploy, and operate Kubernetes-based streaming and batch processing platforms.
  • Lead cloud migration initiatives from on-premise environments to Microsoft Azure.
  • Drive performance tuning, scalability optimization, reliability improvements, and operational excellence across large-scale data workloads.
  • Develop cloud-native solutions supporting enterprise data movement and processing.
  • Collaborate with product owners, architects, cloud engineering teams, and business stakeholders to deliver enterprise-scale data solutions.
  • Contribute to platform modernization, automation, CI/CD implementation, and engineering best practices.
  • Support troubleshooting, production stability, and continuous improvement initiatives for critical data platforms.

Requirements

Basic Qualifications: (What are the skills required for this job with minimum years of experience on each)

  • Minimum 8+ years of overall Software Engineering or Data Engineering experience.
  • Minimum 5+ years of hands-on experience with Apache Spark for large-scale batch and streaming data processing.
  • Minimum 5+ years of hands-on experience with Apache Kafka including event-driven architectures and real-time data streaming solutions.
  • Minimum 3+ years of hands-on experience with Apache Flink for stream processing and real-time analytics workloads.
  • Minimum 5+ years of experience developing applications using Java and/or Scala.
  • Minimum 5+ years of experience writing complex SQL queries and optimizing database performance.
  • Minimum 3+ years of experience with Kubernetes and containerized application deployment.
  • Minimum 3+ years of experience designing and implementing solutions on Microsoft Azure Cloud.
  • Minimum 3+ years of experience building and supporting distributed data platforms using Cassandra, YugabyteDB, PostgreSQL, or similar databases.
  • Minimum 3+ years of experience building event-driven architectures and streaming applications.
  • Minimum 2+ years of experience implementing CI/CD pipelines, source control, and deployment automation using GitLab, GitHub, or similar tools.
  • Minimum 2+ years of experience with cloud-native deployment patterns, containerization, and platform automation.
  • Demonstrated experience in performance tuning, troubleshooting, and supporting mission-critical data platforms.

Nice to Have (But Not a Must)

  • Experience with Apache Airflow orchestration.
  • Experience with metadata-driven data platforms and self-service ingestion frameworks.
  • Experience with enterprise cloud migration and modernization programs.
  • Experience with platform engineering and internal developer platforms.
  • Experience leading technical initiatives, mentoring engineers, or serving as a technical lead.
  • Knowledge of DevSecOps, infrastructure as code, and observability frameworks.

Travel: Minimal travel required. Travel may be necessary based on project and stakeholder requirements. Degree: Bachelor’s degree in Computer Science, Information Technology, Engineering, or equivalent work experience.

About the company

Cargill is committed to providing food and agricultural solutions to nourish the world in a safe, responsible, and sustainable way. Sitting at the heart of the supply chain, we par…

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.careerjet.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:15 min

Empowering domain teams with an open data platform

Sandhya Menon Sandhya Menon · World Congress 2026 Europe

6:36 min

Funding open source through GitHub Accelerator and Sponsors

Stormy Peters · World Congress 2023

6:14 min

Structuring CI/CD pipelines with integrated security and quality checks

Christoph Ruggenthaler · LIVE

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

3:05 min

Audience questions on AI agents and pipeline vectorization

Joy Joy · World Congress 2024

4:54 min

Implementing geographic salary tiers for compensation equity and fairness

Rudi Bauer Rudi Bauer +1 · Cappuccino with HR

Videos

See all

Related articles

See all