> Markdown version of [/jobs/ext/2705430-data-engineer-data-platform-spark-kafka-flink-scala-java](https://www.wearedevelopers.com/jobs/ext/2705430-data-engineer-data-platform-spark-kafka-flink-scala-java). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Data Engineer - Data Platform (Spark/Kafka/Flink/Scala/Java) - **Company:** VDart, Inc. - **Location:** Atlanta, GA, United States - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** Java (Programming Language), Airflow, Microsoft Azure, Batch Processing, Big Data, Cloud Computing, Cloud Engineering, Databases, Continuous Delivery, Continuous Integration, Information Engineering, Data Infrastructure, Data Integration, Extract Transform Load (ETL), Data Transformation, Data Systems, Relational Databases, File Systems, Distributed Data Store, Distributed Systems, Github, PostgreSQL, Performance Tuning, Cloud Services, Software Engineering, Data Streaming, Freeform SQL, Cloud Platform System, Database Optimization, Apache Spark, Software Application Programming, Gitlab, Event Driven Architecture, Containerization, Kubernetes, Infrastructure Automation Frameworks, Information Technology, Apache Flink, Deployment Automation, Cassandra, Apache Kafka, Data Management, Stream Processing, Stream Analytics, Software Version Control, Data Pipelines, Devsecops - **Published:** September 4, 2026 - **Apply:** https://www.careerjet.com/jobad/usb4dcf250cf1f0460acfc6bbea255b3f1 ## About the Role Basic Qualifications: (What are the skills required for this job with minimum years of experience on each) * Minimum 8+ years of overall Software Engineering or Data Engineering experience. * Minimum 5+ years of hands-on experience with Apache Spark for large-scale batch and streaming data processing. * Minimum 5+ years of hands-on experience with Apache Kafka including event-driven architectures and real-time data streaming solutions. * Minimum 3+ years of hands-on experience with Apache Flink for stream processing and real-time analytics workloads. * Minimum 5+ years of experience developing applications using Java and/or Scala. * Minimum 5+ years of experience writing complex SQL queries and optimizing database performance. * Minimum 3+ years of experience with Kubernetes and containerized application deployment. * Minimum 3+ years of experience designing and implementing solutions on Microsoft Azure Cloud. * Minimum 3+ years of experience building and supporting distributed data platforms using Cassandra, YugabyteDB, PostgreSQL, or similar databases. * Minimum 3+ years of experience building event-driven architectures and streaming applications. * Minimum 2+ years of experience implementing CI/CD pipelines, source control, and deployment automation using GitLab, GitHub, or similar tools. * Minimum 2+ years of experience with cloud-native deployment patterns, containerization, and platform automation. * Demonstrated experience in performance tuning, troubleshooting, and supporting mission-critical data platforms. Nice to Have (But Not a Must) * Experience with Apache Airflow orchestration. * Experience with metadata-driven data platforms and self-service ingestion frameworks. * Experience with enterprise cloud migration and modernization programs. * Experience with platform engineering and internal developer platforms. * Experience leading technical initiatives, mentoring engineers, or serving as a technical lead. * Knowledge of DevSecOps, infrastructure as code, and observability frameworks. Travel: Minimal travel required. Travel may be necessary based on project and stakeholder requirements. Degree: Bachelor's degree in Computer Science, Information Technology, Engineering, or equivalent work experience. ## Description * Join a team building next-generation, cloud-native data platforms powering real-time and batch data movement at enterprise scale. * You will design and develop streaming and batch frameworks, Kafka/Flink/Spark-based processing engines, and cloud-native solutions on Azure and Kubernetes. * Ideal candidates have strong distributed systems experience and a passion for platform engineering, automation, and data modernization. Day to Day Job Duties: (What this person will do on a daily/weekly basis) * Design, develop, and support scalable real-time and batch data pipelines using Apache Spark, Apache Flink, Apache Kafka, and Airflow. * Build and enhance metadata-driven self-service data integration platforms and reusable connectors. * Develop and maintain source and target connectors for relational databases, file systems, Kafka, Cassandra, YugabyteDB, and other enterprise data stores. * Design, deploy, and operate Kubernetes-based streaming and batch processing platforms. * Lead cloud migration initiatives from on-premise environments to Microsoft Azure. * Drive performance tuning, scalability optimization, reliability improvements, and operational excellence across large-scale data workloads. * Develop cloud-native solutions supporting enterprise data movement and processing. * Collaborate with product owners, architects, cloud engineering teams, and business stakeholders to deliver enterprise-scale data solutions. * Contribute to platform modernization, automation, CI/CD implementation, and engineering best practices. * Support troubleshooting, production stability, and continuous improvement initiatives for critical data platforms. ## Related Videos - [From Messy Queries to Scalable Systems - How Data Engineering actually works](https://www.wearedevelopers.com/videos/100203-from-messy-queries-to-scalable-systems-how-data-engineering-actually-works) - [WeAreDevelopers LIVE - Modern DevOps for IoT Devices and More](https://www.wearedevelopers.com/videos/1805-wearedevelopers-live-modern-devops-for-iot-devices-and-more) - [Innovating Developer Tools with AI: Insights from GitHub Next](https://www.wearedevelopers.com/videos/1268-innovating-developer-tools-with-ai-insights-from-github-next) - [Enjoying SQL data pipelines with dbt](https://www.wearedevelopers.com/videos/823-enjoying-sql-data-pipelines-with-dbt) - [AI Model Management Life Circles: ML Ops For Generative AI Models From Research to Deployment](https://www.wearedevelopers.com/videos/1152-ai-model-management-life-circles-ml-ops-for-generative-ai-models-from-research-to-deployment) - [Enabling automated 1-click customer deployments with built-in quality and security](https://www.wearedevelopers.com/videos/83-enabling-automated-1-click-customer-deployments-with-built-in-quality-and-security) ## Related Articles - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [Top-Paying Tech Jobs (with Salaries)](https://www.wearedevelopers.com/magazine/372-top-paying-tech-jobs-with-salaries) - [Top Big Data Technologies That You Need to Know](https://www.wearedevelopers.com/magazine/108-top-big-data-technologies-that-you-need-to-know) - [The Most Popular IT Jobs on the Market](https://www.wearedevelopers.com/magazine/376-the-most-popular-it-jobs-on-the-market) - [The 12 Best Jobs for Software Engineers](https://www.wearedevelopers.com/magazine/401-the-12-best-jobs-for-software-engineers)