> Markdown version of [/jobs/ext/2696521-data-engineer-data-platform-spark-kafka-flink-scala-java-onsite-hybrid](https://www.wearedevelopers.com/jobs/ext/2696521-data-engineer-data-platform-spark-kafka-flink-scala-java-onsite-hybrid). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Data Engineer - Data Platform (Spark/Kafka/Flink/Scala/Java) - Onsite Hybrid - **Company:** NTT DATA Corporation - **Location:** Cupertino, CA, United States - **Experience:** Experienced - **Salary:** $96,492.0 - $144,738.0 - **Contract:** Permanent contract - **Skills:** Java (Programming Language), Airflow, Microsoft Azure, Batch Processing, Big Data, Cloud Computing, Cloud Engineering, Databases, Continuous Delivery, Continuous Integration, Information Engineering, Data Infrastructure, Data Integration, Extract Transform Load (ETL), Data Transformation, Data Systems, Relational Databases, File Systems, Distributed Data Store, Distributed Systems, Github, PostgreSQL, Metadata, Performance Tuning, Cloud Services, Software Engineering, Data Streaming, Freeform SQL, Cloud Platform System, Database Optimization, Apache Spark, Software Application Programming, Gitlab, Event Driven Architecture, Containerization, Kubernetes, Infrastructure Automation Frameworks, Information Technology, Apache Flink, Deployment Automation, Cassandra, Apache Kafka, Data Management, Stream Processing, Stream Analytics, Software Version Control, Data Pipelines, Devsecops - **Published:** September 3, 2026 - **Apply:** https://careers-inc.nttdata.com/talentcommunity/apply/1425944800/?locale=en_US ## About the Role NTT DATA strives to hire exceptional, innovative and passionate individuals who want to grow with us. If you want to be part of an inclusive, adaptable, and forward-thinking organization, apply now. We are currently seeking a Data Engineer - Data Platform (Spark/Kafka/Flink/Scala/Java) - Onsite Hybrid to join our team in Cupertino, California (US-CA), United States (US). Join a team building next-generation, cloud-native data platforms powering real-time and batch data movement at enterprise scale. You will design and develop streaming and batch frameworks, Kafka/Flink/Spark-based processing engines, and cloud-native solutions on Azure and Kubernetes. Ideal candidates have strong distributed systems experience and a passion for platform engineering, automation, and data modernization., * Minimum 8+ years of overall Software Engineering or Data Engineering experience. * Minimum 5+ years of hands-on experience with Apache Spark for large-scale batch and streaming data processing. * Minimum 5+ years of hands-on experience with Apache Kafka including event-driven architectures and real-time data streaming solutions. * Minimum 3+ years of hands-on experience with Apache Flink for stream processing and real-time analytics workloads. * Minimum 5+ years of experience developing applications using Java and/or Scala. * Minimum 5+ years of experience writing complex SQL queries and optimizing database performance. * Minimum 3+ years of experience with Kubernetes and containerized application deployment. * Minimum 3+ years of experience designing and implementing solutions on Microsoft Azure Cloud. * Minimum 3+ years of experience building and supporting distributed data platforms using Cassandra, YugabyteDB, PostgreSQL, or similar databases. * Minimum 3+ years of experience building event-driven architectures and streaming applications. * Minimum 2+ years of experience implementing CI/CD pipelines, source control, and deployment automation using GitLab, GitHub, or similar tools. * Minimum 2+ years of experience with cloud-native deployment patterns, containerization, and platform automation. Travel: Minimal travel required. Travel may be necessary based on project and stakeholder requirements. Degree: Bachelor's degree in Computer Science, Information Technology, Engineering, or equivalent work experience. Preferred Skills: * Demonstrated experience in performance tuning, troubleshooting, and supporting mission-critical data platforms. * Experience with Apache Airflow orchestration. * Experience with metadata-driven data platforms and self-service ingestion frameworks. * Experience with enterprise cloud migration and modernization programs. * Experience with platform engineering and internal developer platforms. * Experience leading technical initiatives, mentoring engineers, or serving as a technical lead. * Knowledge of DevSecOps, infrastructure as code, and observability frameworks. ## Description * Design, develop, and support scalable real-time and batch data pipelines using Apache Spark, Apache Flink, Apache Kafka, and Airflow. * Build and enhance metadata-driven self-service * data integration platforms and reusable connectors. * Develop and maintain source and target connectors for relational databases, file systems, Kafka, Cassandra, YugabyteDB, and other enterprise data stores. * Design, deploy, and operate Kubernetes-based streaming and batch processing platforms. * Lead cloud migration initiatives from on-premise environments to Microsoft Azure. * Drive performance tuning, scalability optimization, reliability improvements, and operational excellence across large-scale data workloads. * Develop cloud-native solutions supporting enterprise data movement and processing. * Collaborate with product owners, architects, cloud engineering teams, and business stakeholders to deliver enterprise-scale data solutions. * Contribute to platform modernization, automation, CI/CD implementation, and engineering best practices. * Support troubleshooting, production stability, and continuous improvement initiatives for critical data platforms. ## Related Videos - [From Messy Queries to Scalable Systems - How Data Engineering actually works](https://www.wearedevelopers.com/videos/100203-from-messy-queries-to-scalable-systems-how-data-engineering-actually-works) - [WeAreDevelopers LIVE - Modern DevOps for IoT Devices and More](https://www.wearedevelopers.com/videos/1805-wearedevelopers-live-modern-devops-for-iot-devices-and-more) - [Innovating Developer Tools with AI: Insights from GitHub Next](https://www.wearedevelopers.com/videos/1268-innovating-developer-tools-with-ai-insights-from-github-next) - [Enjoying SQL data pipelines with dbt](https://www.wearedevelopers.com/videos/823-enjoying-sql-data-pipelines-with-dbt) - [AI Model Management Life Circles: ML Ops For Generative AI Models From Research to Deployment](https://www.wearedevelopers.com/videos/1152-ai-model-management-life-circles-ml-ops-for-generative-ai-models-from-research-to-deployment) - [Enabling automated 1-click customer deployments with built-in quality and security](https://www.wearedevelopers.com/videos/83-enabling-automated-1-click-customer-deployments-with-built-in-quality-and-security) ## Related Articles - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [Top-Paying Tech Jobs (with Salaries)](https://www.wearedevelopers.com/magazine/372-top-paying-tech-jobs-with-salaries) - [The Most Popular IT Jobs on the Market](https://www.wearedevelopers.com/magazine/376-the-most-popular-it-jobs-on-the-market) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again)