Data Engineer - Data Platform (Spark/Kafka/Flink/Scala/Java) - Onsite Hybrid

NTT DATA Corporation
Cupertino, CA, United States
3 days ago
Apply on careers-inc.nttdata.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Experienced
Experience required
2 years minimum
Compensation
$96,492.0 - $144,738.0
Working hours
Regular working hours

Tech stack

Java (Programming Language) Airflow Microsoft Azure Batch Processing Big Data Cloud Computing Cloud Engineering Databases Continuous Delivery Continuous Integration Information Engineering Data Infrastructure
+36 more
Data Integration Extract Transform Load (ETL) Data Transformation Data Systems Relational Databases File Systems Distributed Data Store Distributed Systems Github PostgreSQL Metadata Performance Tuning Cloud Services Software Engineering Data Streaming Freeform SQL Cloud Platform System Database Optimization Apache Spark Software Application Programming Gitlab Event Driven Architecture Containerization Kubernetes Infrastructure Automation Frameworks Information Technology Apache Flink Deployment Automation Cassandra Apache Kafka Data Management Stream Processing Stream Analytics Software Version Control Data Pipelines Devsecops

Job description

  • Design, develop, and support scalable real-time and batch data pipelines using Apache Spark, Apache Flink, Apache Kafka, and Airflow.
  • Build and enhance metadata-driven self-service
  • data integration platforms and reusable connectors.
  • Develop and maintain source and target connectors for relational databases, file systems, Kafka, Cassandra, YugabyteDB, and other enterprise data stores.
  • Design, deploy, and operate Kubernetes-based streaming and batch processing platforms.
  • Lead cloud migration initiatives from on-premise environments to Microsoft Azure.
  • Drive performance tuning, scalability optimization, reliability improvements, and operational excellence across large-scale data workloads.
  • Develop cloud-native solutions supporting enterprise data movement and processing.
  • Collaborate with product owners, architects, cloud engineering teams, and business stakeholders to deliver enterprise-scale data solutions.
  • Contribute to platform modernization, automation, CI/CD implementation, and engineering best practices.
  • Support troubleshooting, production stability, and continuous improvement initiatives for critical data platforms.

Requirements

NTT DATA strives to hire exceptional, innovative and passionate individuals who want to grow with us. If you want to be part of an inclusive, adaptable, and forward-thinking organization, apply now.

We are currently seeking a Data Engineer - Data Platform (Spark/Kafka/Flink/Scala/Java) - Onsite Hybrid to join our team in Cupertino, California (US-CA), United States (US).

Join a team building next-generation, cloud-native data platforms powering real-time and batch data movement at enterprise scale. You will design and develop streaming and batch frameworks, Kafka/Flink/Spark-based processing engines, and cloud-native solutions on Azure and Kubernetes. Ideal candidates have strong distributed systems experience and a passion for platform engineering, automation, and data modernization., * Minimum 8+ years of overall Software Engineering or Data Engineering experience.

  • Minimum 5+ years of hands-on experience with Apache Spark for large-scale batch and streaming data processing.
  • Minimum 5+ years of hands-on experience with Apache Kafka including event-driven architectures and real-time data streaming solutions.
  • Minimum 3+ years of hands-on experience with Apache Flink for stream processing and real-time analytics workloads.
  • Minimum 5+ years of experience developing applications using Java and/or Scala.
  • Minimum 5+ years of experience writing complex SQL queries and optimizing database performance.
  • Minimum 3+ years of experience with Kubernetes and containerized application deployment.
  • Minimum 3+ years of experience designing and implementing solutions on Microsoft Azure Cloud.
  • Minimum 3+ years of experience building and supporting distributed data platforms using Cassandra, YugabyteDB, PostgreSQL, or similar databases.
  • Minimum 3+ years of experience building event-driven architectures and streaming applications.
  • Minimum 2+ years of experience implementing CI/CD pipelines, source control, and deployment automation using GitLab, GitHub, or similar tools.
  • Minimum 2+ years of experience with cloud-native deployment patterns, containerization, and platform automation.

Travel:

Minimal travel required. Travel may be necessary based on project and stakeholder requirements.

Degree:

Bachelor’s degree in Computer Science, Information Technology, Engineering, or equivalent work experience.

Preferred Skills:

  • Demonstrated experience in performance tuning, troubleshooting, and supporting mission-critical data platforms.
  • Experience with Apache Airflow orchestration.
  • Experience with metadata-driven data platforms and self-service ingestion frameworks.
  • Experience with enterprise cloud migration and modernization programs.
  • Experience with platform engineering and internal developer platforms.
  • Experience leading technical initiatives, mentoring engineers, or serving as a technical lead.
  • Knowledge of DevSecOps, infrastructure as code, and observability frameworks.

Benefits & conditions

NTT DATA provides a reasonable range of compensation for U.S.-based positions. The starting pay range for this role is $96,492 - $144,738 per annum. Actual compensation will depend on a number of factors, including the candidate’s relevant experience, technical skills, and other qualifications. This position may also be eligible for incentive compensation based on individual and/or company performance.

This position is eligible for company benefits including medical, dental, and vision insurance with an employer contribution, flexible spending or health savings account, life and AD&D insurance, short and long term disability coverage, paid time off, employee assistance, participation in a 401k program with company match, and additional voluntary or legally-required benefits.

About the company

NTT DATA is a $30 billion business and technology services leader, serving 75% of the Fortune Global 100. We are committed to accelerating client success and positively impacting society through responsible innovation. We are one of the world’s leading AI and digital infrastructure providers, with unmatched capabilities in enterprise-scale AI, cloud, security, connectivity, data centers and application services. our consulting and Industry solutions help organizations and society move confidently and sustainably into the digital future. As a Global Top Employer, we have experts in more than 50 countries. We also offer clients access to a robust ecosystem of innovation centers as well as established and start-up partners. NTT DATA is a part of NTT Group, which invests over $3 billion each year in R&D.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on careers-inc.nttdata.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:15 min

Empowering domain teams with an open data platform

Sandhya Menon Sandhya Menon · World Congress 2026 Europe

6:36 min

Funding open source through GitHub Accelerator and Sponsors

Stormy Peters · World Congress 2023

6:14 min

Structuring CI/CD pipelines with integrated security and quality checks

Christoph Ruggenthaler · LIVE

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

3:05 min

Audience questions on AI agents and pipeline vectorization

Joy Joy · World Congress 2024

4:54 min

Implementing geographic salary tiers for compensation equity and fairness

Rudi Bauer Rudi Bauer +1 · Cappuccino with HR

Videos

See all

Related articles

See all