Data Engineer

Mizuho Financial Group, Inc.
Woodbridge Township, NJ, United States
6 days ago

Role details

Contract type
Internship / Graduate position
Employment type
Full-time (> 32 hours)
Experience level
Starter
Experience required
0 years minimum
Compensation
$88,000.0 - $105,000.0
Working hours
Regular working hours

Tech stack

Abstraction Layers Artificial Intelligence Amazon Web Services Microsoft Azure Big Data Cloud Computing Code Review Information Systems Computer Programming Databases Data Architecture Data Deduplication
+28 more
Information Engineering Extract Transform Load (ETL) Data Systems Data Visualization Relational Databases Database Queries Electronic Data Interchange (EDI) Python (Programming Language) Object-Oriented Software Development Raw Data Standard Sql Software Engineering SQL Databases SQL Server Integration Services Cloud Platform System Data Ingestion Azure Data Factory Apache Spark Git Data Lakes Pyspark Core Data Information Technology Apache Kafka Data Management Software Version Control Data Pipelines Databricks

Job description

The IT Data team is responsible for design and development of Data products for the entire firm. The team is embarking on an ambitious new project/implementation. “DEAL”. It is the abbreviation for “Data Exchange and Abstraction Layer”. It’s the next generation, Data Mesh based platform implemented in the Mizuho Azure Cloud on Databricks. Data Mesh is a decentralized data architecture where data is owned and managed by the domain-specific teams that produce the data i.e. Banking, Finance etc., and curate it for downstream consumption. It emphasizes domain-oriented ownership, treating data as a product, providing a self-serve data platform, and using limited federated computational governance from the Data Architecture and Data Management Office.

In this role you will be responsible for development of data solutions for the enterprise using innovative and cutting- edge technologies like AI tools / models. The solutions and software developed will be used for reporting and analytics by the entire firm globally. The data solutions developed will have to be accurate, timely and highly scalable. In this role you will be managing large data-sets with complex interdependencies. This is a hands-on software development role. You will be collaborating with teams firmwide to develop solutions.

You’ll support the development and maintenance of data pipelines on the Databricks Lakehouse platform using the medallion architecture (Bronze/Silver/Gold). You will work under the guidance of senior engineers to ingest, transform, and validate data, growing your skills across the modern data stack., * Assist in building and maintaining ingestion pipelines that land raw data into the Bronze layer.

  • Support Silver layer transformations under guidance: cleansing, deduplication, and schema enforcement.
  • Write SQL and PySpark for defined transformation tasks.
  • Run and monitor scheduled jobs; help investigate and resolve pipeline failures.
  • Document pipeline logic, transformations, and fixes.
  • Participate in code reviews as a reviewer-in-training and incorporate feedback on your own work.
  • Learn team standards for version control, testing, and deployment.

Requirements

  • 0-2 years of experience in data engineering, analytics, or a related technical role (internships and academic projects count).
  • Foundational SQL skills (joins, aggregations, filtering).
  • Python proficiency with basic OOPS knowledge .
  • Understanding of core data concepts (tables, schemas, relational data).
  • Willingness to learn Databricks, Spark, and cloud technologies.
  • Familiarity with Git or a demonstrated ability to learn version control quickly.

Preferred / Nice-to-Have

  • Exposure to Databricks, Apache Spark, or PySpark (coursework or hands-on).
  • Awareness of the medallion architecture and Delta Lake basics.
  • Experience with any cloud platform (Azure, AWS, or GCP).
  • Relevant coursework, bootcamp, or a Databricks certification (e.g., Data Engineer Associate).
  • Any experience with data visualization or BI tools., * Eagerness to learn and take feedback.
  • Attention to detail and care for data accuracy.
  • Basic problem-solving and logical thinking.
  • Communicate issues across teams, * Bachelor’s degree in Computer Science, Engineering, Information Systems, or related field. Master’s degree preferred.
  • Proven experience in data engineering, software development, or related roles.
  • Proficiency in programming languages commonly used in data engineering (e.g., Python, Scala, etc.).
  • Strong knowledge of database systems, data modeling techniques, and SQL proficiency.
  • Proficiency with ETL tools commonly used in data engineering (e.g., SSIS, Databricks, Azure Data Factory).
  • Experience with big data technologies and frameworks (e.g., Spark, Kafka, etc.).
  • Familiarity with cloud platforms and services (e.g., Azure).
  • Excellent problem-solving skills and attention to detail.
  • Effective communication and collaboration skills in a team-oriented environment.
  • Ability to adapt to evolving technologies and business requirements

Benefits & conditions

The expected base salary ranges from $88,000 - $105,000. Salary offers are based on a wide range of factors including relevant skills, training, experience, education, and, where applicable, certifications and licenses obtained. Market and organizational factors are also considered. In addition to salary and a generous employee benefits package, including Medical, Dental and 401K plans, successful candidates are also eligible to receive a discretionary bonus.

About the company

Mizuho Financial Group, Inc. is the 15th largest bank in the world as measured by total assets of ~$2 trillion. Mizuho’s 60,000 employees worldwide offer comprehensive financial services to clients in 35 countries and 800 offices throughout the Americas, EMEA and Asia. Mizuho Americas is a leading provider of corporate and investment banking services to clients in the US, Canada, and Latin America. Through its acquisition of Greenhill , Mizuho provides M&A, restructuring and private capital advisory capabilities across Americas, Europe and Asia. Mizuho Americas employs approximately 3,500 professionals, and its capabilities span corporate and investment banking, capital markets, equity and fixed income sales & trading, derivatives, FX, custody and research. Visit www.mizuhoamericas.com.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on mizuho.wd1.myworkdayjobs.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

6:21 min

Investigating push inefficiencies with upstream Git experts

Jonathan Creamer · Coffee With Developers

4:32 min

Harnessing Spark with Python using PySpark and Py4J

Ayon Roy · LIVE

4:09 min

Challenges of interpreting raw data with language models

Clemens Vasters Clemens Vasters · WWC 2025

3:09 min

Balancing data science skillings alongside systems engineering rigor

Nico Schmidt · LIVE

56 sec

Favorite git commands and the importance of patch commits

Eileen Uchitelle Eileen Uchitelle +1 · Coffee With Developers

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

Videos

See all

Related articles

See all