Data Solutions Engineer, Assistant Vice President

Citi
Irving, TX, United States
2 days ago
Apply on citi.wd5.myworkdayjobs.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
5 years minimum
Compensation
$107,120.0 - $160,680.0
Working hours
Regular working hours

Tech stack

Java (Programming Language) Amazon Web Services Big Data Cloud Computing Cloud Database Information Engineering Data Infrastructure Data Systems Data Warehousing Distributed Computing Environment Apache Hadoop Apache Hive
+19 more
Python (Programming Language) Machine Learning Scrum Methodology SAS (Software) Software Engineering Working Model 2D Google Cloud Cloud Platform System Snowflake Apache Spark Containerization Pyspark Apache Kafka Data Management Machine Learning Operations Software Coding Data Pipelines Docker Databricks

Job description

  • Build and maintain scalable, high-performance data pipelines that support analytics, reporting, and machine learning capabilities across the organization.

  • Develop Big Data solutions using Apache Spark, Hadoop, and cloud-native services to process and manage large volumes of data reliably.

  • Partners with data scientists and analytics teams to engineer data pipelines that support advanced analytics and machine learning use cases.

  • Collaborate with architects and senior engineers to gather requirements and contribute to the implementation of data platform designs.

  • Apply data quality controls, monitoring, and automation to maintain data governance and auditability standards.

  • Research emerging technologies and share recommendations that could improve platform performance, scalability, or maintainability.

  • Support the development of junior engineers by sharing knowledge and technical best practices within the team.

  • Participate in Agile delivery activities including sprint planning, development, testing, and production support.

Requirements

  • 5+ years of experience in data engineering, software engineering, or a closely related discipline.

  • Practical experience designing and building large-scale data pipelines and distributed data processing solutions in production environments.

  • Hands-on development experience using Python and Apache Spark (PySpark) to process and transform data at scale.

  • Experience working with cloud-based data platforms, with exposure to AWS or Google Cloud Platform (GCP).

  • Working knowledge of data warehousing concepts, data modeling, and distributed processing frameworks including Spark, Hadoop, Kafka, or Hive.

  • Drive implementation, consistent patterns, reusable components, and coding standards for data engineering processes

  • Tune Big data applications on Hadoop and non-Hadoop platforms for optimal performance

  • Be the technical expert and partner with team members on Big Data and Cloud Tech stacks.

  • Familiarity with containerization tools such as Docker and Kubernetes.

  • Clear communication skills with the ability to work effectively across technical and non-technical teams.

Beneficial Skills & Qualifications

  • Experience using Snowflake/Databricks for cloud data warehousing and analytics.

  • Exposure to migrating or modernizing legacy data platforms, including SAS-based workflows.

  • Familiarity with machine learning data pipelines and MLOps concepts.

  • Development experience in Scala or Java alongside Python.

  • AWS or GCP cloud certification(s)., Note: Applicants must be authorized to work in the U.S for this position; Citi will not sponsor applicants for U.S. work authorization for this role. Candidate must be located within commuting distance or be willing to relocate to the area.

Benefits & conditions

At Citi, you’ll work as part of a collaborative engineering team on data challenges that have direct business impact across a global organization. This role offers meaningful technical development, exposure to modern cloud and data technologies, and a structured path to grow your career within Citi Technology.

  • Hybrid working model with 3 days in the office and 2 days working remotely, giving you flexibility alongside regular in-person collaboration.

  • Hands-on experience with modern cloud platforms, Big Data technologies, and analytics tooling used at enterprise scale.

  • Structured learning and development opportunities to support your technical growth and career progression within Citi Technology.

  • Exposure to data modernization projects that directly support business decision-making across a global financial institution.

  • Competitive compensation package and comprehensive benefits program., $107,120.00 - $160,680.00

In addition to salary, Citi’s offerings may also include, for eligible employees, discretionary and formulaic incentive and retention awards. Citi offers competitive employee benefits, including: medical, dental & vision coverage; 401(k); life, accident, and disability insurance; and wellness programs. Citi also offers paid time off packages, including planned time off (vacation), unplanned time off (sick leave), and paid holidays. For additional information regarding Citi employee benefits, please visit citibenefits.com. Available offerings may vary by jurisdiction, job level, and date of hire.

About the company

Working at Citi is far more than just a job. A career with us means joining a team of approximately 219,000 dedicated people from around the globe. At Citi, you’ll have the opportunity to grow your career, give back to your community and make a real impact.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on citi.wd5.myworkdayjobs.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:28 min

Defining big data and machine learning fundamentals

Ayon Roy · LIVE

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

4:32 min

Harnessing Spark with Python using PySpark and Py4J

Ayon Roy · LIVE

2:07 min

Inspecting default bridge architectures and custom Docker networks

Oliver Seitz Oliver Seitz · World Congress 2025

2:10 min

Why organizations combine big data and machine learning

Ayon Roy · LIVE

2:34 min

Capabilities of the Apache Spark processing engine

Ayon Roy · LIVE

Videos

See all

Related articles

See all