> Markdown version of [/jobs/ext/3532716-data-solutions-engineer-assistant-vice-president](https://www.wearedevelopers.com/jobs/ext/3532716-data-solutions-engineer-assistant-vice-president). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Data Solutions Engineer, Assistant Vice President - **Company:** Citi - **Location:** Irving, TX, United States (Remote available) - **Experience:** Expert - **Salary:** $107,120.0 - $160,680.0 - **Contract:** Permanent contract - **Skills:** Java (Programming Language), Amazon Web Services, Big Data, Cloud Computing, Cloud Database, Information Engineering, Data Infrastructure, Data Systems, Data Warehousing, Distributed Computing Environment, Apache Hadoop, Apache Hive, Python (Programming Language), Machine Learning, Scrum Methodology, SAS (Software), Software Engineering, Working Model 2D, Google Cloud, Cloud Platform System, Snowflake, Apache Spark, Containerization, Pyspark, Apache Kafka, Data Management, Machine Learning Operations, Software Coding, Data Pipelines, Docker, Databricks - **Published:** October 2, 2026 - **Apply:** https://citi.wd5.myworkdayjobs.com/2/job/Irving-Texas-United-States/Data-Solutions-Engineer--Assistant-Vice-President_26999171/apply ## About the Role * 5+ years of experience in data engineering, software engineering, or a closely related discipline. * Practical experience designing and building large-scale data pipelines and distributed data processing solutions in production environments. * Hands-on development experience using Python and Apache Spark (PySpark) to process and transform data at scale. * Experience working with cloud-based data platforms, with exposure to AWS or Google Cloud Platform (GCP). * Working knowledge of data warehousing concepts, data modeling, and distributed processing frameworks including Spark, Hadoop, Kafka, or Hive. * Drive implementation, consistent patterns, reusable components, and coding standards for data engineering processes * Tune Big data applications on Hadoop and non-Hadoop platforms for optimal performance * Be the technical expert and partner with team members on Big Data and Cloud Tech stacks. * Familiarity with containerization tools such as Docker and Kubernetes. * Clear communication skills with the ability to work effectively across technical and non-technical teams. Beneficial Skills & Qualifications * Experience using Snowflake/Databricks for cloud data warehousing and analytics. * Exposure to migrating or modernizing legacy data platforms, including SAS-based workflows. * Familiarity with machine learning data pipelines and MLOps concepts. * Development experience in Scala or Java alongside Python. * AWS or GCP cloud certification(s)., Note: Applicants must be authorized to work in the U.S for this position; Citi will not sponsor applicants for U.S. work authorization for this role. Candidate must be located within commuting distance or be willing to relocate to the area. ## Description * Build and maintain scalable, high-performance data pipelines that support analytics, reporting, and machine learning capabilities across the organization. * Develop Big Data solutions using Apache Spark, Hadoop, and cloud-native services to process and manage large volumes of data reliably. * Partners with data scientists and analytics teams to engineer data pipelines that support advanced analytics and machine learning use cases. * Collaborate with architects and senior engineers to gather requirements and contribute to the implementation of data platform designs. * Apply data quality controls, monitoring, and automation to maintain data governance and auditability standards. * Research emerging technologies and share recommendations that could improve platform performance, scalability, or maintainability. * Support the development of junior engineers by sharing knowledge and technical best practices within the team. * Participate in Agile delivery activities including sprint planning, development, testing, and production support. ## Related Videos - [PySpark - Combining Machine Learning & Big Data](https://www.wearedevelopers.com/videos/44-pyspark-combining-machine-learning-big-data) - [Alibaba Big Data and Machine Learning Technology](https://www.wearedevelopers.com/videos/37-alibaba-big-data-and-machine-learning-technology) - [From Messy Queries to Scalable Systems - How Data Engineering actually works](https://www.wearedevelopers.com/videos/100203-from-messy-queries-to-scalable-systems-how-data-engineering-actually-works) - [Docker Compose: Rediscovered](https://www.wearedevelopers.com/videos/1978-docker-compose-rediscovered) - [Data Science in Retail](https://www.wearedevelopers.com/videos/586-data-science-in-retail) - [Empowering Retail Through Applied Machine Learning](https://www.wearedevelopers.com/videos/976-empowering-retail-through-applied-machine-learning) ## Related Articles - [Top Big Data Technologies That You Need to Know](https://www.wearedevelopers.com/magazine/108-top-big-data-technologies-that-you-need-to-know) - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [Coffee with Developers - Maria Apazoglou - Making AI understandable for all in production](https://www.wearedevelopers.com/magazine/475-coffee-with-developers-maria-apazoglou-making-ai-understandable-for-all-in-production) - [Making Data Warehouses Fast: A Developer’s Story](https://www.wearedevelopers.com/magazine/107-making-data-warehouses-fast-a-developer-s-story) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Data Analyst Salary in the UK](https://www.wearedevelopers.com/magazine/278-data-analyst-salary-in-the-uk)