Associate Data Engineer 2027

IBM
Chicago, IL, United States
5 days ago
Apply on dejobs.org
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Working hours
Regular working hours
Job source

Tech stack

Java (Programming Language) Agile Methodology Artificial Intelligence Airflow Amazon Web Services Data Analysis IBM System I Microsoft Azure BigQuery Cloud Computing Cloud Engineering Databases
+26 more
Continuous Integration Data Architecture Data Cleansing Information Engineering Data Integration Extract Transform Load (ETL) Data Visualization Data Warehousing Linux Python (Programming Language) Machine Learning Scala (Programming Language) SQL Databases Unstructured Data Data Ingestion Delivery Pipeline Snowflake Prompt Engineering Apache Spark Information Technology Data Management Cloud Optimization Data Inconsistencies Data Pipelines Databricks Programming Languages

Job description

Ready to think boldly, work with some of the world’s most recognized brands, and kick-start your career? Welcome to the IBM’s Associate Program for university hires.

From day one, you’ll collaborate with global clients and contribute to projects that help organizations solve their toughest challenges across digital transformation, cloud strategy, AI adoption, and process redesign - working alongside a global cohort of diverse, ambitious peers. As an Associate, you’ll have access to industry-recognized certifications, digital badges, and a minimum of 40 hours of structured learning per year on IBM’s AI-driven learning platform - supported by coaches, mentors, and thriving professional communities across practices. IBM’s culture of internal mobility means you’ll have the freedom to explore new technologies, industries, and career paths as your interests evolve.

Bring your curiosity. Grow your skills. Build what’s next - like an IBMer.

To give yourself the best opportunity for success, we advise applying only to roles that align with your skills and experience, rather than applying broadly across all entry-level positions. You’ll receive a status update email for each application, so be sure to check your IBM Careers account regularly - it’s the best way to get a centralized view of which roles you have active applications against.

Your role and responsibilities

As an Associate Data Engineer at IBM you will harness the power of data to unveil captivating stories and intricate patterns. You’ll contribute to data gathering, storage, and both batch and real-time processing.

Collaborating closely with diverse teams, you’ll play an important role in deciding the most suitable data management systems and identifying the crucial data required for insightful analysis. As a Data Engineer, you’ll tackle obstacles related to database integration and untangle complex, unstructured data sets., * Assist in designing and implementing scalable data architecture and management systems tailored for modern cloud environments.

  • Work on optimizing existing data pipelines and processes for improved performance.
  • Work with ETL/ELT ingestion pipelines.
  • Collect and analyze data to identify trends, providing clients with actionable insights to enhance marketing, operational, and business practices.
  • Participate in troubleshooting data-related issues, working to solve challenges and data inconsistencies.
  • Create visually compelling and user-friendly data visualizations, dashboards, and reports to effectively communicate findings to both technical and non-technical stakeholders.
  • Ensure the integrity, accuracy, and reliability of data through rigorous data cleaning, validation, and preprocessing procedures.
  • Work with the project team to prioritize and translate Client requirements and define current and future operational scenarios (processes, models, use cases, plans, and solutions). Work collaboratively with the Client and the Architect to ensure proper translation of business requirements to solution requirements.
  • Present analytical findings and recommendations clearly and concisely, demonstrating the value of data-driven decision-making to clients.

Requirements

  • Advocates business process transformation by analyzing application portfolios across the ecosystem to identify process optimization and automation opportunities, leveraging planning, project management, and Agile methodologies to drive effective solutions.
  • Demonstrates strong communication, problem-solving, adaptability, and teamwork skills, with a collaborative mindset, curiosity to learn, and the ability to clearly articulate ideas, ask insightful questions, and recommend solutions.

Required technical and professional expertise

  • Ability to incorporate a variety of statistical and machine learning techniques.
  • Basic understanding of Cloud (AWS, Azure, etc).
  • Ability to use programming languages like Java, Python, Scala, etc., to build pipelines to extract and transform data from a repository to a data consumer.
  • Ability to use Extract, Transform, and Load (ETL) tools and/or data integration, or federation tools to prepare and transform data as needed.
  • Ability to use leading-edge tools such as Linux, SQL, Python, Spark, Hadoopand Java.
  • Exposure to ETL/ELT projects, data warehouse design, analytics pipelines, and capstone data engineering
  • Willingness to travel up to 100%, based on project requirements.

Preferred technical and professional experience

  • Preferred Bachelor’s degree in a related field (Computer Science, Data Science, Statistics, Math, MIS, Engineering).
  • Preferred Skills/Tech: SQL, Python,dbt, Snowflake, Big Query, Airflow, data modeling, CI/CD basics.
  • Preferred Certifications: SnowProcore, Databricks Data Engineer Associate, Google Associate Cloud Engineer or Data Engineer coursework
  • General skills: GenAI literacy - prompt engineering, RAG, fine-tuning, and evaluation of generative models.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on dejobs.org
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:30 min

Leveraging BigQuery ML for scalable SQL-based segmentation experiments

Julian Joseph · LIVE

2:15 min

Empowering domain teams with an open data platform

Sandhya Menon Sandhya Menon · World Congress 2026 Europe

52 sec

Running persistent Linux environments directly on Windows

Ben Breard Ben Breard · World Congress 2025

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

3:27 min

Explaining query execution overhead and caching limitations in BigQuery

Adnan Rahic · JS Congress

3:05 min

Audience questions on AI agents and pipeline vectorization

Joy Joy · World Congress 2024

Videos

See all

Related articles

See all