Data Engineer 5 (Python, SQL, Databricks, Snowflake) (Enterprise Platforms Technology)

Capital One Financial Corporation
Riverwoods, IL, United States
2 days ago
Apply on dejobs.org
Prepare application

Role details

Contract type
Internship / Graduate position
Employment type
Part-time / full-time
Experience level
Experienced
Experience required
2 years minimum
Compensation
$229,900.0 - $262,400.0
Working hours
Regular working hours
Job source

Tech stack

Query Performance Java (Programming Language) Artificial Intelligence Amazon Web Services Data Analysis Microsoft Azure Computer Programming Information Engineering Data Governance Data Infrastructure Dataspaces Data Systems
+25 more
Data Warehousing Software Design Patterns Distributed Data Store Python (Programming Language) Machine Learning Standard Sql Scala (Programming Language) Software Engineering SQL Databases Data Streaming Unstructured Data Google Cloud Data Classification Feature Engineering Large Language Models Snowflake IT Architecture Generative AI Data Layers Information Technology Data Analytics Real Time Data Non-relational Database Data Pipelines Databricks

Job description

We are seeking a Data Engineer to develop the technical vision, architectural design, and implementation of our AI-enabling data ecosystem.

In this strategic role, you will bridge the gap between traditional enterprise data architecture and modern AI capabilities. You will design resilient, scalable data pipelines, and real-time streaming architectures that power LLM workflows, Retrieval-Augmented Generation (RAG) pipelines, and predictive ML models. You will work closely with Data Scientists, Analysts, and other Engineers to establish best-in-class data engineering practices for traditional and AI enabled workloads.

About the Team

The Enterprise Platforms Tech Top of House Data & Analytics team serves as the strategic intelligence backbone for executive leadership and strategy. We operate at the intersection of business strategy, core enterprise infrastructure, and advanced analytics. Our team is responsible for delivering high-impact, enterprise-level data products, decisioning engines, and modern analytics capabilities that power executive level decision-making., * AI Architecture & Data Foundation: Architect, build, and scale clean, reliable, and latency-optimized data pipelines (batch and real-time) designed specifically to supply structured, semi-structured, and unstructured data to AI/ML applications and Large Language Models (LLMs).

  • Technical Leadership & Strategy: Contribute to defining the architectural blueprint for the Top of House AI data layer. Serve as a hands-on technical lead, guiding junior and mid-level engineers in coding standards, design patterns, and engineering excellence.
  • Data Governance, Security, & Lineage: Partner with Enterprise Security and Data Governance teams to implement robust data classification, privacy safeguards, and automated lineage tracking for AI models and training datasets.
  • Feature Engineering & Store: Build and maintain scalable feature stores to support both real-time inferencing and offline model training across executive-facing predictive models.
  • Performance & Cost Optimization: Audit and optimize cross-cloud data pipelines, query performance, and storage infrastructure for efficiency, cost management, and reliability.
  • Cross-Functional Collaboration: Partner directly with Technical Program Managers, Data Scientists, Top of House leads, and C-suite stakeholders to turn executive data needs into production-ready solutions.

Requirements

  • Bachelor’s Degree or higher in Computer Science or a related quantitative field (Statistics, Economics, Operations Research, Analytics, Mathematics, Engineering)
  • At least 6 years of experience in application development (Internship experience does not apply)
  • At least 4 years of experience in distributed data
  • At least 4 years of experience with SQL
  • At least 4 years of experience programming with at least one of the following languages: Python, Java, or Scala
  • At least 4 years of experience designing and developing data pipelines
  • At least 2 years of experience in data modeling and designing end-to-end data solutions using both relational and non-relational database systems, * Master’s Degree in a related field
  • 9+ years of experience in application development including Python, SQL, Scala, or Java
  • 5+ years of experience with a public cloud (AWS, Microsoft Azure, Google Cloud)
  • 5+ year experience working on real-time data and streaming applications
  • 5+ years of data warehousing experience (eg Snowflake)
  • Experience leveraging interactive AI tooling to accelerate productivity, utilizing capabilities beyond basic code completion (Claude, Gemini)

Benefits & conditions

The minimum and maximum full-time annual salaries for this role are listed below, by location. Please note that this salary information is solely for candidates hired to perform work within one of these locations, and refers to the amount Capital One is willing to pay at the time of this posting. Salaries for part-time roles will be prorated based upon the agreed upon number of hours to be regularly worked.

McLean, VA: $229,900 - $262,400 for Data Engineer 5

New York, NY: $250,800 - $286,200 for Data Engineer 5

Plano, TX: $209,000 - $238,500 for Data Engineer 5

About the company

Capital One offers a comprehensive, competitive, and inclusive set of health, financial and other benefits that support your total well-being. Learn more at the Capital One Careers website (https://www.capitalonecareers.com/benefits) . Eligibility varies based on full or part-time status, exempt or non-exempt status, and management level., Capital One Financial is made up of several different entities. Please note that any position posted in Canada is for Capital One Canada, any position posted in the United Kingdom is for Capital One Europe and any position posted in the Philippines is for Capital One Philippines Service Corp. (COPSSC).

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on dejobs.org
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

1:33 min

Integrating internal APIs and maintaining data sovereignty

Mahran Meißner Mahran Meißner · World Congress 2026 Europe

2:27 min

Managing traffic and tracking costs with Databricks Unity Catalog

Viktoria Semaan Viktoria Semaan · World Congress 2026 Europe

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

2:46 min

Transforming data architecture from on-premise to cloud

Sandhya Menon Sandhya Menon · World Congress 2026 Europe

3:09 min

Balancing data science skillings alongside systems engineering rigor

Nico Schmidt · LIVE

2:50 min

Executing LoRA fine-tuning using serverless Databricks AI runtimes

Viktoria Semaan Viktoria Semaan · World Congress 2026 Europe

Videos

See all

Related articles

See all