> Markdown version of [/jobs/ext/3115580-data-engineer-5-python-sql-databricks-snowflake-enterprise-platforms-technology](https://www.wearedevelopers.com/jobs/ext/3115580-data-engineer-5-python-sql-databricks-snowflake-enterprise-platforms-technology). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Data Engineer 5 (Python, SQL, Databricks, Snowflake) (Enterprise Platforms Technology) - **Company:** Capital One Financial Corporation - **Location:** McLean, VA, United States - **Experience:** Experienced - **Salary:** $229,900.0 - $262,400.0 - **Contract:** Internship / Graduate position - **Skills:** Query Performance, Java (Programming Language), Artificial Intelligence, Amazon Web Services, Data Analysis, Microsoft Azure, Computer Programming, Information Engineering, Data Governance, Data Infrastructure, Dataspaces, Data Systems, Data Warehousing, Software Design Patterns, Distributed Data Store, Python (Programming Language), Machine Learning, Standard Sql, Scala (Programming Language), Software Engineering, SQL Databases, Data Streaming, Unstructured Data, Google Cloud, Data Classification, Feature Engineering, Large Language Models, Snowflake, IT Architecture, Generative AI, Data Layers, Information Technology, Data Analytics, Real Time Data, Non-relational Database, Data Pipelines, Databricks - **Published:** September 27, 2026 - **Apply:** https://dejobs.org/x/x/F9B264A7AD2441EFA28E8CBBA3142E6A/job/ ## About the Role * Bachelor's Degree or higher in Computer Science or a related quantitative field (Statistics, Economics, Operations Research, Analytics, Mathematics, Engineering) * At least 6 years of experience in application development (Internship experience does not apply) * At least 4 years of experience in distributed data * At least 4 years of experience with SQL * At least 4 years of experience programming with at least one of the following languages: Python, Java, or Scala * At least 4 years of experience designing and developing data pipelines * At least 2 years of experience in data modeling and designing end-to-end data solutions using both relational and non-relational database systems, * Master's Degree in a related field * 9+ years of experience in application development including Python, SQL, Scala, or Java * 5+ years of experience with a public cloud (AWS, Microsoft Azure, Google Cloud) * 5+ year experience working on real-time data and streaming applications * 5+ years of data warehousing experience (eg Snowflake) * Experience leveraging interactive AI tooling to accelerate productivity, utilizing capabilities beyond basic code completion (Claude, Gemini) ## Description We are seeking a Data Engineer to develop the technical vision, architectural design, and implementation of our AI-enabling data ecosystem. In this strategic role, you will bridge the gap between traditional enterprise data architecture and modern AI capabilities. You will design resilient, scalable data pipelines, and real-time streaming architectures that power LLM workflows, Retrieval-Augmented Generation (RAG) pipelines, and predictive ML models. You will work closely with Data Scientists, Analysts, and other Engineers to establish best-in-class data engineering practices for traditional and AI enabled workloads. About the Team The Enterprise Platforms Tech Top of House Data & Analytics team serves as the strategic intelligence backbone for executive leadership and strategy. We operate at the intersection of business strategy, core enterprise infrastructure, and advanced analytics. Our team is responsible for delivering high-impact, enterprise-level data products, decisioning engines, and modern analytics capabilities that power executive level decision-making., * AI Architecture & Data Foundation: Architect, build, and scale clean, reliable, and latency-optimized data pipelines (batch and real-time) designed specifically to supply structured, semi-structured, and unstructured data to AI/ML applications and Large Language Models (LLMs). * Technical Leadership & Strategy: Contribute to defining the architectural blueprint for the Top of House AI data layer. Serve as a hands-on technical lead, guiding junior and mid-level engineers in coding standards, design patterns, and engineering excellence. * Data Governance, Security, & Lineage: Partner with Enterprise Security and Data Governance teams to implement robust data classification, privacy safeguards, and automated lineage tracking for AI models and training datasets. * Feature Engineering & Store: Build and maintain scalable feature stores to support both real-time inferencing and offline model training across executive-facing predictive models. * Performance & Cost Optimization: Audit and optimize cross-cloud data pipelines, query performance, and storage infrastructure for efficiency, cost management, and reliability. * Cross-Functional Collaboration: Partner directly with Technical Program Managers, Data Scientists, Top of House leads, and C-suite stakeholders to turn executive data needs into production-ready solutions. ## Related Videos - [From Messy Queries to Scalable Systems - How Data Engineering actually works](https://www.wearedevelopers.com/videos/100203-from-messy-queries-to-scalable-systems-how-data-engineering-actually-works) - [How Cisco embraced a DevOps culture within its network engineering team](https://www.wearedevelopers.com/videos/99-how-cisco-embraced-a-devops-culture-within-its-network-engineering-team) - [Harnessing the Power of Open Source's Newest Technologies](https://www.wearedevelopers.com/videos/1448-harnessing-the-power-of-open-source-s-newest-technologies) - [Cutting LLM Costs Without Cutting Quality: How to Beat Proprietary LLMs with Fine-Tuned Open Source](https://www.wearedevelopers.com/videos/100151-cutting-llm-costs-without-cutting-quality-how-to-beat-proprietary-llms-with-fine-tuned-open-source) - [OLTP in the Lakehouse: Redefining Data for AI Workloads](https://www.wearedevelopers.com/videos/2038-oltp-in-the-lakehouse-redefining-data-for-ai-workloads) - [Beyond Dashboards: Fixing Text-to-SQL with Semantic RAG](https://www.wearedevelopers.com/videos/2036-beyond-dashboards-fixing-text-to-sql-with-semantic-rag) ## Related Articles - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Making Data Warehouses Fast: A Developer’s Story](https://www.wearedevelopers.com/magazine/107-making-data-warehouses-fast-a-developer-s-story) - [Top Big Data Technologies That You Need to Know](https://www.wearedevelopers.com/magazine/108-top-big-data-technologies-that-you-need-to-know) - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [The Fastest-Growing Tech Sectors to Look Out for in 2025](https://www.wearedevelopers.com/magazine/373-the-fastest-growing-tech-sectors-to-look-out-for-in-2025) - [7 Cloud Computing Trends Coming in 2025 for Developers](https://www.wearedevelopers.com/magazine/412-7-cloud-computing-trends-coming-in-2025-for-developers)