Data Engineer III - Python, Databricks

JPMorgan Chase & Co.
Houston, TX, United States
26 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Experienced
Experience required
3 years minimum
Working hours
Regular working hours

Tech stack

Java (Programming Language) JavaScript (Programming Language) Artificial Intelligence Amazon Web Services Code Generation Information Engineering Data Security Cursor (Graphical User Interface Elements) Information Lifecycle Management Python (Programming Language) PostgreSQL MongoDB
+18 more
NoSQL Systems Development Life Cycle Software Deployment Software Engineering SQL Databases Systems Integration Enterprise Data Management Data Processing GitHub Copilot Prompt Engineering Data Layers Containerization Data Lakes AWS Glue Physical Data Models Data Pipelines Docker Databricks

Job description

  • Develop workflows and ELT pipelines using Python and Databricks.
  • Support review of controls to ensure sufficient protection of enterprise data.
  • Implement data security using entitlements frameworks.
  • Update logical or physical data models based on new use cases.
  • Use SQL frequently and understand NoSQL databases
  • Uses enterprise-authorized AI capabilities within the work environment to accelerate data pipeline/design analysis and documentation, validating outputs and handling data according to sensitivity and security requirements.
  • Applies reuse-first, AI-assisted practices to strengthen SDLC-quality routines for data pipelines (e.g., test generation and control validation), ensuring traceability/auditability and alignment to resiliency and security expectations.

Requirements

  • Formal training or certification on software engineering concepts and 3 years applied experience.
  • Good working knowledge of AWS, Databricks, and Python, Experience across the data lifecycle.
  • Advanced at SQL, including joins and aggregations, Working understanding of NoSQL databases.
  • Significant experience with statistical data analysis and ability to determine appropriate tools and data patterns for analysis.
  • Utilize AWS Cloud Services for developing, deploying, and managing applications at scale.
  • Proficiency in AI Coding Assistants, Daily use of tools like Cursor, GitHub Copilot, and Claude to accelerate code generation, documentation, and refactoring.
  • Effective Prompt Engineering, Providing AI models with context, clear goals, relevant source material, and - defined output expectations to generate accurate, usable code.
  • Critical Evaluation & Validation, Ability to identify hallucination patterns, security vulnerabilities, and logic errors in AI-generated code, ensuring safety before production deployment.
  • Demonstrated experience using enterprise-authorized AI capabilities within the work environment to support data engineering workflows with strong validation habits and awareness of data sensitivity.
  • Ability to review and validate AI-assisted outputs (e.g., query suggestions, test ideas, or model change summaries) before use, escalating when uncertain and following data handling requirements., * Familiarity with the Standardized data layer practices (Medallion architecture)
  • Exposure to Aurora Postgres and MongoDB
  • Experience developing and supporting AWS GLUE Jobs, Federated Data Lake
  • Skills in designing efficient data models including normalization, denormalization, and schema design and an understanding around relational and star schemas.
  • Augmented Development Workflow: Integrating tools into CI/CD pipelines, containerization (e.g., Docker), and leveraging AI to quickly bridge language gaps (e.g., transitioning between Python, JavaScript, or Java).

Benefits & conditions

We offer a competitive total rewards package including base salary determined based on the role, experience, skill set and location. Those in eligible roles may receive commission-based pay and/or discretionary incentive compensation, paid in the form of cash and/or forfeitable equity, awarded in recognition of individual achievements and contributions. We also offer a range of benefits and programs to meet employee needs, based on eligibility. These benefits include comprehensive health care coverage, on-site health and wellness centers, a retirement savings plan, backup childcare, tuition reimbursement, mental health support, financial coaching and more. Additional details about total compensation and benefits will be provided during the hiring process.

About the company

As a Data Engineer III at JPMorgan Chase within our agile team, you will design and deliver reliable data collection, storage, access, and analytics solutions that are secure, stable, and scalable.. You will develop, test, and maintain essential data pipelines and architectures, supporting various business functions to achieve the firm’s goals. Working with us, you will use your skills to drive innovation and help shape our team culture. Together, we focus on excellence, collaboration, and continuous improvement., JPMorganChase, one of the oldest financial institutions, offers innovative financial solutions to millions of consumers, small businesses and many of the world’s most prominent corporate, institutional and government clients under the J.P. Morgan and Chase brands. Our history spans over 200 years and today we are a leader in investment banking, consumer and small business banking, commercial banking, financial transaction processing and asset management.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on jpmc.fa.oraclecloud.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:37 min

Comparing traditional SQL tables versus NoSQL non-tabular databases

Stanimira Vlaeva · JS Congress

2:07 min

Inspecting default bridge architectures and custom Docker networks

Oliver Seitz Oliver Seitz · WWC 2025

2:01 min

Migrating existing applications from MongoDB to Postgres

Nikita Shamgunov Nikita Shamgunov · WWC 2024

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

3:16 min

Terminology differences between relational and NoSQL databases

Tim Faulkes · LIVE

3:09 min

Balancing data science skillings alongside systems engineering rigor

Nico Schmidt · LIVE

Videos

See all

Related articles

See all