Senior AI Data Engineer

Cortica - Neurodevelopmental
San Diego, CA, United States
2 months ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
1 year minimum
Compensation
$160,000.0 - $200,000.0
Working hours
Regular working hours
Job source

Tech stack

Web Interfaces Application Programming Interfaces (APIs) Artificial Intelligence Amazon Web Services Amazon S3 Data Analysis Unit Testing Microsoft Azure Big Data Data Validation Information Engineering Data Governance
+35 more
Data Infrastructure Extract Transform Load (ETL) Data Security Dataspaces Data Systems Cursor (Graphical User Interface Elements) Database Queries Dimensional Modeling First Data Python (Programming Language) PostgreSQL MySQL Node.Js Operational Databases Query Optimization Power BI Salesforce.Com Software Engineering SQL Stored Procedures Data Streaming Apex Code Flask (Web Framework) Snowflake VeevaCRM Build Management Data Lakes Integration Tests Kubernetes AWS Glue AWS Data Analytics Apache Kafka Graphql Restful APIs Data Pipelines Mulesoft

Job description

The Senior AI Data Engineer will serve as both architect and builder of our data ecosystem. Every initiative will follow a complete engineering lifecycle: gathering stakeholder requirements, designing the solution, building and testing it, and shipping it to production. This role will work across data lakes, analytics pipelines, and lightweight application development- the multi-disciplinary data equivalent of a full-stack developer., The Senior AI Data Engineer will work closely with the data science, finance, and clinical operations teams to design intelligent, automated data solutions that power care decisions, financial planning, and operational efficiency. AI augmentation is not optional - it is the standard working mode., * Engage stakeholders directly to gather, clarify, and document project requirements.

  • Translate requirements into architected data solutions: choose the right storage, pipeline, modeling, and delivery approach for each problem.
  • Own testing end-to-end - unit tests, data quality checks, reconciliation, and integration tests before anything reaches production.
  • Deploy solutions to production and monitor post-deployment health, iterating rapidly based on real-world feedback.

AI-First Data Engineering:

  • Run parallel AI coding sessions (Claude Code, Cursor, Codex) across different facets of a pipeline simultaneously - orchestrate, verify, and integrate the outputs.
  • Build and maintain context files (CLAUDE.md equivalents) for data projects that encode schema conventions, pipeline patterns, and institutional knowledge - making every future AI session smarter.
  • Design verification loops: automated data quality checks, dbt tests, CI hooks, and pipeline monitors that give AI agents concrete feedback on correctness. * Build MCP (Model Context Protocol) or equivalent integrations to connect AI agents directly to Snowflake, Amazon Athena, Postgresql, MySql, Power BI APIs, Salesforce, and internal tooling.
  • Prefer frontier models for complex architectural decisions and rely on AI acceleration to dramatically increase engineering throughput.

Data Platform & Pipeline Engineering:

  • Design and build complex, reliable data pipelines ingesting from AWS, Azure, Salesforce, MuleSoft, and multiple third-party APIs into our AWS Data Lake and Snowflake warehouse.
  • Implement and evolve data models using Kimball methodology to support financial, operational, and clinical analytics.
  • Optimize pipeline performance, manage data quality, and perform root-cause analysis on data anomalies - internal and external. * Develop and maintain orchestration workflows in Python, and AWS Glue.
  • Continuously evolve the data schema as business and engineering requirements change.

Analytics & Reporting Enablement:

  • Build and support Power BI data models and reports; empower analytics team members to self-serve on a reliable data foundation.
  • Work with data analysts and data scientists to build reusable, well-documented pipeline components they can extend independently.
  • Deliver data products that drive clinical care decisions, financial planning, and operational performance improvements.

Application Development:

  • Build lightweight internal data applications and tooling where needed; data entry interfaces, operational dashboards, automation scripts that bridge the gap between data pipelines and end users.
  • Design for agentic workflows: build AI-powered data tools accessible via web interfaces or Slack that surface insights proactively.
  • Integrate with Salesforce Health Cloud and other platforms using APIs and event-driven patterns.

Security, Governance & Collaboration:

  • Ensure data security and HIPAA compliance in all pipeline and application work. Partner with IT to enforce data governance standards.
  • Document decisions, tradeoffs, and architecture clearly so that future engineers (and AI agents) can build on your work effectively.
  • Collaborate across IT, finance, clinical operations, and data science - acting as the connective tissue between data infrastructure and business outcomes., Cortica cares deeply about the well-being of each member of our team, and we have created a passionate, caring, and growth-minded culture that helps teammates thrive! As a Cortica teammate, we’ll support your well-being through medical, dental, and vision insurance, a 401(k) plan with company matching and rapid vesting, paid holidays and wellness days, life insurance, disability insurance options, tuition reimbursements for professional development and continuing education, and referral bonuses. We value you and the experience you bring to your role, and are proud to provide you with a compensation and benefits package designed to enhance all aspects of your life.

Requirements

Do you have experience in Unit testing?, * You have 5+ years of hands-on data engineering experience, including building and operating production data pipelines.

  • You have expert-level Python skills for ETL, pipeline orchestration, and automation.
  • You possess deep SQL proficiency - query optimization, data modeling, stored procedures.
  • You bring 2+ years’ experience working with AI first development workflows.
  • You bring 4+ years’ experience with the following AWS (S3, Glue, Lambda, Redshift), and/or Azure big data services.
  • You have 1+ year of experience with Snowflake.
  • You have 2+ years of experience with orchestration frameworks.
  • You have 2+ years of Salesforce experience with Apex and configurations.
  • You’re experienced with Kimball dimensional modeling - you’ve built star schemas and conformed dimensions in production.
  • You have Power BI (or equivalent BI tool) experience - data model design and report development.
  • You have API integration experience - REST, GraphQL, event streaming (Kafka, Kinesis, or similar).
  • You possess application development literacy - comfortable building lightweight web tooling (Python/Flask, Node, or similar) to complement data products.
  • You reside in one of the following states: CA, TX, NC, WA, ID, NV, AZ, CO, KS, AR, LA, AL, GA, FL, SC, TN, VA, MD, NJ, DE, IL, WI, MI, OH, MA, PA, NH, CT

Benefits & conditions

Pulled from the full job description

  • Referral program
  • Health insurance
  • 401(k) matching
  • Vision insurance
  • Dental insurance
  • Life insurance
  • Disability insurance, The base pay range for this opening is $160,000 to $200,000. According to your skill level, relevant experience, education level, and location, you will receive compensation that fits appropriately within the range. EOE. This posting is not meant to be an exhaustive list of the role and its duties.

About the company

Cortica is a rapidly growing healthcare company pioneering the most effective treatment methods for children with neurodevelopmental differences. Our mission is to design and deliver life-changing care - one child, one family, one community at a time. Ultimately, we envision a world that cultivates the full potential of every child. At Cortica, every team member is instrumental in helping us achieve our mission!

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on indeed.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:18 min

Scaling MySQL databases for massive user growth

Johannes Nicolai Johannes Nicolai +1 · LIVE

2:52 min

Generating APIs with the Neo4j GraphQL library

William Lyon · LIVE

45 sec

Working securely with Node.js path application programming interfaces

Sonya Moisset · WWC 2023

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

1:48 min

Analyzing network packets with database protocol tools

Daniël van Eeden Daniël van Eeden · WWC Europe 2026

1:31 min

Essential AI and human skills for future teams

Alexander Weißhaupt Alexander Weißhaupt +1 · WWC 2025

Videos

See all

Related articles

See all