Forward Deployed Data Engineer

Databricks
Houston, TX, United States
14 days ago
Apply on www.careerbuilder.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Compensation
$180,000.0 - $230,000.0
Working hours
Regular working hours

Tech stack

Unity 3d Application Programming Interfaces (APIs) Artificial Intelligence Airflow Amazon Web Services Architectural Patterns Cloud Computing Code Review Cyber Security Continuous Delivery Continuous Integration Information Engineering
+33 more
Data Infrastructure Data Security Software Debugging Python (Programming Language) Operational Databases Oracle (Applications) Performance Tuning DataOps Requirements Management Software Engineering SQL Stored Procedures PL-SQL SQL Databases Teradata SQL Scripting Google Cloud Enterprise Software Applications Sql Optimization Informatica Powercenter Netezza Large Language Models Data Strategy Git Data Lakes Pyspark Production Code Data Management Virtual Agents Terraform Code Restructuring Cisco Legacy Systems Databricks

Job description

As a Forward Deployed Data Engineer (FDE) on our US Launch Team, you will act as the principal technical anchor embedded directly within high-stakes enterprise client accounts (focusing heavily on Financial Services and Healthcare). You won’t just be writing specs or advising from afar-you will be on the front lines in client environments, bridging the gap between C-suite data strategy and hands-on production code.

Working at the tip of the spear alongside our elite partners at Databricks and Anthropic, you will lead architecture design, modernize brittle legacy ecosystems into cloud-native Lakehouses, build agentic AI data harnesses, and guide integrated “SWAT” pods to ship working software in weeks rather than months., * Embedded Technical Leadership: Serve as the primary technical authority on client engagements. Partner directly with client VP/CTO stakeholders to scope architectures, map data domains, and turn complex requirements into execution-ready engineering plans.

  • Hands-On Lakehouse Modernization: Roll up your sleeves to write and optimize production PySpark, Databricks SQL, and Delta Live Tables (DLT)-actively refactoring complex legacy logic (Informatica, PL/SQL, legacy stored procedures) into Medallion Architectures.
  • Agentic AI & Data Infrastructure: Architect deterministic, secure data “harnesses” that integrate LLM workflows (Anthropic/Claude) into traditional enterprise data pipelines safely and cost-effectively.
  • DataOps & IaC Ownership: Write modular Terraform scripts to provision cloud environments (AWS/GCP), manage governance and lineage via Unity Catalog, and enforce CI/CD rigor across projects.
  • Pod & Delivery Guidance: Partner with our LATAM-based nearshore engineering squads (600+ experts) as the US technical lead-conducting code reviews, debugging performance bottlenecks, and maintaining high engineering standards.
  • Technical Unblocking & Escalation: Act as the ultimate technical safety net. If a pipeline breaks or a deployment stalls at 2:00 AM, you have the hands-on depth to jump into the code, fix the issue, and keep client delivery on track.

Requirements

Technical Mastery & Execution

  • Extensive Hands-On Data Engineering: Deep proficiency in Python and advanced SQL with a proven track record of shipping production data software.
  • Deep Databricks Platform Expertise: Hands-on experience building, scaling, and tuning Databricks Lakehouse environments (Delta Lake, Unity Catalog, PySpark, DLT, Workflows).
  • Modern Stack & IaC Discipline: Expert-level experience with dbt (Core or Enterprise) for data modeling, Terraform for Infrastructure as Code, and Git/Airflow for orchestration and CI/CD.
  • Legacy Migration Experience: Demonstrated background migrating enterprise client workloads from legacy platforms (Teradata, Netezza, Informatica, Oracle) to modern cloud infrastructure.

Client Presence & Startup Grit

  • The “Player-Coach” Mindset: Equal comfort pitching target architectures to a client CTO and debugging a failing PySpark job or dbt macro alongside junior engineers.
  • Executive Communication: Ability to translate messy business requirements into clean architectural patterns and speak authoritatively across both technical and business functions.
  • Scrappy Nation-Builder: High autonomy and resilience-thriving in a fast-paced, Series A growth environment without relying on rigid corporate playbooks or large support structures.

Nice-to-Haves

  • Prior experience in forward-deployed, technical consulting, or ProServe roles at top-tier agencies or hyper-growth vendors.
  • Pragmatic experience building or deploying LLM/GenAI orchestration frameworks (LangChain, LlamaIndex, Anthropic API) into production pipelines.
  • Official certifications in Databricks, AWS, GCP, or dbt.

The anticipated base salary range for this role is $180,000 - $230,000. In addition to base pay, this position may be eligible for an annual discretionary bonus. An individual’s final salary offer will be determined based on a variety of factors, including geographic location, experience, specialized skills, and qualifications. This compensation range is subject to updates or modifications at the company’s discretion

Skills: Amazon Web Services (AWS), Application Programming Interface (API), Architectural Design, Architectural Services, Artificial Intelligence (AI), Cisco Unity, Cloud Computing, Coaching, Code Reviews, Continuous Deployment/Delivery, Continuous Integration, Data Management, Debugging Skills, Ecosystems, Embedded Systems, Enterprise Applications, Financial Services, GCP (Good Clinical Practices), Healthcare, Informatica, Information/Data Security (InfoSec), NCR Teradata, Oracle, Oracle PL-SQL, Python Programming/Scripting Language, Refactoring, Requirements Management, SQL (Structured Query Language), Scripting (Scripting Languages), Software Engineering, Startup, Stored Procedures, Technical Consulting, Technical Leadership

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.careerbuilder.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

6:21 min

Investigating push inefficiencies with upstream Git experts

Jonathan Creamer · Coffee With Developers

1:29 min

Expanding practical knowledge with community sandboxes and resources

Stuart Clark · LIVE

3:09 min

Balancing data science skillings alongside systems engineering rigor

Nico Schmidt · LIVE

56 sec

Favorite git commands and the importance of patch commits

Eileen Uchitelle Eileen Uchitelle +1 · Coffee With Developers

3:45 min

Prototyping deterministic agents with n8n and PyATS

Alfonso Sandoval Rosas Alfonso Sandoval Rosas · Europe 2026 Virtual

Videos

See all

Related articles

See all