> Markdown version of [/jobs/ext/2246138-data-platform-engineer](https://www.wearedevelopers.com/jobs/ext/2246138-data-platform-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Data Platform Engineer - **Company:** Worth AI - **Location:** Miami, FL, United States (Remote available) - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** Java (Programming Language), Geographic Information Systems, Application Programming Interfaces (APIs), Airflow, Amazon Web Services, Application Performance Management, Systems Engineering, BigQuery, Continuous Integration, Data as a Services, Data Infrastructure, Extract Transform Load (ETL), Data Structures, Data Warehousing, Query Languages, DevOps, Graph Database, Python (Programming Language), Machine Learning, Neo4j, Software Engineering, Web Services, Rust (Programming Language), Snowflake, Apache Spark, Build Management, Containerization, Core Data, Kubernetes, Apache Flink, ISO 20022 Standard, Integration Frameworks, Apache Kafka, Graphql, Data Management, Restful APIs, Terraform, Data Pipelines, Docker, Amazon Redshift, Databricks, Golang - **Published:** August 26, 2026 - **Apply:** https://www.careerjet.com/jobad/us14ccd8d28877d45bbfcd60bac6fcffae ## About the Role * Expertise in Graph Ecosystems: Hands-on experience with Graph databases (e.g., Neo4j, AWS Neptune, or TigerGraph) and query languages like Cypher or Gremlin * Identity & Linkage Mastery: Proven experience with Entity Resolution or Record Linkage (e.g., using tools like Senzing, Quantexa, or custom probabilistic matching models) * Schema Design: Ability to design flexible ontologies that handle evolving regulatory data (e.g., changing PEP definitions or Sanction list formats) * API Performance for Graphs: Experience building GraphQL or REST APIs specifically optimized for graph traversals and deep-tree lookups * Experience building centralized data platforms or "data-as-a-service" offerings at scale (e.g., at a large tech or cloud-native company) * Strong software engineering skills in at least one language commonly used for data and services (e.g., Python, Java, Go, Rust) * Hands-on experience building data pipelines and ETL/ELT workflows on a major cloud provider (AWS preferred) * Experience with modern data stack tools such as Spark/Flink, Kafka/Kinesis, Airflow/managed schedulers, and data warehouses (e.g., Snowflake, Redshift, BigQuery, Databricks) * Familiarity with DevOps practices: CI/CD, containerization (Docker), orchestration (Kubernetes), and infrastructure-as-code (Terraform) * Strong focus on observability (metrics, logs, traces), resilience, and building early warning signals * Comfort collaborating cross-functionally and communicating clearly with both technical and non-technical stakeholders. Nice to Have * Background supporting machine learning or real-time decisioning use cases from a platform point of view * Compliance Domain Knowledge: Understanding of AML, CTF, and KYC/KYB data structures (e.g., LEIs, ISO 20022) * Geospatial Data: Experience handling global address normalization and geospatial indexing for risk detection ** All Remote Hires - will be required to travel to Orlando, Florida at least twice per year for Town Halls and team collaboration in addition to orientation in Orlando, Florida ## Description As a Data Platform Engineer, you will design, build, and operate the core data services that power our products and analytics. You'll own end-to-end data pipelines and API services that ingest, process, and expose high-quality data to internal customers (data science, analytics, product, and other engineering teams) and external partners. You'll be part of a small, high-impact team that treats the data platform as a product with strong SLAs, and reliable self-service for internal and external users. Responsibilities What you'll do: * Architect and implement entity resolution logic to de-duplicate and link disparate data points into unified "Golden Records" for businesses and individuals * Design and maintain a high-performance global business knowledge graph and ontology to map complex ownership chains, UBOs, and hidden risk relationships across international borders * Implement a hybrid storage strategy that bridges graph databases for relationship mapping with document and search stores for rich metadata and adverse media content * Optimize the platform for real-time risk assessment, ensuring the ability to traverse multiple levels of ownership in milliseconds to support automated "Go/No-Go" onboarding decisions * Design and build scalable data services and APIs for ingesting, transforming, and serving data across the company * Develop and maintain batch and streaming data pipelines using modern data processing frameworks and AWS cloud-native tooling * Own the reliability, performance, and API first data platform, including monitoring, alerting, and on-call where appropriate * Implement best practices for data modeling, quality, lineage, and governance to ensure trustworthy, well-documented datasets * Work closely with data scientists, analysts, and application engineers to understand their needs and translate them into robust platform capabilities * Drive automation and standardization through CI/CD, model as a service, and reproducible environments * Help define and evolve the architecture of our data platform as a true internal service with clear contracts, SLAs, and versioned APIs, Senior Systems Engineer Endpoint Engineering We are Lennar Lennar is one of the nation's leading homebuilders, dedicated to making an impact and creating an extraordinary exp… + 7 days ago + ## Related Videos - [Putting the Graph In GraphQL With The Neo4j GraphQL Library](https://www.wearedevelopers.com/videos/257-putting-the-graph-in-graphql-with-the-neo4j-graphql-library) - [From DevOps to Scaled DevOps: How We’re Rebuilding Continuous Delivery as a Platform](https://www.wearedevelopers.com/videos/100018-from-devops-to-scaled-devops-how-we-re-rebuilding-continuous-delivery-as-a-platform) - [Docker Compose: Rediscovered](https://www.wearedevelopers.com/videos/1978-docker-compose-rediscovered) - [From Messy Queries to Scalable Systems - How Data Engineering actually works](https://www.wearedevelopers.com/videos/100203-from-messy-queries-to-scalable-systems-how-data-engineering-actually-works) - [Cyber Sleuth: Finding Hidden Connections in Cyber Data](https://www.wearedevelopers.com/videos/893-cyber-sleuth-finding-hidden-connections-in-cyber-data) - [Docker build without Docker](https://www.wearedevelopers.com/videos/100114-docker-build-without-docker) ## Related Articles - [Graph and AI Trends 2026: Why Is AI Running but Not Yet Delivering?](https://www.wearedevelopers.com/magazine/680-graph-and-ai-trends-2026-why-is-ai-running-but-not-yet-delivering) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [Coffee with Developers - Maria Apazoglou - Making AI understandable for all in production](https://www.wearedevelopers.com/magazine/475-coffee-with-developers-maria-apazoglou-making-ai-understandable-for-all-in-production) - [Making Data Warehouses Fast: A Developer’s Story](https://www.wearedevelopers.com/magazine/107-making-data-warehouses-fast-a-developer-s-story) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs)