> Markdown version of [/jobs/ext/1593326-data-engineer-forward-deployed-engineer](https://www.wearedevelopers.com/jobs/ext/1593326-data-engineer-forward-deployed-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Data Engineer, Forward Deployed Engineer - **Company:** Kyndryl, Inc. - **Location:** Dallas, TX, United States - **Salary:** $143,640.0 - $273,000.0 - **Contract:** Permanent contract - **Skills:** Agile Methodology, Artificial Intelligence, Airflow, Amazon Web Services, Automation of Tests, Microsoft Azure, Batch Processing, Cloud Computing, Cloud Engineering, Data Architecture, Information Engineering, Data Infrastructure, Extract Transform Load (ETL), Database Models, Software Debugging, Distributed Data Store, Distributed Systems, Github, Intelligence Analysis, Python (Programming Language), Machine Learning, Meta-Data Management, Rapid Prototyping Process, Search Technologies, Software Systems, Data Streaming, Database Optimization, Technical Debt, Containerization, Information Technology, Apache Kafka, Machine Learning Operations, Virtual Agents, Data Delivery, Software Version Control, Data Pipelines, Docker, Web Api - **Published:** July 18, 2026 - **Apply:** https://www.juju.com/job/00000000ghcsk7 ## About the Role + Strong, production-grade coding proficiency in Python and advanced SQL optimization. + Practical experience designing, building, and orchestrating ETL/ELT pipelines using tools such as Airflow, dbt, and Kafka. + Extensive technical expertise in database modeling, distributed computing, and deploying resources within cloud-native environments (AWS, Azure, or GCP). + Hands-on experience implementing vector databases (e.g., Pinecone, Milvus, Chroma, Weaviate) and vector indexing strategies for high-context search. + Working knowledge of data quality, pipeline lineage, and modern data observability tools (e.g., Great Expectations, DataHub). + Cloud-native technical certifications on AWS, Azure, or GCP, or a demonstrated willingness to achieve certifications during your journey with us. + Bachelor's degree in Computer Science, Data Science, Engineering, or a closely related technical field, or equivalent practical professional experience. Preferred Skills and Experience + Prior experience designing and implementing data pipelines in highly regulated industries, such as Financial Services, Healthcare, or the Public Sector. + Hands-on experience scaling deployments using containerization tools such as Docker and Kubernetes. + Familiarity with emerging semantic modeling techniques, metadata management, and modern integration patterns for AI agent architectures. + Experience building and monitoring CI/CD pipelines, utilizing version control (Git & GitHub), and delivering software under Agile methodologies. + Advanced technical certifications in specialized database engineering, data architecture, or machine learning. + Master's degree in Computer Science, Data Engineering, or a related discipline. ## Description At Kyndryl, we don't just build technology-we deploy it where it matters most, driving scalable technology modernization and mission-critical transformations for the world's leading enterprises. Within our newly established Lab, we design, build, and deploy real-world enterprise technology where it creates the greatest physical impact. We operate under a modern triad delivery model, combining deep engineering, design, and business strategy to deliver outcomes that keep global industries moving forward. As a Forward Deployed Data Engineer, you are the builder whom customers ask for by name. You will serve as a critical bridge between complex business challenges and scalable data solutions, operating at the intersection of AI innovation and real-world impact. Operating as a technical anchor, you will design, deploy, and refine the high-performance data infrastructure, pipelines, and context engines that fuel next-generation artificial intelligence, machine learning, and autonomous workflows in live customer environments. This is a dynamic, highly collaborative, and hands-on role. Operating in a hybrid model out of our Lab with a minimum of three days in-office, you will work directly alongside architects, software engineers, and client stakeholders in rapid-prototype cycles. You will take true ownership of outcomes, moving quickly from scoping and building initial proof-of-concepts to deploying hardened, production-ready systems that generate immediate business value for our customers. Forward Deployed Pipeline Architecture & Optimization + Design, optimize, and maintain scalable ETL/ELT pipelines to support high-volume batch processing and low-latency real-time streaming. + Architect distributed systems and database models across hybrid and cloud-native environments, ensuring they are optimized for performance, scale, and cost. + Diagnose, debug, and resolve complex performance bottlenecks across diverse database schemas and distributed data layers. AI Grounding & Context Engineering + Deploy, index, and manage vector databases and semantic layers specifically tailored for high-context AI search, retrieval-augmented generation (RAG), and agent memory structures. + Construct robust, secure, and high-throughput API endpoints that enable autonomous systems to interact dynamically with complex, distributed datasets. + Design and configure feature stores to support active machine learning pipelines and real-time inference loops. Data Quality, Observability & Compliance + Implement modern data quality and pipeline observability frameworks to ensure clean, high-fidelity data delivery to downstream systems. + Log pipeline activities, track metadata lineage, and set up automated testing loops to systematically prevent schema drift. + Ensure all built technical solutions align seamlessly with strict industry compliance, security, and data protection regulations. Field Intelligence & Platform Evolution + Capture deployment insights, document real-world best practices, and feed these learnings back into Kyndryl's core technology platforms and frameworks to accelerate future customer implementations. + Collaborate seamlessly across global teams and practices, cutting through organizational silos to prioritize the right customer outcome. + Balance rapid software delivery and field prototyping with long-term platform stability, reducing technical debt as solutions mature. Kyndryl currently does not require employees to be fully vaccinated against COVID-19, however, if you are hired to work at a client, customer, or partner location, you may be required to show proof of vaccination to align with their respective COVID-19 vaccination policies. Those who believe they are eligible may apply for a medical or religious accommodation prior to the start of employment. Who You Are Who You Are You're good at what you do and possess the required experience to prove it. However, equally as important - you have a growth mindset; keen to drive your own personal and professional development. You are customer-focused - someone who prioritizes customer success in their work. And finally, you're open and borderless - naturally inclusive in how you work with others. As a Forward Deployed Engineer, you thrive in dynamic, fast-moving environments and are highly comfortable navigating ambiguity. You are passionate about translating messy, real-world data challenges into elegant software solutions, and you value building strong, trust-based relationships with both technical teams and client stakeholders. ## Related Videos - [From Messy Queries to Scalable Systems - How Data Engineering actually works](https://www.wearedevelopers.com/videos/100203-from-messy-queries-to-scalable-systems-how-data-engineering-actually-works) - [Docker Compose: Rediscovered](https://www.wearedevelopers.com/videos/1978-docker-compose-rediscovered) - [Innovating Developer Tools with AI: Insights from GitHub Next](https://www.wearedevelopers.com/videos/1268-innovating-developer-tools-with-ai-insights-from-github-next) - [Enjoying SQL data pipelines with dbt](https://www.wearedevelopers.com/videos/823-enjoying-sql-data-pipelines-with-dbt) - [AI Model Management Life Circles: ML Ops For Generative AI Models From Research to Deployment](https://www.wearedevelopers.com/videos/1152-ai-model-management-life-circles-ml-ops-for-generative-ai-models-from-research-to-deployment) - [Docker build without Docker](https://www.wearedevelopers.com/videos/100114-docker-build-without-docker) ## Related Articles - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [Coffee with Developers - Maria Apazoglou - Making AI understandable for all in production](https://www.wearedevelopers.com/magazine/475-coffee-with-developers-maria-apazoglou-making-ai-understandable-for-all-in-production) - [Making Data Warehouses Fast: A Developer’s Story](https://www.wearedevelopers.com/magazine/107-making-data-warehouses-fast-a-developer-s-story) - [The Most Popular IT Jobs on the Market](https://www.wearedevelopers.com/magazine/376-the-most-popular-it-jobs-on-the-market) - [Navigating the AI Shift](https://www.wearedevelopers.com/magazine/629-navigating-the-ai-shift)