> Markdown version of [/jobs/ext/442775-data-engineer](https://www.wearedevelopers.com/jobs/ext/442775-data-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Data Engineer - **Company:** Nuvitek LLC - **Location:** United States (Remote available) - **Experience:** Experienced - **Salary:** $115,000.0 - $125,000.0 - **Contract:** Permanent contract - **Skills:** Application Programming Interfaces (APIs), Artificial Intelligence, Software Applications, Automated Storage and Retrieval Systems, Data Architecture, Information Engineering, Data Infrastructure, Extract Transform Load (ETL), Data Security, Distributed Computing Environment, Python (Programming Language), Performance Tuning, Cloud Services, Zero Trust Network Access, Search Technologies, Systems Integration, Unstructured Data, AI Infrastructure, Data Logging, Data Processing, Cloud Platform System, Large Language Models, Generative AI, AI Platforms, Machine Learning Operations, Data Pipelines, Docker - **Published:** June 4, 2026 - **Apply:** https://www.indeed.com/viewjob?jk=049035500ed22131 ## About the Role Do you have experience in Technical Proficiency?, The ideal candidate has hands-on experience with retrieval-augmented generation (RAG), contextual augmentation generation (CAG), OCR processing, vector databases, and modern AI data architectures. This role requires strong technical expertise, problem-solving skills, and the ability to work collaboratively within agile pod-based teams., * 4+ years of experience in data engineering, data platform development, or AI/ML infrastructure * Strong experience building RAG and/or CAG pipelines * Hands-on experience with vector databases and semantic retrieval systems * Experience developing document ingestion and OCR processing workflows * Strong understanding of LLM integrations and AI data pipeline architectures * Experience working with structured, semi-structured, and unstructured datasets * Proficiency with Python and modern data engineering frameworks * Familiarity with APIs, ETL/ELT pipelines, and distributed processing systems * Experience building and operating data pipelines in secure federal cloud environments, including FedRAMP Moderate and Zero Trust architectures, with appropriate handling of sensitive data and Controlled Unclassified Information (CUI) * Ability to obtain and maintain a federal Public Trust (or higher) clearance * Strong analytical, troubleshooting, and performance optimization skills * Ability to work effectively in agile or pod-based delivery environments * Excellent communication and collaboration skills * Experience working with historical archives or large-scale document digitization efforts * Familiarity with cloud-native data platforms and AI infrastructure * Experience with search relevance tuning and ranking optimization * Knowledge of embedding models, chunking strategies, and retrieval optimization techniques * Experience with containerization and orchestration technologies such as Docker and Kubernetes * Familiarity with accessibility, governance, and secure data handling practices * Passion for building scalable AI-driven solutions that improve user experiences and operational efficiency ## Description Nüvitek is seeking a highly skilled Data Engineer to support the design, development, and optimization of advanced AI and data processing solutions. This role will focus on building scalable data pipelines that power large language model (LLM) applications, including retrieval systems, document ingestion workflows, and intelligent search capabilities., * Design, develop, and maintain scalable RAG/CAG pipelines for AI-powered applications * Build and optimize document ingestion workflows for structured and unstructured data sources * Manage and maintain vector stores to support semantic search and retrieval capabilities * Develop OCR processing pipelines for historical and modern document collections spanning 1781-2025 * Optimize retrieval performance, relevance tuning, and ranking strategies for LLM-based systems * Build reliable data pipelines that support integrations with large language models and AI services * Collaborate with engineers, UX teams, product owners, and stakeholders to deliver scalable AI solutions * Ensure data quality, integrity, security, and performance across ingestion and retrieval systems * Implement monitoring, logging, and troubleshooting for AI and data processing workflows * Contribute to architecture decisions, technical documentation, and engineering best practices * Participate in agile pod-based development teams and continuous improvement initiatives ## Related Videos - [Docker Compose: Rediscovered](https://www.wearedevelopers.com/videos/1978-docker-compose-rediscovered) - [Crypto-secure Data Management with In-Database Blockchain](https://www.wearedevelopers.com/videos/632-crypto-secure-data-management-with-in-database-blockchain) - [This App Reached 10,000 Users in One Week. Here's How.](https://www.wearedevelopers.com/videos/100329-this-app-reached-10-000-users-in-one-week-here-s-how) - [Data Science, ML & AI in the Oil and Gas Industry at NDT Global - Dr. Katja Träumner](https://www.wearedevelopers.com/videos/1308-data-science-ml-ai-in-the-oil-and-gas-industry-at-ndt-global-dr-katja-traumner) - [Developer Experience, Platform Engineering and AI powered Apps](https://www.wearedevelopers.com/videos/990-developer-experience-platform-engineering-and-ai-powered-apps) - [Supercharge your cloud-native applications with Generative AI](https://www.wearedevelopers.com/videos/950-supercharge-your-cloud-native-applications-with-generative-ai) ## Related Articles - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [MLOps And AI Driven Development](https://www.wearedevelopers.com/magazine/82-mlops-and-ai-driven-development) - [Dev Digest 121 - AI goes offline](https://www.wearedevelopers.com/magazine/456-dev-digest-121-ai-goes-offline) - [Navigating the AI Shift](https://www.wearedevelopers.com/magazine/629-navigating-the-ai-shift) - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk)