> Markdown version of [/jobs/ext/3124770-rag-architect-genai-solutions-architect](https://www.wearedevelopers.com/jobs/ext/3124770-rag-architect-genai-solutions-architect). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # RAG Architect / GenAI Solutions Architect - **Company:** VALCAN IT, INC. - **Location:** United States (Remote available) - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Amazon Web Services, Microsoft Azure, Cloud Computing, Encodings, Continuous Integration, Information Engineering, Graph Database, Python (Programming Language), Machine Learning, Query Optimization, Role-Based Access Control, Cloud Services, Search Technologies, Enterprise Data Management, Google Cloud, Large Language Models, Multi-Agent Systems, Prompt Engineering, Generative AI, Indexer, Kubernetes, Machine Learning Operations, Virtual Agents, Restful APIs, Docker, Microservices - **Published:** September 28, 2026 - **Apply:** https://www.dice.com/job-detail/c161c5d3-eb2b-4c63-9fa7-f4218065e456 ## About the Role 8+ years of software/AI engineering experience with strong architecture experience. Strong hands-on experience with RAG and LLM-based applications. Expertise in Python, LLMs, embeddings, prompt engineering, and NLP. Strong knowledge of Vector Databases such as Pinecone, Weaviate, Milvus, pgvector, or OpenSearch. Experience with LangChain, LangGraph, LlamaIndex, or similar frameworks. Strong understanding of semantic search, hybrid search, reranking, chunking, embeddings, and retrieval optimization. Experience with AWS, Azure, or Google Cloud Platform AI/cloud services. Experience designing REST APIs, microservices, and scalable AI platforms. Knowledge of Docker, Kubernetes, CI/CD, and MLOps. Strong understanding of AI security, data privacy, RBAC, and LLM guardrails. Preferred Skills: Experience with Agentic AI / AI Agents. Knowledge of Graph RAG / Knowledge Graphs. Experience with multimodal RAG. Experience with AWS Bedrock, Azure OpenAI, or Google Vertex AI. Experience with RAG evaluation and observability platforms. ## Description We are seeking an experienced RAG Architect to design and lead scalable Retrieval-Augmented Generation (RAG) solutions using enterprise data, LLMs, vector search, and AI orchestration technologies., Design end-to-end RAG architecture for enterprise AI applications. Architect document ingestion, chunking, embedding, indexing, retrieval, and generation pipelines. Design and optimize vector search and semantic retrieval solutions. Integrate LLMs, embedding models, vector databases, and enterprise data sources. Implement advanced retrieval techniques including hybrid search, reranking, metadata filtering, and query optimization. Design RAG solutions using frameworks such as LangChain, LangGraph, LlamaIndex, or equivalent. Establish RAG evaluation frameworks for relevance, accuracy, groundedness, hallucination, and retrieval quality. Implement security, access control, PII protection, guardrails, and responsible AI practices. Design scalable APIs and microservices for production RAG applications. Collaborate with Data Engineering, ML Engineering, Cloud, Security, and Application teams. Lead technical design, architecture reviews, POCs, and production implementation. ## Related Videos - [RAG's Not Dead, You're Just Using It Wrong! - Phil Nash](https://www.wearedevelopers.com/videos/1906-rag-s-not-dead-you-re-just-using-it-wrong-phil-nash) - [Docker Compose: Rediscovered](https://www.wearedevelopers.com/videos/1978-docker-compose-rediscovered) - [Dynamic Entities in .NET: Building Low-Code Systems on Top of Entity Framework Core](https://www.wearedevelopers.com/videos/100218-dynamic-entities-in-net-building-low-code-systems-on-top-of-entity-framework-core) - [A Brief History of Data Storage](https://www.wearedevelopers.com/videos/974-a-brief-history-of-data-storage) - [Accelerating GenAI Development: Harnessing Astra DB Vector Store and Langflow for LLM-Powered Apps](https://www.wearedevelopers.com/videos/966-accelerating-genai-development-harnessing-astra-db-vector-store-and-langflow-for-llm-powered-apps) - [Docker build without Docker](https://www.wearedevelopers.com/videos/100114-docker-build-without-docker) ## Related Articles - [Graph and AI Trends 2026: Why Is AI Running but Not Yet Delivering?](https://www.wearedevelopers.com/magazine/680-graph-and-ai-trends-2026-why-is-ai-running-but-not-yet-delivering) - [Got AI ideas but no money? Here are 10 free ways to level up your AI skills with Google Cloud](https://www.wearedevelopers.com/magazine/600-got-ai-ideas-but-no-money-here-are-10-free-ways-to-level-up-your-ai-skills-with-google-cloud) - [MLOps And AI Driven Development](https://www.wearedevelopers.com/magazine/82-mlops-and-ai-driven-development) - [ I Gave a Video Editor More Autonomy Than a Trading Bot. On Purpose.](https://www.wearedevelopers.com/magazine/773-i-gave-a-video-editor-more-autonomy-than-a-trading-bot-on-purpose) - [What Are Large Language Models?](https://www.wearedevelopers.com/magazine/304-what-are-large-language-models) - [Dev Digest 121 - AI goes offline](https://www.wearedevelopers.com/magazine/456-dev-digest-121-ai-goes-offline)