GenAI Architect (Onsite)

Hexaware Technologies
Mount Laurel Township, NJ, United States
9 days ago
Apply on www.dice.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Experienced
Experience required
3 years minimum
Working hours
Regular working hours
Job source

Tech stack

LangGraph Framework Artificial Intelligence Microsoft Azure Cloud Engineering Python (Programming Language) Performance Tuning Software Architecture Tensorflow Azure Machine Learning Pinecone Pytorch LangChain
+20 more
Retrieval-Augmented Generation Transfer Learning Large Language Models Prompt Engineering Generative AI Event Driven Architecture Microsoft Copilot Studio Containerization AI Platforms Kubernetes HuggingFace Transformer Architectures Weaviate Machine Learning Operations FAISS Claude Api Design GPT Docker Microservices

Job description

  • Lead the design and development of scalable GenAI solutions leveraging LLMs, diffusion models, and multimodal architectures.
  • Architect end-to-end pipelines involving prompt engineering, vector databases, retrieval-augmented generation (RAG), and LLM fine-tuning.
  • Select and integrate foundational models (e.g., GPT, Claude, LLaMA, Mistral) based on business needs and technical constraints.
  • Lead architectural reviews and provide technical guidance to development teams
  • Design integration patterns between AI services and existing enterprise systems
  • Create technical roadmaps aligned with business objectives and AI strategy
  • Assess technical feasibility of AI use cases and provide effort estimations

Requirements

  • 8+ years of experience in software architecture with 3+ years focused on AI/ML systems
  • Hands-on experience with LLMs, transformers, fine-tuning techniques (LoRA, PEFT), and prompt engineering.
  • Proficient in Python, with libraries/frameworks such as Hugging Face Transformers, LangChain, OpenAI API, PyTorch, TensorFlow.
  • Experience with vector databases (e.g., Pinecone, FAISS, Weaviate) and RAG pipelines.
  • Strong understanding of cloud-native AI architectures (Azure & DBX), containerization (Docker/Kubernetes), and API integration.
  • Deep understanding of LLM architectures, transformer models, and generative AI patterns
  • Experience with MLOps tools and practices (Azure ML, Docker, Kubernetes)
  • Proficiency in Python and familiarity with AI/ML frameworks (PyTorch, TensorFlow, Hugging Face)
  • Experience with Azure AI Foundry, LangChain, LangGraph, and Copilot Studio
  • Strong technical leadership and mentoring capabilities
  • Experience with microservices architecture, API design, and event-driven systems

About the company

Hexaware is a dynamic and innovative IT organization committed to delivering cutting-edge solutions to our clients worldwide. We pride ourselves on fostering a collaborative and inclusive work environment where every team member is valued and empowered to succeed.

Hexaware provides access to a vast array of tools that enhance, revolutionize, and advance professional profiles. We complete the circle with excellent growth opportunities, chances to collaborate with high-profile customers, opportunities to work alongside brilliant minds, and the perfect work-life balance.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.dice.com
Prepare application

Good distractions

Loading talks and stories from around this role…