GenAI Architect (Onsite)
Hexaware Technologies
Mount Laurel Township, NJ, United States
9 days ago
- Discuss this with your agent
- Open in Claude
- Open in ChatGPT
Apply on www.dice.com
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Role details
Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Experienced
Experience required
3 years minimum
Working hours
Regular working hours
Job source
Tech stack
LangGraph Framework
Artificial Intelligence
Microsoft Azure
Cloud Engineering
Python (Programming Language)
Performance Tuning
Software Architecture
Tensorflow
Azure Machine Learning
Pinecone
Pytorch
LangChain
+20 more
Retrieval-Augmented Generation
Transfer Learning
Large Language Models
Prompt Engineering
Generative AI
Event Driven Architecture
Microsoft Copilot Studio
Containerization
AI Platforms
Kubernetes
HuggingFace
Transformer Architectures
Weaviate
Machine Learning Operations
FAISS
Claude
Api Design
GPT
Docker
Microservices
Job description
- Lead the design and development of scalable GenAI solutions leveraging LLMs, diffusion models, and multimodal architectures.
- Architect end-to-end pipelines involving prompt engineering, vector databases, retrieval-augmented generation (RAG), and LLM fine-tuning.
- Select and integrate foundational models (e.g., GPT, Claude, LLaMA, Mistral) based on business needs and technical constraints.
- Lead architectural reviews and provide technical guidance to development teams
- Design integration patterns between AI services and existing enterprise systems
- Create technical roadmaps aligned with business objectives and AI strategy
- Assess technical feasibility of AI use cases and provide effort estimations
Requirements
- 8+ years of experience in software architecture with 3+ years focused on AI/ML systems
- Hands-on experience with LLMs, transformers, fine-tuning techniques (LoRA, PEFT), and prompt engineering.
- Proficient in Python, with libraries/frameworks such as Hugging Face Transformers, LangChain, OpenAI API, PyTorch, TensorFlow.
- Experience with vector databases (e.g., Pinecone, FAISS, Weaviate) and RAG pipelines.
- Strong understanding of cloud-native AI architectures (Azure & DBX), containerization (Docker/Kubernetes), and API integration.
- Deep understanding of LLM architectures, transformer models, and generative AI patterns
- Experience with MLOps tools and practices (Azure ML, Docker, Kubernetes)
- Proficiency in Python and familiarity with AI/ML frameworks (PyTorch, TensorFlow, Hugging Face)
- Experience with Azure AI Foundry, LangChain, LangGraph, and Copilot Studio
- Strong technical leadership and mentoring capabilities
- Experience with microservices architecture, API design, and event-driven systems
About the company
Hexaware is a dynamic and innovative IT organization committed to delivering cutting-edge solutions to our clients worldwide. We pride ourselves on fostering a collaborative and inclusive work environment where every team member is valued and empowered to succeed.
Hexaware provides access to a vast array of tools that enhance, revolutionize, and advance professional profiles. We complete the circle with excellent growth opportunities, chances to collaborate with high-profile customers, opportunities to work alongside brilliant minds, and the perfect work-life balance.
Apply for this position
This job is hosted externally. Click below to view the full posting and apply.
Apply on www.dice.com
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Good distractions
Loading talks and stories from around this roleβ¦