GEN AI Engineer

EXL SERVICE
United States
about 1 month ago
Apply on www.indeed.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
3 years minimum
Compensation
$90,000.0 - $125,000.0
Working hours
Regular working hours
Job source

Tech stack

Application Programming Interfaces (APIs) Agile Methodology Artificial Intelligence Amazon Web Services Microsoft Azure Cloud Computing Code Review Continuous Integration Python (Programming Language) Scrum Methodology Software Engineering Management of Software Versions
+9 more
Software Organization Google Cloud Retrieval-Augmented Generation Large Language Models Model Validation Generative AI Information Technology Machine Learning Operations GPT

Job description

Job Description: Generative AI Engineer is a hands-on technical role focused on building, deploying, and optimizing production-grade generative AI solutions. This role works closely with AI Architects, data scientists, and platform teams to translate business problems into scalable AI-driven applications using large language models and modern GenAI frameworks. The Senior AI Engineer is responsible for designing and implementing components such as prompt workflows, retrieval-augmented generation pipelines, vector search integrations, and lightweight agent-based systems, while ensuring performance, reliability, and cost efficiency in production environments. The role emphasizes strong software engineering practices, cloud-native development, and LLMOps fundamentals, including testing, monitoring, and CI/CD integration. Candidates are expected to contribute ideas, improve existing architectures, and mentor junior engineers, while operating within established architectural guardrails. This position is ideal for a practitioner who enjoys deep technical problem-solving, rapid experimentation, and delivering measurable business impact through practical, scalable generative AI solutions.

  • Responsibilities: Design, develop, and deploy production-grade Generative AI solutions for business use cases
  • Build and maintain LLM-based applications using prompts, chains, agents, and RAG pipelines
  • Implement AI components under defined solution and architectural guidelines
  • Integrate LLMs with enterprise data sources, vector databases, and APIs
  • Optimize model accuracy, latency, scalability, and cost in production environments
  • Collaborate closely with AI Architects, data scientists, and platform engineering teams
  • Apply strong software engineering practices including code reviews, testing, and documentation
  • Support LLMOps activities such as model evaluation, monitoring, versioning, and CI/CD integration
  • Participate in Agile delivery ceremonies including sprint planning, demos, and retrospectives

Requirements

  • Qualifications: Bachelor’s degree in Computer Science, Artificial Intelligence, Engineering, or a related field
  • Master’s degree in AI, Data Science, or a related discipline is a plus
  • 3-6 years of experience in software engineering, AI/ML, or data science roles
  • 3+ years of hands-on experience building Generative AI or NLP-based solutions
  • Demonstrated experience deploying AI / LLM-based applications into production
  • Strong proficiency in Python and modern software development practices
  • Practical experience working with LLM frameworks, vector databases, and RAG-based architectures
  • Experience working on cloud platforms such as AWS, Azure, GCP, or Nvidia ecosystems
  • Familiarity with Agile development methodologies and collaborative delivery modelsStrong problem-solving skills with the ability to translate business requirements into technical solutions

Benefits & conditions

3.73.7 out of 5 stars Florida Hybrid work $90,000 - $125,000 a year - Full-time

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.indeed.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

40 sec

Generative pre-trained transformer models powering code completions

lgonta lgonta +1 · World Congress 2024

3:39 min

Addressing code review surrender and process exploitation

Laura Tacho Laura Tacho · World Congress 2026 Europe

6:10 min

Unlocking free learning credits via Google Cloud Innovators

Asrar Asrar · World Congress 2024

2:08 min

Essential engineering roles in the generative AI space

Mary Grygleski Mary Grygleski · LIVE

2:52 min

Scaling generative AI use cases across large enterprises

Mike Butcher Mike Butcher +3 · World Congress 2024

51 sec

Assessing GPT-4o performance for pull request feedback

Merrill Lutsky Merrill Lutsky · World Congress 2025

Videos

See all

Related articles

See all