> Markdown version of [/jobs/ext/2490076-gen-ai-engineer](https://www.wearedevelopers.com/jobs/ext/2490076-gen-ai-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Gen AI Engineer - **Company:** Elevance Health - **Location:** Nashville, TN, United States - **Experience:** Experienced - **Contract:** Permanent contract - **Skills:** Agile Methodology, Artificial Intelligence, Amazon Web Services, Computer Vision, Big Data, Cloud Computing, Configuration Management, Program Optimization, Data Governance, Data Mining, Python (Programming Language), Machine Learning, Natural Language Processing, Tensorflow, SQL Databases, Unstructured Data, Data Ingestion, Pytorch, Large Language Models, Apache Spark, Caching, Generative AI, Keras, Fastapi, Containerization, Information Technology, HuggingFace, Data Analytics, Azure AKS, Machine Learning Operations, Stream Processing, Artificial Intelligence Markup Language (AIML), AWS EKS - **Published:** August 1, 2026 - **Apply:** https://www.nashvillejobsite.com/job.asp?id=3338252436&tx=JT989TYI&pt=1&aff=0B19D771-A501-4A5E-8338-2A822B784D54&utm_source=Job%20Feed&utm_medium=textkernel&utm_campaign=DE&utm_term=0B19D771-A501-4A5E-8338-2A822B784D54 ## About the Role * Requires a Bachelor's degree in a highly quantitative field (Computer Science, Machine Learning, Operational Research, Statistics, Mathematics, etc.) or equivalent degree and 4 or more years of experience; or any combination of education and experience in configuration management, which would provide an equivalent background. Preferred Skills, Capabilities, and Experiences: * Advanced Python proficiency. * 4+ years of professional hands-on experience leveraging large sets of structured and unstructured data to develop data-driven tactical and strategic analytics and insights using ML, NLP, and computer vision solutions. * Demonstrated 4+ years hands-on experience with Python, SQL, Hugging Face, TensorFlow, Keras, PyTorch, and Spark. * Experience with GCP/AWS cloud platforms. * Strong knowledge of and measurable hands-on experience with developing or tuning Large Language Models (LLM) and Generative AI (GAI) * Experience with NLP, LLMs (extractive and generative), fine-tuning and LLM model development. * Experience developing and optimizing high-quality prompts for NLP applications. * Excellent written & verbal communication and stakeholder management skills. * 4+ years project leadership experience including Agile project management, Scaled Agile Frameworks (SAFE). * LLM Infrastructure & Deployment: LLM serving platforms (vLLM, Text Generation Inference, FastAPI); Model quantization for LLMs (GPTQ, AWQ, bitsandbytes); GPU memory optimization techniques (tensor parallelism, pipeline parallelism); LLM caching strategies for inference optimization; RAG architecture design and implementation. * Advanced cloud infrastructure (AWS EKS/ECS, GCP GKE, Azure AKS) knowledge. * Containerization strategies for ML workloads; Canary deployments for ML models. ## Description Location: This role requires associates to be in-office 1 - 2 days per week, fostering collaboration and connectivity, while providing flexibility to support productivity and work-life balance. This approach combines structured office engagement with the autonomy of virtual work, promoting a dynamic and adaptable workplace. Alternate locations may be considered if candidates reside within a commuting distance from an office. Please note that per our policy on hybrid/virtual work, candidates not within a reasonable commuting distance from the posting location(s) will not be considered for employment, unless an accommodation is granted as required by law. PLEASE NOTE: This position is not eligible for current or future visa sponsorship. The Gen AI Engineer is responsible for analyzing and modeling organizational data for the Artificial Intelligence (AI) function to draw business insights, which can be used to make business decisions. How You Will Make an Impact: * Applies data extraction, transformation and loading techniques in order to connect large data sets from a variety of sources. * LLM development and fine-tuning strategies, best practices, and standards to enhance AI ML model deployment and monitoring efficiency. * Develop roadmap and strategy for NLP, LLM, Gen AI model development and lifecycle implementation. * Responsible for the design and development of custom ML, Gen AI, NLP, LLM Models for batch and stream processing-based AI ML pipelines including data ingestion, preprocessing modules, search and retrieval, Retrieval Augmented Generation (RAG), NLP/LLM model development and ensure the end-to-end solution meets all technical and business requirements, and SLA specifications. * Work closely with the MLOps team to create and maintain robust evaluation solutions and tools to evaluate model performance, accuracy, consistency, reliability, during development, and UAT. * Identify and implement model optimizations to improve system efficiency. * Collaborate closely with the MLOps, product teams, business stakeholders, machine learning engineers, and software engineers for the deployment of machine learning models into production environments, ensuring smooth integration, reliability and scalability. * Ensure the use of standards, governance and best practices in ML model development, and adherence to model and data governance standards. ## Related Videos - [How E.On productionizes its AI model & Implementation of Secure Generative AI.](https://www.wearedevelopers.com/videos/623-how-e-on-productionizes-its-ai-model-implementation-of-secure-generative-ai) - [Intro to FastAPI](https://www.wearedevelopers.com/videos/462-intro-to-fastapi) - [Hosting a modern justice system](https://www.wearedevelopers.com/videos/332-hosting-a-modern-justice-system) - [Machine learning in the browser with TensorFlowjs](https://www.wearedevelopers.com/videos/155-machine-learning-in-the-browser-with-tensorflowjs) - [How to Avoid LLM Pitfalls - Mete Atamel and Guillaume Laforge](https://www.wearedevelopers.com/videos/1328-how-to-avoid-llm-pitfalls-mete-atamel-and-guillaume-laforge) - [Building and Deploying Multi-Agent Systems with ADK and Vertex AI](https://www.wearedevelopers.com/videos/1918-building-and-deploying-multi-agent-systems-with-adk-and-vertex-ai) ## Related Articles - [What Are Large Language Models?](https://www.wearedevelopers.com/magazine/304-what-are-large-language-models) - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [Got AI ideas but no money? Here are 10 free ways to level up your AI skills with Google Cloud](https://www.wearedevelopers.com/magazine/600-got-ai-ideas-but-no-money-here-are-10-free-ways-to-level-up-your-ai-skills-with-google-cloud) - [MLOps And AI Driven Development](https://www.wearedevelopers.com/magazine/82-mlops-and-ai-driven-development) - [MLOps – What’s the deal behind it?](https://www.wearedevelopers.com/magazine/125-mlops-what-s-the-deal-behind-it) - [Graph and AI Trends 2026: Why Is AI Running but Not Yet Delivering?](https://www.wearedevelopers.com/magazine/680-graph-and-ai-trends-2026-why-is-ai-running-but-not-yet-delivering)