> Markdown version of [/jobs/ext/1992703-ai-technical-architect](https://www.wearedevelopers.com/jobs/ext/1992703-ai-technical-architect). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # AI Technical Architect - **Company:** The Coca-Cola Company - **Location:** Dallas, TX, United States - **Experience:** Expert - **Salary:** $102,675.0 - $171,125.0 - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Airflow, Amazon Web Services, Audit Trail, Microsoft Azure, Computer Networks, Continuous Integration, Python (Programming Language), Network Architecture, Tensorflow, Cloud Platform System, Feature Engineering, Data Ingestion, Pytorch, Large Language Models, AI Platforms, Kubernetes, Information Technology, Data Lineage, ONNX (Open Neural Network Exchange) Format, Machine Learning Operations, Serverless Computing - **Published:** August 8, 2026 - **Apply:** https://www.careerjet.com/jobad/us375a8677d560601173c32cb2eee9ec5c ## About the Role * Bachelor's or master's degree in computer science, Engineering, Data Science, AI, or a related field. * 15+ years of overall engineering experience, with at least 4+ years in AI/ML solution architecture. * Proven experience designing and deploying AI systems in production at scale (LLM and/or classical ML). * Strong hands-on proficiency in Python and at least one major cloud platform (AWS, Azure, or GCP). Must-Have Technical Skills AI/ML & LLM Architecture * Designing LLM/RAG systems, including retrieval pipelines, chunking strategies, embeddings, reranking, prompt orchestration, response orchestration, evaluation, and safety. * Deep understanding of the model lifecycle, including fine-tuning, PEFT/LoRA, quantization, distillation, latency optimization, and cost optimization. * Strong ML/NLP expertise, including feature engineering, model selection, training, cross-validation, experimentation, and testing. MLOps / LLMOps * CI/CD for ML, including model versioning, model promotion, feature stores, model registry, lineage tracking, and drift detection. * Inference stacks including PyTorch, TensorFlow, vLLM, TGI, ONNX, GPU orchestration, autoscaling, and APM. * Pipelines and orchestration frameworks such as Airflow, Kubeflow, and MLflow. ## Description * Define AI/ML reference architecture and solution blueprints (batch/streaming ML, LLM + RAG, multimodal). * Lead end-to-end solution design across data ingestion, model training, inference, deployment, and monitoring. * Architect LLM applications (agents, summarization, classification) with RAG, evaluation frameworks, safety controls, and guardrails. * Own MLOps/LLMOps practices, including CI/CD for models, model registry, feature stores, lineage tracking, observability, drift detection, and cost monitoring. * Choose the right cloud and runtime strategy (managed services vs. self-hosted, GPU vs. CPU, serverless vs. containerized). * Establish AI governance standards, including PII handling, encryption, auditability, and Responsible AI practices. * Collaborate with product and business stakeholders to translate requirements into architectural decisions and delivery plans. * Perform technical spikes and POCs, benchmark models and infrastructure, and lead Architecture Reviews. * Create and maintain standards, patterns, and reusable components; mentor engineers across teams. * Drive performance and cost optimization initiatives, including throughput, latency, SLA/SLO management, caching, quantization/distillation, and autoscaling. * Support vendor and product evaluations, including cloud AI services, vector databases, orchestration frameworks, and monitoring platforms., This position is 100% on-site in Allen, Texas. The Enterprise Network Architect is responsible for designing, supporting, managing, and maintaining the network infrastructure. Co… + 14 hours ago ## Related Videos - [DevOps for AI: running LLMs in production with Kubernetes and KubeFlow](https://www.wearedevelopers.com/videos/1222-devops-for-ai-running-llms-in-production-with-kubernetes-and-kubeflow) - [Understanding Kubernetes in a visual way](https://www.wearedevelopers.com/videos/100085-understanding-kubernetes-in-a-visual-way) - [From Messy Queries to Scalable Systems - How Data Engineering actually works](https://www.wearedevelopers.com/videos/100203-from-messy-queries-to-scalable-systems-how-data-engineering-actually-works) - [Photonic Computing: Programming a New Class of AI Accelerators (incl. Live Coding)](https://www.wearedevelopers.com/videos/100196-photonic-computing-programming-a-new-class-of-ai-accelerators-incl-live-coding) - [AI Model Management Life Circles: ML Ops For Generative AI Models From Research to Deployment](https://www.wearedevelopers.com/videos/1152-ai-model-management-life-circles-ml-ops-for-generative-ai-models-from-research-to-deployment) - [LLMOps-driven fine-tuning, evaluation, and inference with NVIDIA NIM & NeMo Microservices](https://www.wearedevelopers.com/videos/1582-llmops-driven-fine-tuning-evaluation-and-inference-with-nvidia-nim-nemo-microservices) ## Related Articles - [MLOps – What’s the deal behind it?](https://www.wearedevelopers.com/magazine/125-mlops-what-s-the-deal-behind-it) - [MLOps And AI Driven Development](https://www.wearedevelopers.com/magazine/82-mlops-and-ai-driven-development) - [What Are Large Language Models?](https://www.wearedevelopers.com/magazine/304-what-are-large-language-models) - [MLops – Deploying, Maintaining And Evolving Machine Learning Models in Production](https://www.wearedevelopers.com/magazine/115-mlops-deploying-maintaining-and-evolving-machine-learning-models-in-production) - [Coffee with Developers - Maria Apazoglou - Making AI understandable for all in production](https://www.wearedevelopers.com/magazine/475-coffee-with-developers-maria-apazoglou-making-ai-understandable-for-all-in-production) - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer)