> Markdown version of [/jobs/ext/2100116-llm-generative-ai-engineer](https://www.wearedevelopers.com/jobs/ext/2100116-llm-generative-ai-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # LLM & Generative AI Engineer - **Company:** Rolejoin - **Location:** London, UK - **Contract:** Permanent contract - **Skills:** Clean Code Principles, Natural Language Processing, Open Source Technology, Search Technologies, Software Construction, Pytorch, Large Language Models, Deep Learning, Generative AI, Low Latency, HuggingFace, Machine Learning Operations - **Published:** August 18, 2026 - **Apply:** https://www.apply4u.co.uk/jobs/x/44291527/ ## About the Role About the RoleJoin our AI Engineering division in London to specialize in LLM fine-tuning, retrieval-augmented generation (RAG), and hosting private models. You will be responsible for tailoring deep learning models to specialized domain tasks.Key ResponsibilitiesFine-tune open-source models (Llama, Mistral, Qwen) for specific domain functionsOptimize model deployment pipelines for low latency and high throughputBuild advanced context management and semantic search solutionsImplement prompt evaluation frameworks and guardrail architecturesRequirements3+ years of experience focusing on Natural Language Processing and Generative AIHands-on experience with PyTorch, Hugging Face Transformers, and parameter-efficient fine-tuning (PEFT/LoRA)Experience deploying models with vLLM, Ollama, or Triton Inference ServerStrong background in software engineering best practices and clean code #J-18808-Ljbffr ## Related Videos - [How to Avoid LLM Pitfalls - Mete Atamel and Guillaume Laforge](https://www.wearedevelopers.com/videos/1328-how-to-avoid-llm-pitfalls-mete-atamel-and-guillaume-laforge) - [Swapping Low Latency Data Storage Under High Load](https://www.wearedevelopers.com/videos/746-swapping-low-latency-data-storage-under-high-load) - [Photonic Computing: Programming a New Class of AI Accelerators (incl. Live Coding)](https://www.wearedevelopers.com/videos/100196-photonic-computing-programming-a-new-class-of-ai-accelerators-incl-live-coding) - [LLMs in the wild: Building an AI agent that survives production](https://www.wearedevelopers.com/videos/100319-llms-in-the-wild-building-an-ai-agent-that-survives-production) - [Unleash the power of 5G in your code: transform your apps](https://www.wearedevelopers.com/videos/1567-unleash-the-power-of-5g-in-your-code-transform-your-apps) - [Serverless deployment of (large) NLP models ](https://www.wearedevelopers.com/videos/158-serverless-deployment-of-large-nlp-models) ## Related Articles - [What Are Large Language Models?](https://www.wearedevelopers.com/magazine/304-what-are-large-language-models) - [The Best Large Language Models on The Market](https://www.wearedevelopers.com/magazine/319-the-best-large-language-models-on-the-market) - [MLOps And AI Driven Development](https://www.wearedevelopers.com/magazine/82-mlops-and-ai-driven-development) - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [Dev Digest 121 - AI goes offline](https://www.wearedevelopers.com/magazine/456-dev-digest-121-ai-goes-offline) - [Got AI ideas but no money? Here are 10 free ways to level up your AI skills with Google Cloud](https://www.wearedevelopers.com/magazine/600-got-ai-ideas-but-no-money-here-are-10-free-ways-to-level-up-your-ai-skills-with-google-cloud)