LLM & Generative AI Engineer
- Discuss this with your agent
- Open in Claude
- Open in ChatGPT
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Role details
Tech stack
Requirements
About the RoleJoin our AI Engineering division in London to specialize in LLM fine-tuning, retrieval-augmented generation (RAG), and hosting private models. You will be responsible for tailoring deep learning models to specialized domain tasks.Key ResponsibilitiesFine-tune open-source models (Llama, Mistral, Qwen) for specific domain functionsOptimize model deployment pipelines for low latency and high throughputBuild advanced context management and semantic search solutionsImplement prompt evaluation frameworks and guardrail architecturesRequirements3+ years of experience focusing on Natural Language Processing and Generative AIHands-on experience with PyTorch, Hugging Face Transformers, and parameter-efficient fine-tuning (PEFT/LoRA)Experience deploying models with vLLM, Ollama, or Triton Inference ServerStrong background in software engineering best practices and clean code #J-18808-Ljbffr
Apply for this position
This job is hosted externally. Click below to view the full posting and apply.
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Good distractions
Talks and stories from around this role — technically off-topic, practically not.
Moments
Explore playlistsVideos
See allRelated articles
See all
The Best Large Language Models on The Market
MLOps And AI Driven Development
How to Become an AI Engineer
Dev Digest 121 - AI goes offline