> Markdown version of [/jobs/ext/2709147-ai-engineer-llm](https://www.wearedevelopers.com/jobs/ext/2709147-ai-engineer-llm). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # AI Engineer - LLM - **Company:** BRIGHTAI CORPORATION - **Location:** Palo Alto, United States - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Document Retrieval, Python (Programming Language), Machine Learning, Open Source Technology, Search Technologies, Reinforcement Learning, Pytorch, Delivery Pipeline, Large Language Models, Prompt Engineering, Deep Learning, Information Technology, Low Latency, HuggingFace, Process Control Systems, GPT - **Published:** September 4, 2026 - **Apply:** https://startup.jobs/senior-ai-engineer-llm-rag-brightai-corporation-7092836 ## About the Role * M.S. or Ph.D. in Computer Science, AI, Machine Learning, or a related field, with specialization in NLP or deep learning. * Strong research or applied background in large language models (LLMs) and retrieval-augmented generation (RAG) systems. Agentic RAG experience is highly desirable., * 5+ years of experience in machine learning or AI with a strong focus on NLP, LLMs, or conversational AI. * Fluency with modern LLMs and open-source foundational models (e.g., LLAMA, Falcon, Mistral, GPT, Claude). * Experience building RAG pipelines with tools like LangChain, LlamaIndex, or custom vector database integrations, with at least one production grade system was built. * Fluency with prompt engineering, instruction tuning, or fine-tuning open-source models. * Deep understanding of document retrieval (semantic search, embedding generation, similarity metrics) and vector stores (e.g., FAISS, Weaviate, Pinecone). * Strong foundation in core machine learning techniques, including experience with reinforcement learning (RL) or decision-making models. * Proficiency with ML development frameworks such as PyTorch, Hugging Face Transformers, or similar. Strong Python programming is a must. * Experience integrating AI systems into real-world applications with user-facing interfaces and operational constraints. * Excellent problem-solving and critical thinking skills; ability to design solutions for complex, ambiguous problems. * Strong written and verbal communication skills, with ability to collaborate cross-functionally with engineers, product managers, and domain experts., * Experience applying LLMs in industrial or physical infrastructure settings (e.g., manufacturing, logistics, utilities, energy). * Knowledge of industrial control systems, maintenance workflows, or technician support processes. * Exposure to multimodal models or integrating textual data with sensor and/or time-series data. * Prior experience in a startup or a fast-paced environment building LLM-powered products from the ground up. ## Description * Lead the architecture and development of RAG systems that combine LLMs (e.g., LLAMA, Mistral, Claude, GPT) with structured and unstructured external information sources. * Develop AI-powered assistants to support technicians in diagnosing and resolving anomalies or failures in factory, plant, or industrial settings. * Build pipelines to ingest, preprocess, and index large corpora of documents (manuals, logs, notes, procedures) for semantic search and grounding. * Customize and fine-tune foundational models to incorporate domain-specific language, tone, and logic for industrial troubleshooting scenarios. * Collaborate with product, data, and cloud teams to design scalable, privacy-compliant, and latency-sensitive LLM applications. * Design evaluation strategies to measure performance, accuracy, and user experience of RAG-enabled systems in production settings. * Stay up to date with the latest advances in LLM architectures, retrieval methods, and prompt engineering, and integrate emerging techniques into the product roadmap. ## Related Videos - [How to Avoid LLM Pitfalls - Mete Atamel and Guillaume Laforge](https://www.wearedevelopers.com/videos/1328-how-to-avoid-llm-pitfalls-mete-atamel-and-guillaume-laforge) - [Photonic Computing: Programming a New Class of AI Accelerators (incl. Live Coding)](https://www.wearedevelopers.com/videos/100196-photonic-computing-programming-a-new-class-of-ai-accelerators-incl-live-coding) - [ Evaluating AI models for code comprehension](https://www.wearedevelopers.com/videos/1462-evaluating-ai-models-for-code-comprehension) - [Coffee with Developers - Maria Apazoglou](https://www.wearedevelopers.com/videos/1209-coffee-with-developers-maria-apazoglou) - [Serverless deployment of (large) NLP models ](https://www.wearedevelopers.com/videos/158-serverless-deployment-of-large-nlp-models) - [Streaming AI Responses in Real-Time with SSE in Next.js & NestJS](https://www.wearedevelopers.com/videos/1630-streaming-ai-responses-in-real-time-with-sse-in-next-js-nestjs) ## Related Articles - [What Are Large Language Models?](https://www.wearedevelopers.com/magazine/304-what-are-large-language-models) - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [MLOps And AI Driven Development](https://www.wearedevelopers.com/magazine/82-mlops-and-ai-driven-development) - [13 AI Tools You Have to Try](https://www.wearedevelopers.com/magazine/219-13-ai-tools-you-have-to-try) - [Graph and AI Trends 2026: Why Is AI Running but Not Yet Delivering?](https://www.wearedevelopers.com/magazine/680-graph-and-ai-trends-2026-why-is-ai-running-but-not-yet-delivering) - [Got AI ideas but no money? Here are 10 free ways to level up your AI skills with Google Cloud](https://www.wearedevelopers.com/magazine/600-got-ai-ideas-but-no-money-here-are-10-free-ways-to-level-up-your-ai-skills-with-google-cloud)