> Markdown version of [/jobs/ext/1458787-senior-ai-engineer](https://www.wearedevelopers.com/jobs/ext/1458787-senior-ai-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Senior AI Engineer - **Company:** On behalf of Next Deavor - **Location:** United States (Remote available) - **Experience:** Expert - **Salary:** $150,000.0 - $220,000.0 - **Contract:** Permanent contract - **Skills:** Application Programming Interfaces (APIs), Artificial Intelligence, Software System Penetration Testing, Code Review, Decision Support Systems, Mobile Application Software, Python (Programming Language), Open Source Technology, Open Web Application Security, Red Team (Cyber Security), Reverse Engineering, Pytorch, Large Language Models, Multi-Agent Systems, Kubernetes, Low Latency, HuggingFace, TensorRT, Data Generation - **Published:** July 27, 2026 - **Apply:** https://www.indeed.com/viewjob?jk=22656f8f77b7f720 ## About the Role * 5+ years building production ML/AI systems, including at least 2 years working directly on LLMs or LLM-powered agents * Deep Python and strong production engineering practices (testing, code review, observability) * Hands-on fine-tuning experience: SFT, preference optimization (DPO, GRPO, RLHF/RLAIF), data curation, and synthetic data generation * Strong grasp of transformer architectures and training stack (PyTorch, Hugging Face, DeepSpeed or FSDP, accelerate) * Experience designing and shipping multi-agent or tool-using LLM systems in production * Rigorous evaluation design experience: building harnesses, tracking experiments, and data-driven decision making * Inference optimization experience (vLLM, TensorRT-LLM, quantization, throughput/latency tradeoffs) * Familiarity with retrieval pipelines, vector stores, and structured memory for agents * Kubernetes and containerized deployment fluency * Genuine interest in offensive security and the ability to ramp quickly on OWASP Top 10, API/web/mobile pentesting concepts Here's What Else Might Help You Out * Offensive security certifications or experience (OSCP/OSWE/OSWA, CTF, bug bounty, red team) * Research publications at top ML/security venues or open source contributions to agent/LLM tooling * Experience with adversarial ML or red-teaming AI systems * Familiarity with mobile app reverse engineering or binary analysis ## Description * Design, implement, and iterate on named agents, including orchestration patterns, hand-offs, planning loops, tool use, and shared memory * Contribute to model training and fine-tuning across data curation, supervised fine-tuning (SFT), preference optimization (DPO/GRPO/RLHF-style), and evaluation * Extend the co-evolutionary self-training (Javelin) loop so the system improves from its engagements * Build self-improvement systems (false-positive detection, tiered skill learning, agent directives, code-patch proposals) and pipelines for human approval * Design security-specific evaluations covering OWASP Top 10, exploit chaining, finding accuracy, and agent reliability; track performance over model and agent changes * Contribute to multimodal (vision) and mobile (iOS/Android) coverage and BYOK support efforts * Own production reliability: latency, cost, observability, failure-mode analysis, and Kubernetes-based deployment * Improve customer-facing accuracy surfaces and live accuracy gauges exposed to customers ## Related Videos - [Tour de Force: Open-Source LLM Inference Optimization from Simple to Sophisticated](https://www.wearedevelopers.com/videos/100099-tour-de-force-open-source-llm-inference-optimization-from-simple-to-sophisticated) - [Understanding Kubernetes in a visual way](https://www.wearedevelopers.com/videos/100085-understanding-kubernetes-in-a-visual-way) - [Photonic Computing: Programming a New Class of AI Accelerators (incl. Live Coding)](https://www.wearedevelopers.com/videos/100196-photonic-computing-programming-a-new-class-of-ai-accelerators-incl-live-coding) - [How AI Models Get Smarter](https://www.wearedevelopers.com/videos/1374-how-ai-models-get-smarter) - [Instant KAI Sandboxes with vCluster: Multi-Tenant, Multi-Scheduler GPU Sharing](https://www.wearedevelopers.com/videos/100333-instant-kai-sandboxes-with-vcluster-multi-tenant-multi-scheduler-gpu-sharing) - [Efficient deployment and inference of GPU-accelerated LLMs​](https://www.wearedevelopers.com/videos/929-efficient-deployment-and-inference-of-gpu-accelerated-llms) ## Related Articles - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [MLOps And AI Driven Development](https://www.wearedevelopers.com/magazine/82-mlops-and-ai-driven-development) - [What Are Large Language Models?](https://www.wearedevelopers.com/magazine/304-what-are-large-language-models) - [From Prototype to Production: Build AI Agents with This Free 4-Course Learning Path](https://www.wearedevelopers.com/magazine/655-from-prototype-to-production-build-ai-agents-with-this-free-4-course-learning-path) - [Dev Digest 137 - AI'm not sure about this](https://www.wearedevelopers.com/magazine/485-dev-digest-137-ai-m-not-sure-about-this) - [MLOps – What’s the deal behind it?](https://www.wearedevelopers.com/magazine/125-mlops-what-s-the-deal-behind-it)