> Markdown version of [/jobs/ext/1762822-software-engineer-ii-ai-engineer-w-skills-in-deployment](https://www.wearedevelopers.com/jobs/ext/1762822-software-engineer-ii-ai-engineer-w-skills-in-deployment). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Software Engineer II - AI Engineer (w/ skills in Deployment) - **Company:** Robert Half - **Location:** San Francisco, CA, United States - **Salary:** $85,000.0 - $124,000.0 - **Contract:** Permanent contract - **Skills:** Java (Programming Language), JavaScript (Programming Language), A/B Testing, Application Programming Interfaces (APIs), Artificial Intelligence, Amazon Web Services, Microsoft Azure, C Sharp (Programming Language), Databases, Continuous Integration, Data Integration, Software Debugging, Fault Tolerance, Interaction Design, Python (Programming Language), Systems Development Life Cycle, Cloud Services, Search Technologies, Software Deployment, Software Engineering, SQL Databases, Systems Architecture, Management of Software Versions, Data Logging, Large Language Models, Prompt Engineering, Caching, Generative AI, Machine Learning Operations - **Published:** July 9, 2026 - **Apply:** https://dejobs.org/x/x/2E0F273DFDE74BD8B0CC69D78C6EEB5F/job/ ## About the Role * 4+ years of experience in IT or a related field. * 2+ years of software engineering experience. * 1+ year of experience in GenAI deployment. * Experience with AI coding agent-augmented development. * Experience with cost optimization, including token and caching strategies. * Experience with Python, Java, C#, JavaScript, or SQL. * Experience building and deploying applications. * Knowledge of cloud platforms, containers, and CI/CD. * Understanding of SDLC, APIs, and system architecture. * Knowledge of databases and data integration. * Understanding of LLM fundamentals and token behavior. * Experience with LLMOps/MLOps, including versioning and experiment tracking. * Experience with prompt engineering techniques. * Experience with RAG pipelines, including embeddings and vector search. * Familiarity with evaluation metrics, including accuracy and hallucination risk. * Knowledge of GenAI debugging and safety mechanisms. * Knowledge of deployment governance, including access control and compliance. * Experience with observability, including logging, tracing, and monitoring. * Experience with Azure, AWS, or GCP. * Strong communication and requirements-gathering skills. Preferred Generative AI Skills * Experience with model orchestration, including multi-step workflows and agents. * Experience with LangChain, Semantic Kernel, or AutoGen. * Experience designing embeddings and semantic search solutions. * Experience with experimentation and A/B testing of prompts and models. * Experience with conversational UX and human-AI interaction design. * Experience with incident handling, including retries and graceful degradation. ## Description * Build prompt workflows, retrieval layers, APIs, and cloud services. * Troubleshoot production issues, including latency, hallucinations, and errors. * Provide Level II production support for deployed systems. * Design components, including LLM integrations and RAG pipelines. * Implement CI/CD pipelines, containerization, and release processes. * Develop RAG pipelines with embeddings, chunking, and vector search. * Apply prompt engineering techniques, including few-shot prompting and structured outputs. * Evaluate models for accuracy, relevance, and hallucination risk. * Implement safety guardrails, including PII protection and prompt-injection defense. * Execute testing, including unit, integration, and GenAI evaluation testing. * Monitor production systems for latency, cost, usage, and errors. * Support incident management with fallback and recovery strategies. ## Related Videos - [Bringing the power of AI to your application.](https://www.wearedevelopers.com/videos/1010-bringing-the-power-of-ai-to-your-application) - [HTTP headers that make your website go faster](https://www.wearedevelopers.com/videos/1676-http-headers-that-make-your-website-go-faster) - [Kubernetes and Microservices with Multi-Model Databases](https://www.wearedevelopers.com/videos/382-kubernetes-and-microservices-with-multi-model-databases) - [The State of GenAI & Machine Learning in 2025](https://www.wearedevelopers.com/videos/1383-the-state-of-genai-machine-learning-in-2025) - [Event based cache invalidation in GraphQL](https://www.wearedevelopers.com/videos/433-event-based-cache-invalidation-in-graphql) - [Should we build Generative AI into our existing software?](https://www.wearedevelopers.com/videos/1129-should-we-build-generative-ai-into-our-existing-software) ## Related Articles - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [Got AI ideas but no money? Here are 10 free ways to level up your AI skills with Google Cloud](https://www.wearedevelopers.com/magazine/600-got-ai-ideas-but-no-money-here-are-10-free-ways-to-level-up-your-ai-skills-with-google-cloud) - [The Prompt Engineer ✍️](https://www.wearedevelopers.com/magazine/216-the-prompt-engineer) - [What is Agentic Programming and Why Should Developers Care?](https://www.wearedevelopers.com/magazine/625-what-is-agentic-programming-and-why-should-developers-care) - [What Are Large Language Models?](https://www.wearedevelopers.com/magazine/304-what-are-large-language-models) - [Prompt Engineering is a Job of the Past](https://www.wearedevelopers.com/magazine/342-prompt-engineering-is-a-job-of-the-past)