Software Engineer II - AI Engineer (w/ skills in Deployment)

Robert Half
San Francisco, CA, United States
about 1 month ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience required
1 year minimum
Compensation
$85,000.0 - $124,000.0
Working hours
Regular working hours
Job source

Tech stack

Java (Programming Language) JavaScript (Programming Language) A/B Testing Application Programming Interfaces (APIs) Artificial Intelligence Amazon Web Services Microsoft Azure C Sharp (Programming Language) Databases Continuous Integration Data Integration Software Debugging
+17 more
Fault Tolerance Interaction Design Python (Programming Language) Systems Development Life Cycle Cloud Services Search Technologies Software Deployment Software Engineering SQL Databases Systems Architecture Management of Software Versions Data Logging Large Language Models Prompt Engineering Caching Generative AI Machine Learning Operations

Job description

  • Build prompt workflows, retrieval layers, APIs, and cloud services.
  • Troubleshoot production issues, including latency, hallucinations, and errors.
  • Provide Level II production support for deployed systems.
  • Design components, including LLM integrations and RAG pipelines.
  • Implement CI/CD pipelines, containerization, and release processes.
  • Develop RAG pipelines with embeddings, chunking, and vector search.
  • Apply prompt engineering techniques, including few-shot prompting and structured outputs.
  • Evaluate models for accuracy, relevance, and hallucination risk.
  • Implement safety guardrails, including PII protection and prompt-injection defense.
  • Execute testing, including unit, integration, and GenAI evaluation testing.
  • Monitor production systems for latency, cost, usage, and errors.
  • Support incident management with fallback and recovery strategies.

Requirements

  • 4+ years of experience in IT or a related field.
  • 2+ years of software engineering experience.
  • 1+ year of experience in GenAI deployment.
  • Experience with AI coding agent-augmented development.
  • Experience with cost optimization, including token and caching strategies.
  • Experience with Python, Java, C#, JavaScript, or SQL.
  • Experience building and deploying applications.
  • Knowledge of cloud platforms, containers, and CI/CD.
  • Understanding of SDLC, APIs, and system architecture.
  • Knowledge of databases and data integration.
  • Understanding of LLM fundamentals and token behavior.
  • Experience with LLMOps/MLOps, including versioning and experiment tracking.
  • Experience with prompt engineering techniques.
  • Experience with RAG pipelines, including embeddings and vector search.
  • Familiarity with evaluation metrics, including accuracy and hallucination risk.
  • Knowledge of GenAI debugging and safety mechanisms.
  • Knowledge of deployment governance, including access control and compliance.
  • Experience with observability, including logging, tracing, and monitoring.
  • Experience with Azure, AWS, or GCP.
  • Strong communication and requirements-gathering skills.

Preferred Generative AI Skills

  • Experience with model orchestration, including multi-step workflows and agents.
  • Experience with LangChain, Semantic Kernel, or AutoGen.
  • Experience designing embeddings and semantic search solutions.
  • Experience with experimentation and A/B testing of prompts and models.
  • Experience with conversational UX and human-AI interaction design.
  • Experience with incident handling, including retries and graceful degradation.

Benefits & conditions

The typical annual salary range for this position is shown below and is negotiable depending upon experience and location. The position is eligible for a discretionary annual bonus.

$85,000.00 - $124,000.00

We offer exceptional earning potential and a competitive benefits package, including group health insurance benefits (medical, vision, dental), FSA and HSA healthcare accounts, life and accident insurance, adoption and fertility assistance, paid parental leave of up to 6 weeks, and short/long term disability. Robert Half provides paid time off for vacation, personal needs, and sick time. The amount of Choice Time Off (CTO) our people receive varies based on their years of service and is pro-rated based on the hours worked per week. A new hire earns up to 17 days of CTO per calendar year. Our people also receive up to 11 paid holidays per calendar year. We also offer the opportunity to contribute to our company 401(k) savings and investment plan or deferred compensation plan (if eligible), with an employer match of 100% on the first 3% of your contributions for eligible employees. Learn more at https://roberthalfbenefits.com .

Robert Half Inc. is an Equal Opportunity Employer. M/F/Disability/Veteran

As part of Robert Half’s Corporate Services facility employment process, any offer of employment is contingent upon successful completion of a background check.

Our recruiters use their expertise and may utilize AI to help with their evaluation of candidates.

About the company

Robert Half is seeking a Software Engineer II - AI Engineer who will analyze, design, program, debug, test, implement, deploy, and support software enhancements and new applications using Generative AI technologies. This role contributes to the development and production deployment of GenAI-enabled applications, including LLM-powered workflows, RAG pipelines, and AI-driven user experiences., For positions located in San Francisco, CA: Robert Half will consider qualified applicants with criminal histories in a manner consistent with the requirements of the San Francisco Fair Chance Ordinance.

For positions located in Los Angeles County, CA: Robert Half will consider for employment qualified applicants with arrest or conviction records in accordance with the Los Angeles County Fair Chance Ordinance for Employers and the California Fair Chance Act.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on dejobs.org

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:15 min

Reversing the caching model for artifact delivery

Thijs Feryn Thijs Feryn · WWC Europe 2026

3:04 min

Database evolution and the funding behind vector databases

Erik Bamberg · LIVE

54 sec

Generating multiple hook options for outreach A/B testing

Leandro Gomes da Silva Leandro Gomes da Silva · WWC 2025

2:08 min

Essential engineering roles in the generative AI space

Mary Grygleski Mary Grygleski · LIVE

2:33 min

Maintaining prompt structures for prefix caching

Douglas Reiser Douglas Reiser · Europe 2026 Virtual

3:22 min

Evaluating advanced artificial intelligence platforms for daily recruitment

Rudi Bauer Rudi Bauer +1 · Cappuccino with HR

Videos

See all

Related articles

See all