AI/Ml Engineer

Robert Half
Houston, TX, United States
23 days ago

Role details

Contract type
Temporary contract
Employment type
Full-time (> 32 hours)
Experience level
Experienced
Experience required
3 years minimum
Working hours
Regular working hours
Job source

Tech stack

Application Programming Interfaces (APIs) Artificial Intelligence Amazon Web Services Software Applications Microsoft Azure Cloud Computing Continuous Integration Information Engineering Python (Programming Language) Machine Learning Natural Language Processing Open Source Technology
+9 more
Tensorflow Pytorch Large Language Models Generative AI AI Platforms Kubernetes Information Technology Machine Learning Operations Databricks

Job description

We are seeking an AI/ML Engineer with a focus on Generative AI to design, develop, and deploy advanced AI solutions that drive business innovation. This role will leverage large language models (LLMs), machine learning, and cloud-based AI technologies to build intelligent applications, automate workflows, and enhance data-driven decision-making., * Design, develop, and deploy Generative AI solutions using LLMs such as OpenAI, Claude, Gemini, or open-source models.

  • Build and optimize AI/ML pipelines for model training, fine-tuning, evaluation, and inference.
  • Develop Retrieval-Augmented Generation (RAG) architectures and integrate vector databases.
  • Collaborate with software engineers, data engineers, and business stakeholders to deliver AI-powered applications.
  • Implement prompt engineering strategies and model optimization techniques to improve performance and accuracy.
  • Monitor, troubleshoot, and enhance AI models in production environments.
  • Ensure AI solutions adhere to security, governance, and responsible AI best practices.

Requirements

  • Bachelor’s degree in Computer Science, Data Science, Engineering, or related field.
  • 3+ years of experience in Machine Learning, AI, or Data Engineering.
  • Proficiency in Python and ML frameworks such as PyTorch, TensorFlow, LangChain, or LlamaIndex.
  • Experience with cloud platforms (Azure, AWS, or GCP) and AI services.
  • Knowledge of LLMs, NLP, vector databases, embeddings, and RAG frameworks.
  • Strong understanding of MLOps, APIs, and scalable system design.

Preferred Skills:

  • Experience fine-tuning foundation models and deploying AI solutions in production.
  • Familiarity with Azure AI Foundry, Databricks, MLflow, Kubernetes, and CI/CD pipelines.
  • Excellent problem-solving, communication, and stakeholder management skills. Technology Doesn’t Change the World, People Do.®

About the company

Robert Half is the world’s first and largest specialized talent solutions firm that connects highly qualified job seekers to opportunities at great companies. We offer contract, temporary and permanent placement solutions for finance and accounting, technology, marketing and creative, legal, and administrative and customer support roles.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on dejobs.org

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:35 min

Preventing remote code execution in PyTorch models

Balázs Kiss · WWC 2023

2:28 min

Understanding Kubernetes architecture and core cluster components

Marc Nimmerrichter · WWC 2022

1:39 min

Fundamentals of tensors and the TensorFlow library

Håkan Silfvernagel · LIVE

2:08 min

Essential engineering roles in the generative AI space

Mary Grygleski Mary Grygleski · LIVE

4:41 min

Replacing PyTorch with ONNX runtime for AWS Lambda deployments

Marek Suppa · LIVE

1:31 min

Essential AI and human skills for future teams

Alexander Weißhaupt Alexander Weißhaupt +1 · WWC 2025

Videos

See all

Related articles

See all