Machine Learning Engineer

WebAI, Inc.
Austin, TX, United States
about 2 months ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
4 years minimum
Working hours
Regular working hours
Job source

Tech stack

Artificial Intelligence Computer Vision Profiling Nvidia CUDA Databases Distributed Systems General-Purpose Computing on Graphics Processing Units Machine Learning Tensorflow Systems Integration Pytorch Generative AI
+4 more
HuggingFace Machine Learning Operations Microservices Data Generation

Job description

We are seeking a Senior Machine Learning Engineer to support our Public Sector initiatives focused on building and optimizing production ready AI systems for secure and distributed environments.

You will be responsible for transforming prototype models into scalable, efficient, and reliable production systems that operate seamlessly across a spectrum of hardware from government cloud infrastructure to edge devices in restricted or disconnected environments.

Responsibilities:

  • Design, develop, and deploy agentic workflows to orchestrate multi-step reasoning, tool use, and decision-making across production systems.
  • Productionize AI models from research prototypes into scalable, deployable systems used in real world applications.
  • Engineer adaptive ML systems using LoRA, PEFT, and on-device inference strategies, leveraging PyTorch, TensorFlow, and Hugging Face Transformers for model development, fine-tuning, and optimization.
  • Implement model optimization techniques such as quantization, pruning, distillation, and hardware specific acceleration.
  • Build and maintain Retrieval Augmented Generation (RAG) pipelines, including vector database integration for contextual retrieval.
  • Work with multi-modal AI systems across computer vision, audio, and natural language domains.
  • Optimize model execution for distributed and resource constrained environments, ensuring reliability under variable connectivity conditions., We at webAI are committed to living out the core values we have put in place as the foundation on which we operate as a team. We seek individuals who exemplify the following:
  • Truth - Emphasizing transparency and honesty in every interaction and decision.
  • Ownership - Taking full responsibility for one’s actions and decisions, demonstrating commitment to the success of our clients.
  • Tenacity - Persisting in the face of challenges and setbacks, continually striving for excellence and improvement.
  • Humility - Maintaining a respectful and learning-oriented mindset, acknowledging the strengths and contributions of others.

Requirements

  • Active US Security clearance
  • 4+ years of experience in applied AI, ML engineering, or production AI systems.
  • Deep proficiency in PyTorch, TensorFlow, or Hugging Face Transformers.
  • Proven experience deploying AI models across cloud, edge, and mobile hardware environments.
  • Expertise in model compression and optimization (quantization, pruning, distillation).
  • Experience building RAG pipelines and integrating vector databases (e.g., Quadrant, ChromaDB, FAISS, Milvus, Pinecone).
  • Familiarity with multi-modal models and synthetic data generation methods.
  • Strong algorithmic and problem solving skills, especially in distributed or constrained compute environments.

Preferred Skills:

  • Experience with edge AI, federated learning, or offline inference systems.
  • Understanding of AI governance and compliance frameworks relevant to public sector deployments.
  • Experience integrating models into large scale distributed systems or microservice architectures.
  • Excellent communication and technical documentation skills for collaboration across multi disciplinary teams.
  • Strong understanding of GPU computing, CUDA, and performance profiling.

Benefits & conditions

  • Competitive salary
  • Comprehensive health, dental, and vision benefits package
  • 401(k) match (U.S.-based employees only)
  • $200/month Health & Wellness stipend
  • Continuing Education support
  • $500/year Function Health subscription (U.S.-based employees only)
  • Free parking for in-office employees
  • Flexible Time Off (FTO)
  • Parental leave for eligible employees
  • Supplemental life insurance

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on dice.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:03 min

Solving complex engineering challenges in artificial intelligence deployment

Nico Axtmann · WWC 2022

2:35 min

Preventing remote code execution in PyTorch models

Balázs Kiss · WWC 2023

47 sec

Profiling native execution calls with async-profiler

Gonzalo Ortiz Jaureguizar Gonzalo Ortiz Jaureguizar · WWC Europe 2026

3:04 min

Database evolution and the funding behind vector databases

Erik Bamberg · LIVE

2:15 min

Open-source community and machine learning frameworks

Gian Marco Iodice Gian Marco Iodice · WWC 2025

4:41 min

Replacing PyTorch with ONNX runtime for AWS Lambda deployments

Marek Suppa · LIVE

Videos

See all

Related articles

See all