Machine Learning Engineer

EVLO, INC.
United States
17 days ago
Apply on arc.dev
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Experienced
Experience required
3 years minimum
Working hours
Regular working hours
Job source

Tech stack

Artificial Intelligence Amazon Web Services Microsoft Azure Big Data Cloud Computing Code Review Distributed Computing Environment Python (Programming Language) Machine Learning Open Source Technology Tensorflow Pytorch
+11 more
Large Language Models Deep Learning Backend Containerization Kubernetes Information Technology Low Latency Hardware Acceleration Machine Learning Operations Data Pipelines Docker

Job description

The role owns the architecture, development, and scaling of machine learning systems, driving the transition of advanced AI models from research into high-throughput production environments.

The team collaborates closely with applied scientists and backend engineers to ensure models achieve optimal performance, low latency, and robust reliability under heavy enterprise workloads., * Design and implement scalable machine learning pipelines for model training, validation, and inference using Python, PyTorch, and distributed computing frameworks

  • Deploy, monitor, and scale models on cloud platforms like AWS SageMaker or GCP Vertex AI with automated CI/CD pipelines
  • Optimize model inference latency, throughput, and memory footprint through quantization, pruning, and hardware acceleration techniques
  • Build feature and data ingestion pipelines handling large-scale datasets, ensuring consistency between training and production feature stores
  • Implement comprehensive monitoring frameworks to track model performance, data drift, and anomaly detection in real-time production environments
  • Write rigorous unit and integration tests, conduct code reviews, and establish engineering best practices for the broader machine learning team

Requirements

  • 3-6 years of professional software and machine learning engineering experience with a track record of deploying models to production
  • Strong proficiency in Python and hands-on experience with deep learning frameworks such as PyTorch or TensorFlow
  • Solid understanding of MLOps best practices, containerization with Docker, and orchestration using Kubernetes
  • Experience with cloud infrastructure (AWS, GCP, or Azure) and modern feature stores or vector databases
  • Bachelor’s or Master’s degree in Computer Science, Machine Learning, Statistics, or a related quantitative field
  • Bonus: Experience fine-tuning large language models, contributing to open-source ML projects, or publishing research at top-tier AI conferences

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on arc.dev
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

5:28 min

Defining MLOps and its role in production systems

Hauke Brammer · World Congress 2023

2:07 min

Inspecting default bridge architectures and custom Docker networks

Oliver Seitz Oliver Seitz · World Congress 2025

1:52 min

Structuring and scaling the backend engineering team

Stefan Lingler Stefan Lingler +1 · Coffee With Developers

2:35 min

Preventing remote code execution in PyTorch models

Balázs Kiss · World Congress 2023

2:08 min

Essential engineering roles in the generative AI space

Mary Grygleski Mary Grygleski · LIVE

2:34 min

Docker sandbox architecture and microVM environment integration

Manuel de la Peña Manuel de la Peña · World Congress 2026 Europe

Videos

See all

Related articles

See all