AI Infrastructure Specialist - Fully Remote

Mercor, Inc.
New York, NY, United States
3 months ago

Role details

Contract type
Temporary contract
Employment type
Part-time (≤ 32 hours)
Experience level
Expert
Experience required
5 years minimum
Compensation
$187,200.0 - $270,400.0
Working hours
Regular working hours
Job source

Tech stack

Training Data Artificial Intelligence Amazon Web Services Program Optimization Software Debugging DevOps Distributed Systems AI Infrastructure Pytorch Machine Learning Operations

Job description

  • Guide research and engineering teams to close knowledge gaps and improve AI model performance in MLOps, training infrastructure, and ML framework-level topics.
  • Design challenging, domain-relevant tasks across multiple specializations. Write accurate and well-structured solutions to MLOps and ML systems problems.
  • Evaluate MLOps tasks and solutions. Provide clear, written technical feedback.
  • Develop guidelines and detailed rubrics/evaluation frameworks to assess training pipeline design, distributed systems reasoning, and kernel-level optimization across tasks.
  • Collaborate with other subject matter experts to ensure consistency and accuracy in training data., PS: Our team reviews applications daily. Please complete your AI interview and application steps to be considered for this opportunity., Position: Infrastructure Technical Specialist (ITS) FileCloud & AWS Location: 100% remote Duration: 12 Months Infrastructure Technical Specialist (ITS) FileCloud & AWS Key …
  • 3 hours ago
  • Apply easily

Requirements

Must-Have

  • 5+ years of dedicated professional experience in ML infrastructure, MLOps, or ML systems engineering at a recognized, top-tier organization.
  • Hands-on production experience with JAX and/or PyTorch at scale, including distributed training strategies (FSDP, tensor parallelism, pipeline parallelism), memory optimization, and framework-level debugging.
  • Experience writing or optimizing custom GPU kernels using Pallas (JAX) or Triton, including tiling strategies, memory layout design, and kernel fusion.
  • Demonstrable career progression.
  • Ability to engage reliably for at least 30 hours/week during weekdays.
  • Strong written communication skills and the ability to explain complex technical decisions clearly.

About the company

Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors include Benchmark, General Catalyst, Peter Thiel, Adam D’Angelo, Larry Summers, and Jack Dorsey., © 2026 Careerjet All rights reserved

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on careerjet.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:17 min

Mapping the maturity roadmap for scaled devops adoption

Dominik Krichbaum Dominik Krichbaum · WWC Europe 2026

2:35 min

Preventing remote code execution in PyTorch models

Balázs Kiss · WWC 2023

2:20 min

Architecting language translation with focused training data

Jaroslaw Kutylowski Jaroslaw Kutylowski +1 · WWC 2023

2:08 min

Essential engineering roles in the generative AI space

Mary Grygleski Mary Grygleski · LIVE

1:24 min

Comprehensive AI infrastructure stacks at the Linux Foundation

Matt White Matt White · WWC 2025

3:18 min

Scaling global network engineering through DevOps culture

Stuart Clark · LIVE

Videos

See all

Related articles

See all