> Markdown version of [/jobs/ext/2963031-ml-systems-engineer-fully-remote](https://www.wearedevelopers.com/jobs/ext/2963031-ml-systems-engineer-fully-remote). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # ML Systems Engineer - Fully Remote - **Company:** Mercor, Inc. - **Location:** New York, NY, United States (Remote available) - **Experience:** Expert - **Salary:** $187,200.0 - $249,600.0 - **Contract:** Permanent contract - **Skills:** Training Data, Artificial Intelligence, Profiling, Nvidia CUDA, Software Debugging, Pytorch, Machine Learning Operations - **Published:** September 17, 2026 - **Apply:** https://www.careerjet.com/jobad/us869f3b927c1765c53452745e17f747f4 ## About the Role * 2+ years of hands-on professional experience in ML systems, ML infrastructure, or GPU performance engineering. * Practical experience in GPU kernels (e.g., CUDA, Triton) or performance profiling (e.g., Kineto, torch.profiler). * Working production experience with JAX and/or PyTorch. * Familiarity with modern accelerators such as A100, H100, or TPU. * Strong written communication skills. ## Description * Design challenging tasks across GPU kernels, performance profiling, debugging, and inference serving. Write accurate, well-structured solutions. * Guide research and engineering teams to close knowledge gaps and improve AI model performance on ML systems and training infrastructure. * Evaluate MLOps and ML systems tasks. Provide clear, written technical feedback that stands up to reviewer scrutiny. * Develop guidelines and detailed rubrics covering kernel-level optimization, profiler output interpretation, and serving throughput and latency trade-offs. * Collaborate with other subject matter experts to ensure training data consistency and accuracy., PS: Our team reviews applications daily. Please complete your AI interview and application steps to be considered for this opportunity. ## Related Videos - [Coffee with Developers - Stephen Jones - NVIDIA](https://www.wearedevelopers.com/videos/1303-coffee-with-developers-stephen-jones-nvidia) - [Nemotron: NVIDIA's open model strategy for developers](https://www.wearedevelopers.com/videos/100064-nemotron-nvidia-s-open-model-strategy-for-developers) - [Photonic Computing: Programming a New Class of AI Accelerators (incl. Live Coding)](https://www.wearedevelopers.com/videos/100196-photonic-computing-programming-a-new-class-of-ai-accelerators-incl-live-coding) - [Profiling Symfony & PHP apps with Blackfire](https://www.wearedevelopers.com/videos/265-profiling-symfony-php-apps-with-blackfire) - [Introduction to Azure Machine Learning](https://www.wearedevelopers.com/videos/368-introduction-to-azure-machine-learning) - [Serverless deployment of (large) NLP models ](https://www.wearedevelopers.com/videos/158-serverless-deployment-of-large-nlp-models) ## Related Articles - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [What Are Large Language Models?](https://www.wearedevelopers.com/magazine/304-what-are-large-language-models) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again)