> Markdown version of [/jobs/ext/1314834-generative-ai-ml-system-engineering](https://www.wearedevelopers.com/jobs/ext/1314834-generative-ai-ml-system-engineering). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Generative AI - ML System Engineering - **Company:** Meshy LLC - **Location:** United States - **Salary:** $175,000.0 - $300,000.0 - **Contract:** Permanent contract - **Skills:** Clean Code Principles, Application Programming Interfaces (APIs), Artificial Intelligence, Systems Engineering, C++ (Programming Language), Nvidia CUDA, Software Debugging, Distributed Computing Environment, Python (Programming Language), Machine Learning, Graphics Processing Unit (GPU), Pytorch, Generative AI, Machine Learning Operations - **Published:** July 17, 2026 - **Apply:** https://www.indeed.com/viewjob?jk=a92418a54ab42844 ## About the Role * Experience in machine learning or high performance graphics. * Solid practical understanding of at least one machine learning framework (e.g. PyTorch, JAX). * Strong ability to write beautiful and maintainable code in Python and/or C++. * Ability to learn fast and dive into new concepts or complex codebases. * Performance and efficiency oriented mindset, with a strong interest in the tiniest detail. * Strong communication skills for working in a globally distributed team., * A strong passion to navigate through the PyTorch internals, with hands-on experience in areas like torch.compile , fully_shard (FSDP2) APIs. * Experience with building Triton kernels. * Experiences with large-scale distributed training, familiarity with modern parallelization techniques: DP, TP, CP, PP, zero redundancy optimizers, etc. * Experience with diffusion models in 3D or video. * Experience with low precision bf16 or fp8 training. ## Description We are looking for Machine Learning Systems Engineers who can help us build the world's largest end-to-end 3D native machine learning systems. You will help us build our end to end ML framework dedicated for 3D, from pretraining, to finetuning, inferencing, etc. We expect a combination of strong hands on engineering skills, eagerness to learn new things, and thrives in a fast-paced, high-ownership environment., * Work closely with researchers to co-design the next frontier of 3D & Spatial AI. * Build and debug on top of modern PyTorch, for maximum parallelism and efficiency, and build clean and intuitive training infrastructure for our in-house foundational models. * Identifying bottlenecks and optimizing for high throughput & efficient distributed model training across hundreds to thousands of GPUs. * Implementing and maintaining 3D specific custom operators in Triton or CUDA. * Implementing and maintaining novel data-loading framework and libraries. On the inference side * Building efficient inference endpoints with complex multi-stage model pipelines. * Optimizing models through compilation, fusion, quantization, etc., * Soon afterwards, you will receive an online assessment of your knowledge about various engineering topics around training, inference, transformer architecture, and simple numpy coding exercises. * We will then schedule a 45 minutes - 2 hr interview slot for a technical coding round. The questions will revolve around performant C++ programming, tensor / array programming in PyTorch, and some practical hands-on open-book / open-internet training exercises in our GPU-enabled jupyter notebook. * Finally, you will be invited for a 3 hr onsite interview where we'd like you to present a previous work that you are proud of, then you will demonstrate your debugging skills and performance sense in a session with one of our engineers. Finally, you will talk to one of our leaders and our CEO about our culture, your background, and whether we have matching vibes. ## Related Videos - [Coffee with Developers - Stephen Jones - NVIDIA](https://www.wearedevelopers.com/videos/1303-coffee-with-developers-stephen-jones-nvidia) - [Your Next AI Needs 10,000 GPUs. Now What?](https://www.wearedevelopers.com/videos/1590-your-next-ai-needs-10-000-gpus-now-what) - [Your imaginations is (no longer) the limit: how Generative AI empowers people to be creative](https://www.wearedevelopers.com/videos/741-your-imaginations-is-no-longer-the-limit-how-generative-ai-empowers-people-to-be-creative) - [Photonic Computing: Programming a New Class of AI Accelerators (incl. Live Coding)](https://www.wearedevelopers.com/videos/100196-photonic-computing-programming-a-new-class-of-ai-accelerators-incl-live-coding) - [Geometric deep learning for drug discovery](https://www.wearedevelopers.com/videos/264-geometric-deep-learning-for-drug-discovery) - [Nemotron: NVIDIA's open model strategy for developers](https://www.wearedevelopers.com/videos/100064-nemotron-nvidia-s-open-model-strategy-for-developers) ## Related Articles - [MLOps And AI Driven Development](https://www.wearedevelopers.com/magazine/82-mlops-and-ai-driven-development) - [Got AI ideas but no money? Here are 10 free ways to level up your AI skills with Google Cloud](https://www.wearedevelopers.com/magazine/600-got-ai-ideas-but-no-money-here-are-10-free-ways-to-level-up-your-ai-skills-with-google-cloud) - [Graph and AI Trends 2026: Why Is AI Running but Not Yet Delivering?](https://www.wearedevelopers.com/magazine/680-graph-and-ai-trends-2026-why-is-ai-running-but-not-yet-delivering) - [What Are Large Language Models?](https://www.wearedevelopers.com/magazine/304-what-are-large-language-models) - [What is Agentic Programming and Why Should Developers Care?](https://www.wearedevelopers.com/magazine/625-what-is-agentic-programming-and-why-should-developers-care) - [Everything a Developer Needs to Know About MCP with Neo4j](https://www.wearedevelopers.com/magazine/604-everything-a-developer-needs-to-know-about-mcp-with-neo4j)