> Markdown version of [/jobs/ext/531133-ml-infrastructure-engineer](https://www.wearedevelopers.com/jobs/ext/531133-ml-infrastructure-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # ML Infrastructure Engineer - **Company:** Tennr Incorporated - **Location:** New York, NY, United States - **Experience:** Expert - **Salary:** $190,000.0 - $230,000.0 - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Systems Engineering, Program Optimization, Nvidia CUDA, Continuous Integration, Data Systems, Distributed Systems, Python (Programming Language), Machine Learning, Software Architecture, Tensorflow, Software Systems, TypeScript, Data Logging, Pytorch, Large Language Models, Model Validation, Machine Learning Operations, TensorRT, Data Pipelines - **Published:** June 10, 2026 - **Apply:** https://www.indeed.com/viewjob?jk=25a8b8ac3680ba13 ## About the Role Do you have experience in Systems engineering?, * 5+ years of experience in ML model deployment, infrastructure, and scaling in production environments * Strong software engineering fundamentals, with proficiency in Python and TypeScript * Experience in software design and architecture for highly available ML systems for use cases like inference, evaluation, and experimentation * Strong knowledge of observability, including logging, metrics, tracing, model performance monitoring, and alerting * Experience with distributed systems, reliability, and production incident response * Comfortable working in ambiguity with high ownership, moving quickly in a fast-paced startup environment, and proactively driving projects from idea to production * Nice to have: + Experience working with ML CI/CD and common ML frameworks like Pytorch, Tensorflow, etc. + Experience working with common inference frameworks like vLLM, TensorRT, Triton, etc + Experience with GPU orchestration, including managing GPU workloads/scheduling, cost management, cluster utilization, etc + Experience with GPU optimization (training/inference) involving CUDA profiling, memory optimization, multi-GPU communication, etc ## Description As the first ML Ops Engineer at Tennr, you'll play a crucial role in building and iterating on foundational Machine Learning and AI systems. You'll own building machine learning training and inference pipelines that can handle increasing traffic demands and proliferation of product surface as we grow. You will be critical in ensuring our AI-driven healthcare platform is powered by robust, scalable, and efficiently deployed models. Our Machine Learning team owns and develops multiple in-house, proprietary VLMs, LLMs, and other models that are purpose-built for the ambitious problems we are solving in the healthcare space. This is not a role where you are repackaging and wrapping old innovations, but an opportunity to be on the cutting edge of experimentation and productization of net new capabilities. You'll make impactful contributions and influence fundamental elements of our ML and data systems, expanding Tennr's ability to rapidly iterate and solve critical problems for patients and providers., * Architect, design, and implement ML software systems for deploying and managing models at scale. * Develop and maintain infrastructure that supports efficient ML operations, including data pipelines, model evaluations, deployments, and training at scale. * Collaborate closely with ML engineers, software engineers, and cross-functional teams to ensure seamless integration of models with data pipelines and products. * Troubleshoot production issues and continuously improve systems to enhance performance and efficiency. * Create tooling for online and offline evaluation of ML & LLM systems., * Drive Impact: one of our company values is Cowboy, meaning you set the pace. You won't just talk about things, you'll get them done. And feel the impact. * Develop Operational Expertise: learn the inner workings of scaling systems, tools, and infrastructure * Innovate with Purpose: we're not just doing this for fun (although we do have a lot of fun). At Tennr, you'll join a high-caliber team maniacally focused on reducing patient delays across the U.S. healthcare system. * Build Relationships: collaborate and connect with like-minded, driven individuals in our Hudson Square office 4 days/week * Free lunch! Plus a pantry full of snacks. ## Related Videos - [Tour de Force: Open-Source LLM Inference Optimization from Simple to Sophisticated](https://www.wearedevelopers.com/videos/100099-tour-de-force-open-source-llm-inference-optimization-from-simple-to-sophisticated) - [LLMOps-driven fine-tuning, evaluation, and inference with NVIDIA NIM & NeMo Microservices](https://www.wearedevelopers.com/videos/1582-llmops-driven-fine-tuning-evaluation-and-inference-with-nvidia-nim-nemo-microservices) - [Do TypeScript without TypeScript](https://www.wearedevelopers.com/videos/327-do-typescript-without-typescript) - [Photonic Computing: Programming a New Class of AI Accelerators (incl. Live Coding)](https://www.wearedevelopers.com/videos/100196-photonic-computing-programming-a-new-class-of-ai-accelerators-incl-live-coding) - [Efficient deployment and inference of GPU-accelerated LLMs​](https://www.wearedevelopers.com/videos/929-efficient-deployment-and-inference-of-gpu-accelerated-llms) - [Trends, Challenges and Best Practices for AI at the Edge](https://www.wearedevelopers.com/videos/630-trends-challenges-and-best-practices-for-ai-at-the-edge) ## Related Articles - [What Are Large Language Models?](https://www.wearedevelopers.com/magazine/304-what-are-large-language-models) - [MLOps And AI Driven Development](https://www.wearedevelopers.com/magazine/82-mlops-and-ai-driven-development) - [MLOps – What’s the deal behind it?](https://www.wearedevelopers.com/magazine/125-mlops-what-s-the-deal-behind-it) - [MLops – Deploying, Maintaining And Evolving Machine Learning Models in Production](https://www.wearedevelopers.com/magazine/115-mlops-deploying-maintaining-and-evolving-machine-learning-models-in-production) - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [Best US AI Conferences for CTOs in 2026: Build vs. Buy, Vendor Evaluation, and Peer Intelligence](https://www.wearedevelopers.com/magazine/736-best-us-ai-conferences-for-ctos-in-2026-build-vs-buy-vendor-evaluation-and-peer-intelligence)