> Markdown version of [/jobs/ext/2072813-lead-data-scientist-nvidia](https://www.wearedevelopers.com/jobs/ext/2072813-lead-data-scientist-nvidia). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Lead Data Scientist (Nvidia) - **Company:** SoftServe, Inc. - **Location:** United States (Remote available) - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Amazon Web Services, Microsoft Azure, Profiling, Nvidia CUDA, Distributed Systems, Python (Programming Language), Machine Learning, Language Modeling, NumPy, Tensorflow, Software Engineering, AI Infrastructure, Pytorch, Large Language Models, Deep Learning, Generative AI, Pandas, Kubernetes, Information Technology, Low Latency, ONNX (Open Neural Network Exchange) Format, HuggingFace, Machine Learning Operations, TensorRT, Hardware Infrastructure, Virtual Agents, Nim (Programming Language) - **Published:** August 15, 2026 - **Apply:** https://arc.dev/remote-jobs/j/redirect/pb8tnh3vqm ## About the Role * 6+ years of experience in AI consulting, Generative AI, Agentic AI, Machine Learning, or Deep Learning, including ownership of client-facing engagements * Bachelor's or Master's degree in Computer Science, Applied Mathematics, Physics, Engineering, or related technical field preferred * Advanced expertise in Generative AI, Agentic AI, multimodal AI, transformers, Large Language Models (LLMs), and Vision Language Models (VLMs) * Hands-on experience with Python and modern AI/ML frameworks, including PyTorch, TensorFlow, Hugging Face, Pandas, and NumPy * Strong experience with NVIDIA AI technologies, including at least three of the following: NeMo, NIM, Triton, TensorRT-LLM, Riva, DeepStream, Metropolis, or Omniverse * Practical experience in deploying AI workloads on Kubernetes using Helm, NVIDIA GPU Operator, GPU device plugins, MIG/vGPU partitioning, and modern inference platforms such as vLLM or Ollama * Working knowledge of model quantization, inference optimization, and GPU profiling tools, including NVIDIA Nsight Systems, Nsight Compute, DCGM, PyTorch Profiler, and Triton or vLLM monitoring * Proven skill in analyzing GPU performance, identifying compute, memory, or I/O bottlenecks, and optimizing AI infrastructure for performance and cost efficiency * Experience in designing and deploying enterprise AI solutions on AWS, Azure, or GCP using CUDA, TensorRT, Triton Inference Server, DeepStream, and ONNX * Solid understanding of enterprise architecture, distributed systems, MLOps, AI governance, and modern software engineering practices * Strong advisory and stakeholder management skills * English proficiency for leading technical discussions with global clients and stakeholders ## Description In this role, you will combine deep hands-on expertise in GPU-accelerated AI, Generative AI, and modern AI infrastructure with the opportunity to shape enterprise AI strategies for global clients. You will lead technical engagements from discovery through production, influence architectural decisions, drive AI adoption, and contribute to NVIDIA-focused go-to-market initiatives while collaborating with client stakeholders and multidisciplinary engineering teams., * Lead end-to-end AI engagements, from discovery and solution strategy through architecture design, implementation planning, and production delivery * Translate complex business challenges into AI use cases, solution roadmaps, and scalable enterprise architectures * Design and validate production-ready AI solutions leveraging NVIDIA technologies across cloud and on-premises environments * Define reference architectures for Generative AI and Agentic AI solutions, including GPU infrastructure, Kubernetes orchestration, and inference optimization strategies * Evaluate and optimize AI inference performance by analyzing GPU utilization, latency, throughput, scalability, and infrastructure efficiency * Benchmark and optimize LLM serving frameworks and deployment configurations * Lead pre-sales activities, including technical discovery, workshops, solution positioning, proposals, and proof-of-concept initiatives * Collaborate with NVIDIA stakeholders and internal teams to develop reusable accelerators, solution blueprints, and industry offerings * Drive technical thought leadership through whitepapers, technical content, conference presentations, and mentorship ## Related Videos - [Advanced Typing in TypeScript](https://www.wearedevelopers.com/videos/496-advanced-typing-in-typescript) - [Vectorize all the things! Using linear algebra and NumPy to make your Python code lightning fast.](https://www.wearedevelopers.com/videos/562-vectorize-all-the-things-using-linear-algebra-and-numpy-to-make-your-python-code-lightning-fast) - [Photonic Computing: Programming a New Class of AI Accelerators (incl. Live Coding)](https://www.wearedevelopers.com/videos/100196-photonic-computing-programming-a-new-class-of-ai-accelerators-incl-live-coding) - [Your Next AI Needs 10,000 GPUs. Now What?](https://www.wearedevelopers.com/videos/1590-your-next-ai-needs-10-000-gpus-now-what) - [How to implement convenient Python bindings to C++](https://www.wearedevelopers.com/videos/618-how-to-implement-convenient-python-bindings-to-c) - [AI That Fits Your Business, Not the Other Way Around](https://www.wearedevelopers.com/videos/100148-ai-that-fits-your-business-not-the-other-way-around) ## Related Articles - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Got AI ideas but no money? Here are 10 free ways to level up your AI skills with Google Cloud](https://www.wearedevelopers.com/magazine/600-got-ai-ideas-but-no-money-here-are-10-free-ways-to-level-up-your-ai-skills-with-google-cloud) - [MLOps And AI Driven Development](https://www.wearedevelopers.com/magazine/82-mlops-and-ai-driven-development) - [MLOps – What’s the deal behind it?](https://www.wearedevelopers.com/magazine/125-mlops-what-s-the-deal-behind-it) - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [Best US AI Conferences for CTOs in 2026: Build vs. Buy, Vendor Evaluation, and Peer Intelligence](https://www.wearedevelopers.com/magazine/736-best-us-ai-conferences-for-ctos-in-2026-build-vs-buy-vendor-evaluation-and-peer-intelligence)