> Markdown version of [/jobs/ext/1242199-llm-platform-engineer](https://www.wearedevelopers.com/jobs/ext/1242199-llm-platform-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # LLM Platform Engineer - **Company:** Whatnot Inc. - **Location:** San Francisco, CA, United States (Remote available) - **Salary:** $245,000.0 - $345,000.0 - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Amazon Web Services, Amazon Elastic Compute Cloud, Amazon S3, Automated Storage and Retrieval Systems, Databases, Continuous Integration, Data Visualization, Amazon DynamoDB, Elasticsearch, Python (Programming Language), PostgreSQL, Machine Learning, Redis, Software Engineering, Datadog, Data Logging, Cloud Platform System, Large Language Models, Grafana, Information Technology, Apache Flink, Apache Kafka, Machine Learning Operations, Functional Programming - **Published:** July 11, 2026 - **Apply:** https://www.indeed.com/viewjob?jk=26363a8e2499a65a ## About the Role As our next AI/ML Engineer you should have 4+ years of professional experience developing machine learning systems and algorithms, plus: * Bachelor's degree in Computer Science, Statistics, Applied Mathematics or a related technical field, or equivalent work experience. * 3+ years of software engineering experience building and maintaining production systems for consumer-scale loads. * 1+ years of professional experience developing software in Python * Ability to work autonomously and drive initiatives across multiple product areas and communicate findings with leadership and product teams. * Experience with operational, search, and key-value databases such as PostgreSQL, DynamoDB, Elasticsearch, Redis. * Firm grasp of visualization tools for monitoring and logging e.g. DataDog, Grafana. * Familiarity with cloud computing platforms and managed services such as AWS Sagemaker, Lambda, Kinesis, S3, EC2, EKS/ECS, Apache Kafka, Flink. * Professionalism around collaborating in a remote working environment and well tested, reproducible work. * Exceptional documentation and communication skills. ## Description We're looking for builders-intellectually curious, highly entrepreneurial engineers eager to shape the future of AI and ML at Whatnot. You'll design and scale the core infrastructure that powers large language model applications across the company, working side by side with machine learning scientists to bring cutting-edge models into production and unlock entirely new product experiences. This means building systems that make AI dependable and fast at scale-from building retrieval systems to more effectively ground LLM responses in Whatnot's business context to developing scalable LLM evaluation frameworks and human-in-the-loop feedback mechanisms., * Own the infrastructure powering LLMs across critical business surfaces- supporting growth, recommendations, trust and safety, fraud, seller tooling, and more. * Create robust and scalable LLM evaluation frameworks to measure model performance, guide iteration, and prevent regression via CI/CD. * Deploy RAG systems and MCP servers to more effectively ground LLM responses in Whatnot's business context while enforcing rigorous PII controls. * Design efficient human-in-the-loop feedback pipelines that can be used to inform scalable LLM evaluation * Bridge the gap between research and production, helping to transform experimental ideas into scalable solutions * Stretch beyond your comfort zone to take on new technical challenges as we scale AI across Whatnot's ecosystem. US Based: We offer flexibility to work from home or from one of our global office hubs, and we value in-person time for planning, problem-solving, and connection. Team members in this role must live within commuting distance of our New York, Seattle, Los Angeles, and San Francisco hubs. ## Related Videos - [Alibaba Big Data and Machine Learning Technology](https://www.wearedevelopers.com/videos/37-alibaba-big-data-and-machine-learning-technology) - [5 steps for running a Kubernetes environment at scale](https://www.wearedevelopers.com/videos/88-5-steps-for-running-a-kubernetes-environment-at-scale) - [Debugging in the Dark](https://www.wearedevelopers.com/videos/1658-debugging-in-the-dark) - [Reducing LLM Calls with Vector Search Patterns - Raphael De Lio (Redis)](https://www.wearedevelopers.com/videos/1714-reducing-llm-calls-with-vector-search-patterns-raphael-de-lio-redis) - [Accelerating Authentication Architecture: Taking Passwordless to the Next Level](https://www.wearedevelopers.com/videos/733-accelerating-authentication-architecture-taking-passwordless-to-the-next-level) - [Software Engineering Social Connection: Yubo’s lean approach to scaling an 80M-user infrastructure](https://www.wearedevelopers.com/videos/1583-software-engineering-social-connection-yubo-s-lean-approach-to-scaling-an-80m-user-infrastructure) ## Related Articles - [What Are Large Language Models?](https://www.wearedevelopers.com/magazine/304-what-are-large-language-models) - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [MLops – Deploying, Maintaining And Evolving Machine Learning Models in Production](https://www.wearedevelopers.com/magazine/115-mlops-deploying-maintaining-and-evolving-machine-learning-models-in-production) - [MLOps And AI Driven Development](https://www.wearedevelopers.com/magazine/82-mlops-and-ai-driven-development) - [MLOps – What’s the deal behind it?](https://www.wearedevelopers.com/magazine/125-mlops-what-s-the-deal-behind-it)