> Markdown version of [/jobs/ext/2590960-ml-and-ai-knowledge-systems-engineer](https://www.wearedevelopers.com/jobs/ext/2590960-ml-and-ai-knowledge-systems-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # ML and AI Knowledge Systems Engineer - **Company:** Advanced Micro Devices, Inc. - **Location:** San Jose, CA, United States - **Salary:** $240,000.0 - **Contract:** Permanent contract - **Skills:** Application Programming Interfaces (APIs), Artificial Intelligence, Systems Engineering, Automated Storage and Retrieval Systems, Confluence, JIRA, Compilers, Encodings, Computer Engineering, Microarchitecture, Software Debugging, Firmware, Hardware Design, Information Retrieval, Knowledge-Based Systems, Machine Learning, Natural Language Processing, Operational Databases, Performance Tuning, Regression Testing, Graphics Processing Unit (GPU), Large Language Models, Indexer, Information Technology, Free and Open-Source Software, Search Engines - **Published:** August 21, 2026 - **Apply:** https://jobs.localjobnetwork.com/apply/add/88108498/1 ## About the Role * Deep relevant experience building production ML, search, ranking, or information-retrieval systems. * Hands-on experience with RAG, embeddings, vector and keyword search, reranking, chunking, and retrieval evaluation. * Practical experience with LLM prompting, structured generation, tool calling, evaluation, and hallucination reduction. * Experience designing ML evaluation datasets, metrics, experiments, and regression tests. * Demonstrated ability to translate current ML or information-retrieval research into rigorous experiments and production improvements. * Strong software and systems engineering skills, including APIs, asynchronous processing, observability, and production operations. * Experience leading architecture, mentoring engineers, and influencing senior technical stakeholders. * Experience with distributed inference systems such as vLLM and GPU performance optimization. * Experience with AMD Instinct GPUs or comparable accelerators. * Experience with vector databases, hybrid search, distributed indexing, or authorization-aware retrieval. * Familiarity with MCP, coding agents, or tool-based agent architectures. * Experience with hardware design, EDA, RTL, microarchitecture, firmware, compilers, or chip verification. * Technical thought leadership demonstrated through publications, patents, open-source contributions, conference participation, or influential production systems. Publications are valued but not required. PREFERRED ACADEMIC CREDENTIALS: * Master's or Doctoral degree in machine learning, information retrieval, NLP, computer science, electrical or computer engineering, or a related field. ## Description We are building GoldenEye, an internal AI knowledge platform that lets engineers ask natural-language questions across GPU design knowledge, including RTL, microarchitecture specifications, verification collateral, Confluence, and JIRA, and receive trustworthy answers with citations. A prototype is already used by hardware and software teams. We are now scaling it to thousands of users, with the potential to serve more than 40,000 engineers worldwide. We are seeking a hands-on PMTS-level ML engineer to lead GoldenEye's retrieval, ranking, and answer-quality architecture. You will turn a promising prototype into an authoritative production platform while setting technical direction, mentoring engineers, and partnering with hardware, software, infrastructure, and security teams. GoldenEye will make critical engineering knowledge easier to find, verify, and use, accelerating GPU design, verification, debugging, and bring-up. You will serve as the technical anchor for its ML core and shape its adoption across the engineering organization., * Own the architecture for embeddings, chunking, hybrid retrieval, reranking, multi-query planning, and reciprocal-rank fusion. * Build retrieval workflows across specifications, RTL, verification artifacts, wikis, and issue trackers. * Improve search for domain-specific content such as ISA mnemonics, registers, signal names, acronyms, and code identifiers. * Reduce hallucinations through grounded generation, citation enforcement, source validation, and conflict resolution. * Define source-authority and freshness policies for conflicting or outdated information. * Build an evaluation framework for retrieval relevance, answer correctness, citation faithfulness, latency, cost, and regressions. * Use expert feedback, production data, hard negatives, and targeted failure cases to create representative evaluation datasets. * Track advances in retrieval, RAG, agentic systems, evaluation, and efficient inference, and evaluate promising techniques against production requirements. * Optimize embedding, reranking, and inference workloads on AMD Instinct GPUs, including multi-GPU and multi-node serving. * Partner with platform teams on scalable services, indexing, data reconciliation, observability, and cluster orchestration. * Advance Model Context Protocol (MCP) tools used by coding agents. * Ensure retrieval, caching, and answer generation respect per-user access controls. * Mentor engineers, review designs, and communicate technical decisions across organizations., AMD may use Artificial Intelligence to help screen, assess or select applicants for this position. AMD's "Responsible AI Policy" is available here. ## Related Videos - [Give Your LLMs a Left Brain](https://www.wearedevelopers.com/videos/1160-give-your-llms-a-left-brain) - [Improving quality with Agentic AI with Rovo Dev and Xray](https://www.wearedevelopers.com/videos/2005-improving-quality-with-agentic-ai-with-rovo-dev-and-xray) - [Playing Pong on a shoulder press machine](https://www.wearedevelopers.com/videos/100140-playing-pong-on-a-shoulder-press-machine) - [Optimizing Discovery: PostgreSQL's Role in Transforming GetYourGuide's Search](https://www.wearedevelopers.com/videos/1647-optimizing-discovery-postgresql-s-role-in-transforming-getyourguide-s-search) - [Coffee with Developers - Maria Apazoglou](https://www.wearedevelopers.com/videos/1209-coffee-with-developers-maria-apazoglou) - [Collaboration Quantified: Lessons from Open Source Developer Networks](https://www.wearedevelopers.com/videos/1422-collaboration-quantified-lessons-from-open-source-developer-networks) ## Related Articles - [What Are Large Language Models?](https://www.wearedevelopers.com/magazine/304-what-are-large-language-models) - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [MLOps And AI Driven Development](https://www.wearedevelopers.com/magazine/82-mlops-and-ai-driven-development) - [MLOps – What’s the deal behind it?](https://www.wearedevelopers.com/magazine/125-mlops-what-s-the-deal-behind-it) - [Got AI ideas but no money? Here are 10 free ways to level up your AI skills with Google Cloud](https://www.wearedevelopers.com/magazine/600-got-ai-ideas-but-no-money-here-are-10-free-ways-to-level-up-your-ai-skills-with-google-cloud) - [Graph and AI Trends 2026: Why Is AI Running but Not Yet Delivering?](https://www.wearedevelopers.com/magazine/680-graph-and-ai-trends-2026-why-is-ai-running-but-not-yet-delivering)