> Markdown version of [/jobs/ext/3460439-senior-ai-infrastructure-engineer-eda-infrastructure](https://www.wearedevelopers.com/jobs/ext/3460439-senior-ai-infrastructure-engineer-eda-infrastructure). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Senior AI infrastructure engineer - EDA Infrastructure - **Company:** NVIDIA Corporation - **Location:** Santa Clara, CA, United States - **Experience:** Expert - **Salary:** $184,000.0 - $287,500.0 - **Contract:** Permanent contract - **Skills:** Java (Programming Language), Artificial Intelligence, Software as a Service, Configuration Management Databases, Computer Engineering, Data Infrastructure, Python (Programming Language), Machine Learning, Operational Data Store, Cloud Services, Software Engineering, TypeScript, Workflow Management Systems, AI Infrastructure, AI Platforms, Information Technology, Virtual Agents, Data Pipelines, Golang, Programming Languages - **Published:** September 1, 2026 - **Apply:** https://www.jofdav.com/jobs/59518610-senior-ai-infrastructure-engineer-eda-infrastructure ## About the Role * BS degree in Computer Science, Computer Engineering, or a related technical field, or equivalent experience. * 8+ years of experience in infrastructure security, platform engineering, or security tooling. * Proficiency in one or more programming languages such as Python, Go, Typescript, or Java. * Strong understanding of software and infrastructure principles, with experience applying them in production environments. * Ability to lead cross-functional initiatives that span internal teams and external partners in varying disciplines across engineering, product, finance, and security. Ways to stand out from the crowd: * Experience building and operating incident-management processes with internally built and/or externally vended SaaS tools * Experience working with building, deploying, and maintaining ML models in production systems along with familiarity with AI agent frameworks or orchestration tools * Experience building and operating modern observability platforms to deliver scalable metrics, logs, traces, and profiling Experience working with service catalog and configuration management databases (CMDB) ## Description AI Infrastructure Engineers at NVIDIA build the systems, tooling, and data infrastructure that enable operation of our GPU cloud services. We are enabling engineering teams to innovate while proactively identifying, tracking, and mitigating risks across the entire technical task. This role is ideal for engineers who thrive at the intersection of product, infrastructure, and software engineering and who want to build automated, intelligence-driven systems that protect NVIDIA's most critical AI platforms. What you'll be doing: Telemetry * Build and operate scalable telemetry pipelines for metrics, logs, traces, and events across on-premise, CSP, and NCP clusters. * Establish common instrumentation, collection, storage, and access patterns so teams can generate and consume telemetry consistently. * Deliver dashboards, alerting, and analysis capabilities that improve service visibility, detection, and troubleshooting. Operational Excellence * Standardize and automate incident, maintenance, service on-call, and support on-call workflows across HWInf. * Integrate operational data and lifecycle signals to improve ownership, escalation, communication, and post-incident learning. * Build reporting and AI-assisted tooling that reduces manual toil and improves operational responsiveness. Cataloging and Inventory * Build and maintain physical hardware and software catalogs as trusted sources of truth for infrastructure inventory, service ownership, dependencies, and documentation. * Create consistent data models and integration pipelines that connect clusters, hardware, services, teams, and operational workflows. * Provide self-service discovery capabilities so engineers can quickly identify what they operate, who owns it, and how to support it. ## Related Videos - [Pioneering AI Assistants in Banking](https://www.wearedevelopers.com/videos/1627-pioneering-ai-assistants-in-banking) - [Go with the Flow: Stop the Leaks Before Your Memory's a Waterfall!](https://www.wearedevelopers.com/videos/100073-go-with-the-flow-stop-the-leaks-before-your-memory-s-a-waterfall) - [Do TypeScript without TypeScript](https://www.wearedevelopers.com/videos/327-do-typescript-without-typescript) - [This App Reached 10,000 Users in One Week. Here's How.](https://www.wearedevelopers.com/videos/100329-this-app-reached-10-000-users-in-one-week-here-s-how) - [From AI Assistance to Agentic Systems: Scaling Sovereign AI in Banking](https://www.wearedevelopers.com/videos/100070-from-ai-assistance-to-agentic-systems-scaling-sovereign-ai-in-banking) - [Building the Nervous System of AI - Michael Kagan (NVIDIA)](https://www.wearedevelopers.com/videos/2133-building-the-nervous-system-of-ai-michael-kagan-nvidia) ## Related Articles - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [Stephan Gillich - Bringing AI Everywhere](https://www.wearedevelopers.com/magazine/489-stephan-gillich-bringing-ai-everywhere) - [Navigating the AI Shift](https://www.wearedevelopers.com/magazine/629-navigating-the-ai-shift) - [Got AI ideas but no money? Here are 10 free ways to level up your AI skills with Google Cloud](https://www.wearedevelopers.com/magazine/600-got-ai-ideas-but-no-money-here-are-10-free-ways-to-level-up-your-ai-skills-with-google-cloud) - [Best US AI Conferences for CTOs in 2026: Build vs. Buy, Vendor Evaluation, and Peer Intelligence](https://www.wearedevelopers.com/magazine/736-best-us-ai-conferences-for-ctos-in-2026-build-vs-buy-vendor-evaluation-and-peer-intelligence) - [What Industries Outside of AI Are Hiring The Most AI Experts?](https://www.wearedevelopers.com/magazine/98-what-industries-outside-of-ai-are-hiring-the-most-ai-experts)