> Markdown version of [/jobs/ext/2254086-principal-software-engineer-infrastructure](https://www.wearedevelopers.com/jobs/ext/2254086-principal-software-engineer-infrastructure). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Principal Software Engineer - Infrastructure - **Company:** NVIDIA Corporation - **Location:** Santa Clara, CA, United States - **Salary:** $248,000.0 - **Contract:** Permanent contract - **Skills:** Java (Programming Language), Application Programming Interfaces (APIs), Artificial Intelligence, Big Data, Cloud Computing, Configuration Management Databases, Configuration Management, Continuous Integration, Data Centers, Software Debugging, Linux, Disaster Recovery, Distributed Systems, Python (Programming Language), Key Management, Cloud Services, Ansible, Software Engineering, System Programming, Enterprise Data Management, Cloud Platform System, Delivery Pipeline, Concurrency, Kubernetes, Infrastructure Automation Frameworks, Information Technology, Databricks, Golang - **Published:** August 26, 2026 - **Apply:** https://nvidia.wd5.myworkdayjobs.com/NVIDIAExternalCareerSite/job/US-CA-Santa-Clara/Principal-Software-Engineer---Infrastructure_JR2024352-1 ## About the Role * 15+ years of progressive software engineering experience, with a sustained record of delivering complex, business-critical platforms and distributed systems. * Bachelor's or Master's degree in Computer Science, Engineering, or a similar domain, or a Master's degree or equivalent experience. * Deep knowledge of enterprise configuration management or infrastructure automation using technologies such as Ansible Automation Platform, AWX, Salt or equivalent platforms. * Deep hands-on expertise in Go, Python, Java, or a comparable systems programming language, including APIs, concurrency, distributed systems, testing, debugging, and performance engineering. * Strong hands-on experience in Linux, Kubernetes, containers, cloud and hybrid infrastructure, CI/CD and Infrastructure as Code. * Strong experience developing and implementing enterprise data and automation pipelines using databricks or comparable large-scale data platforms. * Proven experience defining, owning, and evolving the architecture of large-scale infrastructure platforms operating across multiple teams, data centers, or cloud environments. * Demonstrated ability to identify organization-wide business and technical challenges, establish a clear strategy, and drive implementation across teams. * Outstanding communication and technical leadership skills, with a proven ability to build consensus, influence senior leaders, mentor experienced engineers, and lead through ambiguity. Ways to Stand Out from the Crowd: * Experience in architecting, deploying and managing Ansible Automation Platform at enterprise scale. * Experience developing platforms that manage large global infrastructure fleets across data centers, public clouds, compute, storage, and networking environments. * A history of leading the development and enterprise-wide adoption of an AI-enabled configuration management, infrastructure automation, or autonomous operations platform. * Experience using AI to solve infrastructure challenges such as configuration drift, compliance, change-risk analysis, incident diagnosis, capacity management, predictive operations, or automated remediation. ## Description * Define the multi-year technical vision and architecture for IT infrastructure automation, configuration management, orchestration, and self-service platforms. * Set the technical direction for infrastructure automation and configuration management across technologies such as Ansible Automation Platform, AWX, Salt or equivalent platforms. * Establish the strategy for applying AI to configuration management, including configuration intelligence, drift and compliance analysis, change-risk identification, root-cause assistance, intelligent recommendations, and guarded automated remediation. * Remain deeply hands-on by developing prototypes and production software, reviewing critical code and designs, and resolving the most challenging technical and scalability problems. * Build secure and scalable integrations across infrastructure platforms, cloud services, CMDB, secrets management, observability, and enterprise data systems. * Set and drive configuration management strategy across networking, storage, and compute domains. Bring deep cross-domain technical expertise to identify gaps and challenge current methods. Establish architectures and engineering standards. Lead organizations in implementation and adoption within large-scale environments. * Ensure platforms meet enterprise requirements for availability, scalability, performance, security, disaster recovery, observability, and operational support. * Act as a force multiplier by mentoring senior and staff engineers, raising engineering standards, facilitating architectural decisions, and developing technical leaders across teams. * Partner with engineering and executive leadership to translate critical business challenges into technical strategy, prioritized roadmaps, and measurable business outcomes. ## Related Videos - [The Software Engineer 2030: From Coder To AI Orchestrator? - Patrick Schnell](https://www.wearedevelopers.com/videos/1825-the-software-engineer-2030-from-coder-to-ai-orchestrator-patrick-schnell) - [Dev & Test in the Cloud? Deploy your cloud environments with Ansible & Terraform](https://www.wearedevelopers.com/videos/1607-dev-test-in-the-cloud-deploy-your-cloud-environments-with-ansible-terraform) - [Go with the Flow: Stop the Leaks Before Your Memory's a Waterfall!](https://www.wearedevelopers.com/videos/100073-go-with-the-flow-stop-the-leaks-before-your-memory-s-a-waterfall) - [Docker network without Docker](https://www.wearedevelopers.com/videos/1418-docker-network-without-docker) - [#90DaysOfDevOps - The DevOps Learning Journey](https://www.wearedevelopers.com/videos/548-90daysofdevops-the-devops-learning-journey) - [Shipping Faster with Less: Render on Cloud Hosting, AI Workloads, and the Future of DevOps](https://www.wearedevelopers.com/videos/1894-shipping-faster-with-less-render-on-cloud-hosting-ai-workloads-and-the-future-of-devops) ## Related Articles - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [What is Software Engineering in the Age of AI?](https://www.wearedevelopers.com/magazine/640-what-is-software-engineering-in-the-age-of-ai) - [Navigating the AI Shift](https://www.wearedevelopers.com/magazine/629-navigating-the-ai-shift) - [Best US AI Conferences for CTOs in 2026: Build vs. Buy, Vendor Evaluation, and Peer Intelligence](https://www.wearedevelopers.com/magazine/736-best-us-ai-conferences-for-ctos-in-2026-build-vs-buy-vendor-evaluation-and-peer-intelligence) - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again)