Senior Manager, Software Engineering - Agentic IT Operations

NVIDIA Corporation
Santa Clara, CA, United States
1 day ago
Apply on nvidia.wd5.myworkdayjobs.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
5 years minimum
Compensation
$248,000.0
Working hours
Regular working hours

Tech stack

Application Programming Interfaces (APIs) Artificial Intelligence Amazon Web Services Microsoft Azure Code Review Continuous Delivery Continuous Integration DevOps Monitoring of Systems Information Technology Operations Python (Programming Language) Automation of Marketing
+15 more
Reliability Engineering Prometheus Software Engineering Software Systems Systems Integration Datadog Large Language Models Grafana Containerization Kubernetes Production Code Build Process Data Pipelines Pagerduty Servicenow

Job description

As a Senior Engineering Manager, you will own the technical vision, architecture, and delivery of enterprise-scale automation platforms that eliminate manual workflows and enable NVIDIA to operate at 10x scale without proportional headcount growth. You will build and lead a team of IT engineers, set the technical bar, and personally contribute to system design and code. Core responsibilities include:

  • Architect and ship agentic AI systems using LLM-based agents, tool calling, RAG, and orchestration frameworks delivering production-grade AI-assisted operations across enterprise IT domains including employee support, endpoint services, and IT support operations.
  • Design and deploy autonomous AI agents that execute complex, multi-step enterprise workflows end-to-end coordinating approvals, vendor handoffs, cross-system data reconciliation, and exception handling with human-in-the-loop controls delivering measurable improvements in availability, cycle time, cost, and compliance.
  • Engineer robust integration and automation platforms spanning ServiceNow, ERP and procurement systems, endpoint-management platforms, Own the full stack infrastructure, data pipelines, APIs, and user-facing applications.
  • Set the engineering standard through hands-on technical leadership co-authoring production code, conducting rigorous code reviews, and personally driving system design for the most critical components.
  • Recruit, develop, and retain top-tier engineering talent. Build a high-performing team culture grounded in engineering excellence, ownership, and continuous delivery.
  • Define and execute a multi-quarter technical roadmap for automation and agentic operations across enterprise IT, with each initiative tied to quantifiable business outcomes (cost reduction, throughput, SLA improvement, headcount avoidance).
  • Drive disciplined execution-project prioritization, milestone tracking, capacity planning, and on-time delivery-while maintaining engineering velocity in a fast-moving environment.
  • Own talent strategy for the team, including hiring pipelines, performance calibration, and career development that builds a deep bench of engineering leaders.

Requirements

We are seeking a hands-on technical leader to build and lead a high-performance engineering organization that architects, delivers, and operates production-grade software systems at global scale. You will be responsible for transforming enterprise IT operations from manual, reactive workflows into fully automated, AI-driven platforms that scale with NVIDIA’s hyper-growth. This role demands deep software engineering expertise, systems thinking, and the ability to drive large-scale technical transformation with measurable business outcomes. Exceptional interpersonal, written, and verbal communication skills are vital for success., * Bachelor’s or Master’s degree in a related field, or equivalent experience

  • 10+ overall years of hands-on software engineering experience, with deep expertise in at least one of: Infrastructure, SRE, DevOps, or Production Engineering. 5+ years leading engineering teams, with direct experience hiring, growing, and managing IT engineers.
  • Demonstrated ability to build engineering teams from zero and scale them in a high-growth, high-ambiguity environment.
  • Deep expertise in designing and shipping production software systems-including integrations, automation platforms, and data pipelines-for complex enterprise operations at scale.
  • Track record of modernizing enterprise IT operations platforms (e.g., asset management, endpoint services, IT supply chain, infrastructure operations) and deploying agentic AI into production-including multi-step autonomous execution, human-in-the-loop safeguards, exception handling, and governance frameworks with measurable business outcomes.
  • Production-grade proficiency with infrastructure-as-code, CI/CD, containerization (Kubernetes), and cloud platforms (AWS, GCP, or Azure).
  • Experience with monitoring and observability tools (Prometheus, Grafana, Datadog, PagerDuty, or similar).
  • Fluent in Python, Go, or equivalent languages-able to architect, write, and review production-quality code, not just scripts.
  • Executive-level communication skills with the ability to influence technical direction across engineering, product, and senior leadership.
  • Proven ability to translate complex technical capabilities into quantifiable business value and present to VP/C-level audiences.

Benefits & conditions

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 248,000 USD - 391,000 USD.

About the company

For over 25 years, NVIDIA has been at the forefront of transforming computer graphics, PC gaming, and accelerated computing, driven by a legacy of continuous innovation and exceptional talent. We are now leveraging the immense potential of AI to usher in the next era of computing, where our GPUs power the “brains” of computers, robots, and autonomous vehicles that can comprehend the world! This pioneering work demands vision, innovation, and the world’s best talent. Join our diverse and supportive environment, where NVIDIANs are inspired to excel and make a profound global impact., NVIDIA is widely considered to be one of the technology world’s most desirable employers. We have some of the most forward-thinking and hardworking people in the world working for us. If you’re creative and autonomous, we want to hear from you!

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on nvidia.wd5.myworkdayjobs.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:17 min

Mapping the maturity roadmap for scaled devops adoption

Dominik Krichbaum Dominik Krichbaum · World Congress 2026 Europe

10:40 min

Visualizing Prometheus open metrics using custom Grafana dashboards

Stijn Polfliet · LIVE

1:36 min

Visualizing memory limits and isolating suspicious endpoints

Dina Matveev Dina Matveev · Europe 2026 Virtual

2:33 min

Architecting CUDA and the AI software stack

Michael Kagan Michael Kagan +1 · World Congress 2026 Europe

3:18 min

Scaling global network engineering through DevOps culture

Stuart Clark · LIVE

4:36 min

Hiring passionate software engineers to tackle unprecedented scaling challenges

Dana Lawson Dana Lawson +1 · World Congress 2026 Europe

Videos

See all

Related articles

See all