Senior Software Engineer, Cloud Automation

NVIDIA Ltd.
Santa Clara, United States of America
yesterday

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Working hours
Regular working hours
Languages
English
Experience level
Senior
Compensation
$ 319K

Job location

Remote
Santa Clara, United States of America

Tech stack

Amazon Web Services (AWS)
Automation of Tests
Bash
Cloud Computing
Cloud Database
Continuous Integration
Data Centers
Software Debugging
DevOps
Groovy
Image Management
Python
Network Service
Ansible
Workflow Management Systems
Data Logging
Scripting (Bash/Python/Go/Ruby)
Cloud Platform System
Multi-Agent Systems
Software Troubleshooting
Multi-Cloud
Kubernetes
Infrastructure Automation Frameworks
Information Technology
Data Analytics
Terraform
Multiplatform
Docker
Jenkins
Go
Microservices

Job description

NVIDIA Cloud platform engineering is seeking exceptionally talented DevOps infrastructure automation engineers to deliver our NVIDIA GPU Cloud services. Join a team that accomplishes over 250K automation runs per week, operating 24x7 with thousands of automation projects. This outstanding opportunity will allow you to work on multi-OSes, multi-cloud platforms, and virtualization technologies simultaneously.

Our team is ambitious, constantly seeking to improve our automation infrastructure through new DevOps and cloud technologies, processes, and tools. You will collaborate closely with application developers and QA to build flawless automation infrastructure for deployment, testing, monitoring, and CI/CD operations.

What You'll Be Doing

  • Collaborate with Configuration Management tools and Workflow management tools such as Ansible and Stackstorm to handle and deploy COLO multi-platform server clusters in NVIDIA Data Centers worldwide.
  • Build and develop infrastructures, tools, and automation scripts to improve automation support for various activities such as image management, switch management, deployment, data analytics, automated testing, logging, monitoring, and alerts for different micro-services.
  • Apply agentic AI to develop AI orchestration skills and manage Data Center operations at scale using infrastructure tools like netbox, foreman, and Mellanox switches.
  • Employ the latest infrastructure management technologies like Kubernetes and Docker to ensure fast and consistent delivery.
  • Own the infrastructure for micro-services and provide operational support to application teams, focusing on automation service, infrastructure and security improvement, and live service troubleshooting.

Requirements

  • BS in Computer Science/Engineering/Math/Physics or equivalent experience
  • 3+ years of proven experience
  • Outstanding agentic AI experience with Codex, Claude code
  • Proficient in scripting languages: Python, Bash, Groovy, GOLANG
  • Outstanding debugging skills
  • Cloud experience with AWS Compute, Containers, and Networking services is preferable.
  • CI/CD experience with Jenkins and Jenkins pipeline
  • Experience working with Configuration Management tools such as Ansible is highly advantageous.
  • Experience with Packer, Terraform, and StackStorm is advantageous.
  • Experience with Kubernetes, Docker, and Helm is highly desirable.

Benefits & conditions

Widely considered to be one of the technology world's most desirable employers, NVIDIA offers highly competitive salaries and a comprehensive benefits package. As you plan your future, see what we can offer to you and your family www.nvidiabenefits.com/

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. For Poland: The base salary range is 183,750 PLN - 318,500 PLN for Level 3, and 240,000 PLN - 416,000 PLN for Level 4. , , JR2019580

About the company

NVIDIA is widely considered one of the technology world's most desirable employers. We have some of the most experienced and talented individuals working for us. If you're creative and autonomous, we want to hear from you!

Apply for this position