Solutions Architect, Data Center Infrastructure - NVIS
- Discuss this with your agent
- Open in Claude
- Open in ChatGPT
Role details
Tech stack
+3 more
Job description
NVIDIA is seeking a Solutions Architect in Data Center Infrastructure to join our Infrastructure Specialists team. Academic and commercial groups worldwide are using NVIDIA products to redefine deep learning, data analytics, and power data centers. Join the team building many of the world’s largest and fastest data centers and supercomputers! NVIDIA is looking for someone who can lead planning and deployments of AI data centers including power/cooling systems, cabling and network provisioning and bring-up/validation.
As the NVIS Solutions Architect for Datacenter Infrastructure, you will focus on data center audit, planning and deployment ensuring the integrity of NVIDIA platform infrastructure. Your primary goal will be to guarantee that all aspects of the data center’s physical infrastructure are meticulously planned, implemented, and validated to meet NVIDIA reference architectures, operational requirements, and industry standards. This infrastructure includes architectural systems, power distribution, liquid/air cooling systems, compute, network and cabling (fiber and copper), and telemetry systems.
What you will be doing:
- NVIS Datacenter Engineering and planning: Collaborate with other teams to plan and implement data center infrastructure solutions based on NVIDIA Datacenter reference architecture, including power distribution, cooling systems, network architecture, server hardware, and storage systems.
- Plan and manage deployment of NVIDIA’s pioneering AI infrastructure solutions including highly complex rack-scale, liquid cooled compute and networking hardware systems, in a fluid and fast paced environment.
- Conduct pre-deployment planning including reviewing cluster and data center architecture, plan network port mapping and fiber optic cabling BOM, identify potential risks, train vendors and find areas for improvement.
- Evaluate customers’ and partners’ infrastructure design proposals for consistency with industry standards and regulatory requirements. Provide feedback and recommendations to improve performance, scalability, and cost-effectiveness.
- Perform testing, troubleshooting and validation of compute systems based on collaboration with product and engineering teams.
- Act as the NVIS mentor providing guidance, mentorship, and support to ensure the NVIS team’s success in their respective roles.
- Quality Assurance: Establish and enforce quality assurance processes to verify that deployments meet established specifications and performance benchmarks. Conduct thorough bring-up, testing, and validation to validate the functionality and reliability of infrastructure components.
- Continuous Improvement: Drive continuous improvement initiatives to enhance data center infrastructure efficiency for NVIDIA data center reference architecture and deployment blueprint, resilience, and sustainability. Find opportunities to streamline processes, automate repetitive tasks, and leverage emerging technologies to optimize infrastructure operations.
- Collaboration and Communication: Collaborate and communicate across internal teams, external vendors, and customers to facilitate the seamless integration of data center infrastructure solutions. Serve as a domain expert and point of contact for infrastructure-related inquiries and blocking issues.
Requirements
Do you have experience in Team leadership?, * Bachelor’s degree (or equivalent experience) in Engineering, Computer Science, Information Technology, or a related field.
- Minimum 3+ years of overall experience in enterprise and/or hyperscale data centers with continual infrastructure deployment experience, preferably for high density AI/HPC data centers.
- Working experience in data center operations, or infrastructure management roles, focusing on large-scale data center deployments.
- Strong technical knowledge and experience in the data center stack - power distribution, liquid cooling, servers, networking, storage and pre-deployment planning
- Relevant certification - preferred
- Demonstrated technical and project leadership under fluid situations, ability to adapt to unknowns and change.
- Excellent analytical, problem-solving, and decision-making skills, keen attention to detail, and a commitment to quality.
- Excellent communication and interpersonal abilities, capable of engaging with various collaborators like customers to enable productive discussions.
- Organization & Time Management - able to plan, schedule, and organize tasks related to the job to achieve goals within or ahead of established time frames.
- Willingness to travel (up to 40%).
Way to stand out from the crowd:
- Linux system administration skills
- Strong knowledge of whole data center Infrastructure stack
- Flexible/agile and enjoys solving challenging problems
Benefits & conditions
Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 124,000 USD - 195,500 USD for Level 3, and 148,000 USD - 235,750 USD for Level 4.
About the company
NVIDIA is widely considered one of the world’s most desirable employers in technology. We have some of the world’s most forward-thinking and passionate people working for us. If you’re creative and autonomous, we want to hear from you!
Apply for this position
This job is hosted externally. Click below to view the full posting and apply.
Apply on indeed.comGood distractions
Talks and stories from around this role — technically off-topic, practically not.
Moments
Explore playlistsVideos
See allRelated articles
See all
Highest Paying Tech Companies for Developers
7 Cloud Computing Trends Coming in 2025 for Developers
What Industries Outside of AI Are Hiring The Most AI Experts?
Got AI ideas but no money? Here are 10 free ways to level up your AI skills with Google Cloud