Network Systems Engineer

Super Micro Computer, Inc.
San Jose, CA, United States
5 days ago
Apply on www.indeed.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience required
1 year minimum
Compensation
$120,000.0 - $140,000.0
Working hours
Regular working hours
Job source

Tech stack

Microsoft Windows Artificial Intelligence Amazon Web Services Microsoft Azure BIOS Software Quality Nvidia CUDA Computer Programming Computer Engineering Data Centers Software Debugging Linux
+20 more
DevOps Programming Tools Networking Hardware Machine Learning Network Configuration and Change Management Routing OpenShift OpenStack Software Reliability Testing Shell Script Planning Software Graphics Processing Unit (GPU) Computer Networking Systems Computer Network Technologies Large Language Models Deep Learning Kubernetes Information Technology Slurm Docker

Job description

As a Network Engineer, you will be assisting roll out and maintain business critical applications and services for Supermicro. You will work with senior engineer to resolve escalated service issues, work with other engineers to resolutions, engineering and implementing complex projects., * Execute system-level rack tests on latest NVidia and AMD GPUs, ARM-based, Intel Xeon, and AMD EPYC processors, encompassing functionality, compatibility, performance, stress, and reliability testing, leveraging proprietary in-house tools.

  • Familiar with HPC/AI applications and benchmarks, address customer support issues, demonstrating innovative problem-solving skills and building robust processes and procedures for HPC/AI solutions.
  • May work on conduct proof of concept design and testing, providing optimized benchmarks for HPC/AI applications in a timely manner. Fine-tune BIOS settings, optimize OS/network configurations, and develop diverse simulation configurations to enhance efficiency across various workloads.
  • Deliver on-site deployment services, ensuring customer acceptance verification and providing post-level 1&2 support. Create and maintain technical documentation, including technical notes, blogs, and diagrams, to facilitate knowledge dissemination.
  • Identify and document hardware and software quality issues and collaborate with Product Management and other Engineering teams to integrate customer feedback into future product enhancements.
  • Proactively engage in HPC roadmap development, planning software and hardware upgrades to sustain exceptional HPC infrastructure performance.
  • Document and analyze test plans, reports, logs, and actively contribute to the development of test utilities and automation scripts to streamline testing processes.

Requirements

  • BS/MS in Electrical Engineering, Computer Engineering or Computer Science
  • 1+ years of work-related experience in Deep Learning and Machine Learning
  • Familiar with Linux/networking debugging/testing or relevant experience preferred
  • Familiar with data center, enterprise, or telecommunication working on routing and switching networking technologies.
  • Knowledge with DevOps or in cloud environments, including but not limited to Docker/Containers and Kubernetes
  • Hands-on experience with workload/scheduler Managers (Slurm) for rack/cluster
  • Familiar with MLPerf Training/Inference benchmark, LLM, HPL-AI or RCCL/NCCL
  • Programming experience with windows and Linux shell scripting
  • Strong sense of teamwork and good team player, strong communication skills

Desired Skills:

  1. Familiar with Intel/AMD/NVIDIA development tool kits such as CUDA, oneAPI, ROCm
  2. Relevant certifications such as CCIE, JNCIE, or Arista ACE are highly desirable
  3. Experience with server/network hardware debugging and troubleshooting
  4. CCNA, OpenStack, OpenShift, Azure or AWS

Benefits & conditions

$120,000 - $140,000

The salary offered will depend on several factors, including your location, level, education, training, specific skills, years of experience, and comparison to other employees already in this role. In addition to a comprehensive benefits package, candidates may be eligible for other forms of compensation, such as participation in bonus and equity award programs.

About the company

Supermicro® is a Top Tier provider of advanced server, storage, and networking solutions for Data Center, Cloud Computing, Enterprise IT, Hadoop/ Big Data, Hyperscale, HPC and IoT/Embedded customers worldwide. We are the #5 fastest growing company among the Silicon Valley Top 50 technology firms. Our unprecedented global expansion has provided us with the opportunity to offer a large number of new positions to the technology community. We seek talented, passionate, and committed engineers, technologists, and business leaders to join us.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.indeed.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:22 min

Infrastructure barriers and compliance risks in research

Jeremy Murray Jeremy Murray · World Congress 2026 Europe

41 sec

Massive client data loss and bio-digital storage

Chris Heilmann Chris Heilmann +1 · LIVE

52 sec

Running persistent Linux environments directly on Windows

Ben Breard Ben Breard · World Congress 2025

2:04 min

Insights on transitioning from supercomputing to technical education

Andrew Holway · LIVE

1:34 min

Pivoting careers into specialized platform engineering roles

Xavier Portilla Edo · LIVE

1:51 min

Managing GPU quotas and multi-tenancy with Kueue

Jeremy Murray Jeremy Murray · World Congress 2026 Europe

Videos

See all

Related articles

See all