> Markdown version of [/jobs/ext/2706578-network-engineer](https://www.wearedevelopers.com/jobs/ext/2706578-network-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Network Engineer - **Company:** NuScale Power Corporation - **Location:** United States - **Experience:** Expert - **Salary:** $150,000.0 - $210,000.0 - **Contract:** Permanent contract - **Skills:** Link Aggregation (Ethernet), Artificial Intelligence, Border Gateway Protocol, Common Lisp Object Systems, Cloud Computing, Cyber Security, Continuous Integration, Data Centers, Network Address Translation, Ethernet, Github, InfiniBand, Subnetting, Virtual Private Networks (VPN), Python (Programming Language), Overlay Transport Virtualization, Remote Direct Memory Access, Ansible, Wide Area Networks, High Performance Computing, Juniper, Gitlab-ci, Git Flow, Low Latency, Terraform, Open Network Automation Platform - **Published:** September 4, 2026 - **Apply:** https://startup.jobs/senior-network-engineer-nscale-8335933 ## About the Role * 8+ years of network engineering experience, with depth in HPC, AI, or hyperscale data centre environments. * Extensive hands-on experience with RDMA-aware networking (InfiniBand, RoCE) for AI/HPC workloads, including subnet managers (OpenSM, UFM) and fabric orchestration. * Expert-level knowledge of modern DC routing and control planes (BGP, EVPN-VxLAN, Clos/spine-leaf), with production experience on platforms such as Cumulus, Nokia, or Arista EOS.Strong automation skills: Python and Ansible, Git-based workflows, and familiarity with modern IaC and pipeline tooling (Terraform, GitLab CI/GitHub Actions); you treat the network as code rather than managing devices by hand. * Strong design and engineering experience with firewall platforms - Juniper SRX and/or Palo Alto - including security policy architecture, HA design, and multi-tenant segmentation. * Experience designing network telemetry and observability for high-throughput environments. * Proven ability to work cross-functionally with systems, storage, and HPC/AI workload teams, and comfortable leading incident response at senior escalation levels. * Hands-on, adaptable, and comfortable in a fast-paced environment building next-generation infrastructure for ML scale-out. ## Description The Network Engineering Team is responsible for the design, validation, and ongoing operation of all networking services that underpin both the internal management platform and the customer-facing cloud infrastructure - including high-performance Ethernet fabrics, InfiniBand, WAN connectivity, and DC networking. The team acts as a 3rd/4th line escalation point for the support organisation., As a Senior Network Engineer, you will own the design, automation, and in-service operation of our AI-optimised network fabrics - low-latency, high-bandwidth InfiniBand and Ethernet networks supporting large-scale training and inference workloads. You'll take ownership of technical areas end to end, act as a senior escalation point, and help raise the bar on how the team automates, documents, and operates the network., * Design, validate, and operate large-scale InfiniBand/RoCE and Ethernet fabric architectures at rack, row, and DC scale, with tight integration to bare-metal provisioning and cluster management systems. * Apply deep expertise in high-performance Ethernet fabrics (BGP, EVPN, VxLAN, LACP, QoS) and contribute to reference architectures and standards implemented consistently across sites. * Design and engineer perimeter and security infrastructure - firewalls, NAT, VPN, and security policy architecture - across WAN and DC edge environments. * Build and maintain network automation in a GitOps model: Python/Ansible tooling for provisioning, configuration validation, and compliance, with version-controlled configuration and CI/CD-driven change across multi-vendor environments. * Drive operational excellence: resolve complex escalations, lead root-cause analysis for performance and stability issues, and reduce reactive toil through runbooks, automation, and measurable SLOs. * Improve network observability - telemetry, monitoring, and alerting that give clear visibility into fabric health and traffic patterns. * Maintain the accuracy of source-of-truth network inventory and configuration data, with all changes flowing through structured change management. * Collaborate with deployment, DC operations, platform engineering, and vendors on new site delivery, and mentor engineers across the team through reviews and knowledge sharing. ## Related Videos - [Innovating Developer Tools with AI: Insights from GitHub Next](https://www.wearedevelopers.com/videos/1268-innovating-developer-tools-with-ai-insights-from-github-next) - [Dev & Test in the Cloud? Deploy your cloud environments with Ansible & Terraform](https://www.wearedevelopers.com/videos/1607-dev-test-in-the-cloud-deploy-your-cloud-environments-with-ansible-terraform) - [The Gashlycrumb Tinies of AI Networking You Must Know (or Languish!)](https://www.wearedevelopers.com/videos/2067-the-gashlycrumb-tinies-of-ai-networking-you-must-know-or-languish) - [Embracing the Hybrid Cloud: Unlocking Success with Ansible](https://www.wearedevelopers.com/videos/932-embracing-the-hybrid-cloud-unlocking-success-with-ansible) - [Bringing AI Model Testing and Prompt Management to Your Codebase with GitHub Models](https://www.wearedevelopers.com/videos/1536-bringing-ai-model-testing-and-prompt-management-to-your-codebase-with-github-models) - [Running Secure Life Science Research at Scale using Hybrid GPU HPC and Kubernetes 🧬](https://www.wearedevelopers.com/videos/100355-running-secure-life-science-research-at-scale-using-hybrid-gpu-hpc-and-kubernetes) ## Related Articles - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [7 Cloud Computing Trends Coming in 2025 for Developers](https://www.wearedevelopers.com/magazine/412-7-cloud-computing-trends-coming-in-2025-for-developers) - [Everything a Developer Needs to Know About MCP with Neo4j](https://www.wearedevelopers.com/magazine/604-everything-a-developer-needs-to-know-about-mcp-with-neo4j) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Best US AI Conferences for CTOs in 2026: Build vs. Buy, Vendor Evaluation, and Peer Intelligence](https://www.wearedevelopers.com/magazine/736-best-us-ai-conferences-for-ctos-in-2026-build-vs-buy-vendor-evaluation-and-peer-intelligence) - [Graph and AI Trends 2026: Why Is AI Running but Not Yet Delivering?](https://www.wearedevelopers.com/magazine/680-graph-and-ai-trends-2026-why-is-ai-running-but-not-yet-delivering)