> Markdown version of [/jobs/ext/1467900-network-engineer](https://www.wearedevelopers.com/jobs/ext/1467900-network-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Network Engineer - **Company:** Together Ai - **Location:** San Francisco, CA, United States - **Experience:** Expert - **Salary:** $190,000.0 - $280,000.0 - **Contract:** Temporary to permanent - **Skills:** Artificial Intelligence, Amazon Web Services, Application Layers, Microsoft Azure, Border Gateway Protocol, Cloud Computing, Computer Clusters, Code Review, Complex Networks, Computer Networks, Continuous Integration, Data Centers, Linux, Distributed Data Store, InfiniBand, Networking Hardware, Multi-protocol Systems, Python (Programming Language), Network Troubleshooting, CURL, Linux System Administration, Network Architecture, Network Planning and Design, Routing, Nmap, Open Shortest Path First (OSPF), Operational Databases, Overlay Transport Virtualization, Remote Direct Memory Access, Software Tools, Ansible, Software Engineering, TCP/IP, Tcpdump, Wireshark, Virtual Local Area Networks, Multitopology Routing Configuration, Computer Networking Systems, Google Cloud, High Performance Computing, Reliability of Systems, Juniper, Git, Kubernetes, Low Latency, Code Inspection, Open Network Automation Platform, Cisco - **Published:** July 28, 2026 - **Apply:** https://www.dice.com/job-detail/02de89e7-138c-4f33-8f45-5fb7e29582c1 ## About the Role * 8+ years of professional experience designing, building, and supporting large-scale production data center, cloud, service-provider, or high-performance computing networks (excluding enterprise networks). * Deep understanding of TCP/IP and strong experience with technologies such as BGP, OSPF, VXLAN, EVPN, ECMP, and QoS. * Experience designing and supporting multi-tenant network environments using technologies such as VRFs, VLANs, overlays, and policy-based segmentation. * Hands-on experience deploying and troubleshooting network platforms from vendors such as Arista, Cisco, Juniper, and NVIDIA. * Strong troubleshooting skills using tools such as Wireshark, tcpdump, MTR, curl, nmap, and standard Linux networking utilities. * Ability to diagnose connectivity, latency, packet-loss, routing, and performance issues across the network, host, and application layers. * Experience developing or maintaining network automation using Python, Ansible, or similar tools. * Experience working through a Git-based software development lifecycle, including branching, code review, validation, linting, testing, CI/CD, deployment, and rollback. * Working knowledge of Kubernetes networking, including pods, services, CNIs, and basic connectivity troubleshooting. * Foundational knowledge of RDMA networking and technologies such as RoCE or InfiniBand. * Experience with cloud networking in AWS, Google Cloud Platform, or Azure. * Strong Linux administration and troubleshooting skills., * Hands-on experience deploying or operating RoCE and/or InfiniBand fabrics. * Experience supporting GPU clusters, HPC environments, distributed storage, or other high-bandwidth and latency-sensitive workloads. * Understanding of AI training and inference traffic patterns and the demands they place on network infrastructure. * Experience operating networks spanning thousands of devices, multiple data centers, and multiple geographic regions. * Familiarity with AI-assisted engineering tools and the ability to validate, test, and safely deploy AI-generated automation or code. ## Description Together AI is looking for a Senior Network Engineer to design, deploy, and operate the global network infrastructure supporting our production services and high-performance AI compute environments. This is a hands-on engineering role for someone with deep networking expertise who can also troubleshoot across Linux, Kubernetes, automation, and application boundaries. You will work on large-scale, multi-vendor data center networks and help ensure they remain highly available, reliable, scalable, and performant. The ideal candidate has strong networking fundamentals, experience operating complex networks at scale, and a structured, evidence-based approach to troubleshooting. You should be comfortable owning problems from initial investigation through root cause and resolution, including situations where the issue may extend beyond the network itself., * Design, deploy, operate, and maintain global, multi-vendor, multi-protocol networks supporting high-performance AI compute infrastructure. * Troubleshoot complex network and application-connectivity issues, identify root causes, and drive problems through resolution. * Analyze telemetry, packet captures, logs, and performance data to identify network degradation, congestion, packet loss, and capacity constraints. * Participate in architecture and design reviews to ensure solutions meet requirements for performance, availability, scalability, security, and operational supportability. * Develop and maintain automation, validation, and operational tooling that improves network reliability and reduces manual effort. * Evaluate network hardware, software, optics, and emerging technologies for use in production environments. * Establish standards and operational best practices for network design, deployment, monitoring, change management, and incident response. * Lead projects addressing complex technical challenges and contribute directly to the network engineering roadmap. * Partner with infrastructure, systems, security, and application teams to troubleshoot issues that cross traditional ownership boundaries. ## Related Videos - [The Gashlycrumb Tinies of AI Networking You Must Know (or Languish!)](https://www.wearedevelopers.com/videos/2067-the-gashlycrumb-tinies-of-ai-networking-you-must-know-or-languish) - [Don’t Insert Crazy! On cURL and AI Slop - Daniel Stenberg](https://www.wearedevelopers.com/videos/1796-don-t-insert-crazy-on-curl-and-ai-slop-daniel-stenberg) - [How a Small Team Shrank a Microsoft Monorepo by 94%](https://www.wearedevelopers.com/videos/1236-how-a-small-team-shrank-a-microsoft-monorepo-by-94) - [Building the Nervous System of AI - Michael Kagan (NVIDIA)](https://www.wearedevelopers.com/videos/2133-building-the-nervous-system-of-ai-michael-kagan-nvidia) - [Capture the Flag 101](https://www.wearedevelopers.com/videos/416-capture-the-flag-101) - [Git for Code Reviews](https://www.wearedevelopers.com/videos/429-git-for-code-reviews) ## Related Articles - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Best US AI Conferences for CTOs in 2026: Build vs. Buy, Vendor Evaluation, and Peer Intelligence](https://www.wearedevelopers.com/magazine/736-best-us-ai-conferences-for-ctos-in-2026-build-vs-buy-vendor-evaluation-and-peer-intelligence) - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [Graph and AI Trends 2026: Why Is AI Running but Not Yet Delivering?](https://www.wearedevelopers.com/magazine/680-graph-and-ai-trends-2026-why-is-ai-running-but-not-yet-delivering) - [7 Cloud Computing Trends Coming in 2025 for Developers](https://www.wearedevelopers.com/magazine/412-7-cloud-computing-trends-coming-in-2025-for-developers) - [Everything a Developer Needs to Know About MCP with Neo4j](https://www.wearedevelopers.com/magazine/604-everything-a-developer-needs-to-know-about-mcp-with-neo4j)