Principal Network Reliability Engineer

VeeRteq Solutions Inc
Plano, United States
8 days ago
Apply on www.dice.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
10 years minimum
Working hours
Regular working hours
Job source

Tech stack

Application Programming Interfaces (APIs) Amazon Web Services Microsoft Azure Software as a Service Cloud Computing Cloud Engineering Computer Networks Data Centers DevOps Distributed Systems Domain Name System (DNS) Python (Programming Language)
+26 more
Network Architecture Networking Basics Network Service Reliability Engineering Ansible Prometheus Zero Trust Network Access Wide Area Networks Cloud-native Network Functions (CNF) Computer Networking Systems Google Cloud Load Balancing Grafana Multi-Cloud Reliability of Systems Data Center Networking Kubernetes Infrastructure Automation Frameworks SDN Network Hashicorp Api Gateway Terraform Open Network Automation Platform Dynatrace Cisco Microservices

Job description

Principal Network Reliability Engineer responsible for architecting, securing, automating, and optimizing highly available network platforms across hybrid and multi-cloud environments, leveraging SRE, Platform Engineering, and DevOps principles to drive operational excellence and network resiliency., * Design and evolve highly available, resilient network platforms supporting enterprise and cloud-native healthcare applications.

  • Architect secure, scalable networking solutions across hybrid data centers and public cloud environments while remaining cloud-provider agnostic.
  • Lead modernization initiatives supporting the evolution from traditional enterprise networking to software-defined, cloud-native networking.
  • Develop Infrastructure as Code, network automation, and policy-as-code solutions that eliminate manual configuration and improve deployment consistency.
  • Design and maintain resilient connectivity supporting Kubernetes platforms, container networking, service meshes, APIs, and microservices-based applications.
  • Establish enterprise network observability through telemetry, flow analytics, synthetic monitoring, distributed tracing, and proactive ing.
  • Lead root cause analysis and drive long-term corrective actions following critical network and cloud connectivity incidents.
  • Partner with Platform Engineering, Security, and Application Engineering teams to improve production readiness, Zero Trust networking, and cloud connectivity.
  • Develop engineering standards, reusable automation, and technical documentation that improve operational consistency and reduce complexity.
  • Evaluate emerging networking technologies and recommend architectural improvements that increase resiliency, scalability, security, and operational efficiency.
  • Mentor engineers while serving as the highest-level technical authority for networking strategy, automation, and cloud networking architecture.
  • Support strategic cloud initiatives across AWS and Google Cloud while designing solutions that remain portable across cloud providers where practical.

Requirements

  • Engineering Degree BE/ME/BTech/MTech/BSc/MSc.
  • Technical certification in multiple technologies is desirable.

Skills: -

Mandatory skills

  • Min10+ years of experience designing and supporting enterprise network platforms in production environments.
  • Expert knowledge of enterprise routing, switching, SD-WAN, data center networking, and cloud networking technologies.
  • Strong experience designing hybrid and public cloud network architectures across AWS, Google Cloud, Azure, or equivalent cloud platforms.
  • Hands-on expertise with Kubernetes networking, container networking, service mesh technologies, and modern application connectivity patterns.
  • Experience implementing Infrastructure as Code (IaC), network automation, and DevOps practices using Terraform, Ansible, Python, or similar technologies.
  • Strong background supporting enterprise SaaS platforms, distributed systems, and high-availability network architectures.
  • Experience participating in incident response, root cause analysis, reliability engineering initiatives, and operational excellence programs.
  • Proven ability to mentor engineers and influence enterprise architecture, engineering standards, and networking strategies.
  • Strong understanding of Site Reliability Engineering (SRE), Platform Engineering, Zero Trust networking, and cloud-native networking principles.

Good to Have Skills

  • Healthcare, SaaS, or regulated industry experience.
  • Expertise designing secure, highly available, cloud-native network architectures.
  • Experience with Software-Defined Networking (SDN), SD-WAN, Zero Trust, and cloud-native networking services.
  • Familiarity with observability platforms such as Dynatrace, Prometheus, Grafana, OpenTelemetry, ThousandEyes, or similar network telemetry solutions.
  • Experience implementing GitOps, CI/CD pipelines, and infrastructure automation for network platforms.
  • Strong understanding of DNS, load balancing, API gateways, ingress controllers, and traffic management technologies.
  • Cisco CCIE (Enterprise Infrastructure, Security, or Data Center) or equivalent expert-level networking certification.
  • Google Cloud Professional Cloud Network Engineer or AWS Advanced Networking Specialty certification.
  • Certified Kubernetes Administrator (CKA).
  • HashiCorp Terraform Associate or equivalent Infrastructure as Code certification.
  • Relevant SRE, Zero Trust, or cloud networking certifications

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.dice.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

1:34 min

Pivoting careers into specialized platform engineering roles

Xavier Portilla Edo · LIVE

1:29 min

Expanding practical knowledge with community sandboxes and resources

Stuart Clark · LIVE

2:17 min

Mapping the maturity roadmap for scaled devops adoption

Dominik Krichbaum Dominik Krichbaum · World Congress 2026 Europe

10:40 min

Visualizing Prometheus open metrics using custom Grafana dashboards

Stijn Polfliet · LIVE

3:18 min

Scaling global network engineering through DevOps culture

Stuart Clark · LIVE

1:20 min

Identifying multi-disciplinary talent for developer experience engineering roles

Hazal Mestci +1 · Coffee With Developers

Videos

See all

Related articles

See all