> Markdown version of [/jobs/ext/2105212-site-reliability-engineer-ii](https://www.wearedevelopers.com/jobs/ext/2105212-site-reliability-engineer-ii). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Site Reliability Engineer II - **Company:** PROS, Inc. - **Location:** Houston, TX, United States - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Applications Architecture, Automation of Tests, Microsoft Azure, Border Gateway Protocol, Cloud Computing, Dynamic Host Configuration Protocol, Network Address Translation, Domain Name System (DNS), Enhanced Interior Gateway Routing Protocol, IP Addressing, Intrusion Detection Systems, Subnetting, Network Security, Network Layer, Network Monitoring, Routing, Network Segmentation, Packet Analyzer, Network Protocols, Nginx, Open Shortest Path First (OSPF), Peering, Release Management, Reliability Engineering, Software Reliability Testing, Remote Access Technology, Systems Architecture, Load Balancing, Cloud Platform System, Reliability of Systems, Firewalls (Computer Science), Amazon Virtual Private Cloud (VPC), Information Technology, Performance Monitor, Fortinet, Firewall Services Module - **Published:** August 18, 2026 - **Apply:** https://pros.wd5.myworkdayjobs.com/PROS_Careers/job/USA-TX-Houston-Office/Site-Reliability-Engineer-II_R3567 ## About the Role We are looking for candidates who possess the combination of the following achievements, skills, and behaviors: * 5+ years of experience in enterprise networking, including hands-on work with routing, switching, firewalls, load balancers, and VPN technologies. * Strong understanding of cloud networking architectures across including VPC/VNet design, peering, private link, and hybrid connectivity models. * Experience with network security technologies, such as security groups, NACLs, firewall policies, WAF, IDS/IPS, and micro-segmentation. * Proficiency in Layer 2 and Layer 3 network protocols, including BGP, OSPF, EIGRP, DNS, DHCP, NAT, and IP addressing/subnetting. * Hands-on experience with load balancers and ingress technologies, including F5, NGINX, Azure Application Gateway, ALB/NLB, or equivalent. * Strong troubleshooting skills using packet analyzers tools, flow logs, and network monitoring platforms. * Skilled in analyzing performance trends and identifies optimization opportunities. * Collaborates with teams to improve monitoring coverage. * Ability to participate in structured reliability testing and analysis. * Able to evaluate system components for resilience. * Contributes to reliability-focused design discussions. * Skilled in analyzing trends to inform service improvements. * Collaborates with teams to align SLOs with user expectations. * Develops moderately complex automation tools. * Skill in building internal self-service capabilities. * Evaluates automation opportunities for operational efficiency. * Skilled in analyzing capacity data to inform scaling decisions. * Able to recommend improvements for resource utilization. * Ensures scalability is considered in feature development. * Follow predefined procedures to deploy PROS products and third-party applications to the Cloud environments. * Contribute to the release management documentation. * Gain understanding of application architecture and interaction between system components. Highly Preferred: * Bachelor's Degree in Computer Science, Information Technology, or a related field * Practical experience with Fortigate firewalls and F5 appliances is highly desirable ## Description The Site Reliability Engineer II optimizes service performance, actively participates in reliability improvements, and conducts in-depth SLO and capacity analysis. This position exists to enhance system reliability and scalability while contributing to automation and self-service tool development. . * Performance Monitoring: Monitor service performance, assist in troubleshooting production issues, and learn system architecture. * Reliability Participation: Monitor service reliability, participate in resolving basic issues, and learn disaster recovery testing procedures. * SLO Implementation: Understand SLO concepts, monitor and analyze SLO patterns, and assist in implementing SLO visualization and alerting. * Capacity Analysis: Perform basic capacity analysis, identify trends in system capacity, and participate in capacity planning. * Automation Deployment: Deploy and maintain existing automation tools, create simple scripts, and troubleshoot automation scripts., * Understand core AI concepts and apply them ethically to enhance productivity, insights, and decision-making. * Craft effective prompts to optimize the quality and relevance of AI-generated outputs. * Explore and apply agentic AI systems, using or managing autonomous agents to streamline workflows and automate tasks. * Leverage AI tools to boost efficiency, creativity, and innovation in their daily work. * Stay curious and adaptable, continuously experimenting with AI-driven solutions to elevate team performance and customer impact. ## Related Videos - [SRE Methods In an Agency Environment](https://www.wearedevelopers.com/videos/348-sre-methods-in-an-agency-environment) - [Transforming Education: A Journey from interactive Markdown to Remote-Labs](https://www.wearedevelopers.com/videos/941-transforming-education-a-journey-from-interactive-markdown-to-remote-labs) - [Post-Quantum Cryptography: Preparing for Q-Day](https://www.wearedevelopers.com/videos/100179-post-quantum-cryptography-preparing-for-q-day) - [Creating a routing app with Google Maps API from scratch](https://www.wearedevelopers.com/videos/831-creating-a-routing-app-with-google-maps-api-from-scratch) - [Azure-Well Architected Framework - designing mission critical workloads in practice](https://www.wearedevelopers.com/videos/1529-azure-well-architected-framework-designing-mission-critical-workloads-in-practice) - [How Cisco embraced a DevOps culture within its network engineering team](https://www.wearedevelopers.com/videos/99-how-cisco-embraced-a-devops-culture-within-its-network-engineering-team) ## Related Articles - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [What Are The Top Skills Required For Azure Developers?](https://www.wearedevelopers.com/magazine/77-what-are-the-top-skills-required-for-azure-developers) - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Why Upskilling And Reskilling is Important For Developers](https://www.wearedevelopers.com/magazine/428-why-upskilling-and-reskilling-is-important-for-developers) - [The 12 Best Jobs for Software Engineers](https://www.wearedevelopers.com/magazine/401-the-12-best-jobs-for-software-engineers) - [Where To Find Software Engineering Jobs](https://www.wearedevelopers.com/magazine/396-where-to-find-software-engineering-jobs)