> Markdown version of [/jobs/ext/3619061-site-reliability-engineer](https://www.wearedevelopers.com/jobs/ext/3619061-site-reliability-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Site Reliability Engineer - **Company:** F5 Networks, Inc. - **Location:** Seattle, WA, United States - **Experience:** Starter - **Salary:** $137,300.0 - $205,900.0 - **Contract:** Permanent contract - **Skills:** Amazon Web Services, Software as a Service, Cloud Computing, Databases, Linux, Monitoring of Systems, Hypertext Transfer Protocols (HTTP), Issue Tracking Systems, JSON, Python (Programming Language), PostgreSQL, Uptime, Networking Basics, Reliability Engineering, Prometheus, Scripting, Cloud Platform System, System Availability, Grafana, Reliability of Systems, Kubernetes, Infrastructure Automation Frameworks, Information Technology, Performance Monitor, Web Technologies, Restful APIs, Model Inference, Terraform - **Published:** October 8, 2026 - **Apply:** https://ffive.wd5.myworkdayjobs.com/f5jobs/job/Seattle/Site-Reliability-Engineer_RP1038966 ## About the Role The ideal candidate has a strong technical foundation, thrives in a fast-paced troubleshooting environment, and is passionate about automation and delivering an exceptional customer experience. This role provides the opportunity to run, support, and scale an AI Security Public SaaS platform, operating AI inference workloads at scale., * Bachelor's degree in Computer Science, Information Technology, or a related field (or equivalent practical experience) * 1-3+ years of experience in technical support, systems administration, or a similar role * Strong understanding of SaaS environments and cloud-based architectures (preferably AWS) * Proficiency in at least one scripting language (e.g., Python) * Solid understanding of web technologies (HTTP, REST APIs, JSON, etc.) * Experience working with ticketing systems * Strong problem-solving and analytical skills * Excellent written and verbal communication skills * Ability to work independently and collaboratively in a team environment * Willingness to learn new technologies and adapt to evolving requirements, * Experience with monitoring and observability tools (e.g., Prometheus, Grafana) * Familiarity with configuration management tools (e.g., Terraform) * Experience working with cloud infrastructure technologies * Exposure to SRE principles and reliability engineering practices * Strong understanding of networking fundamentals * Experience with databases (PostgreSQL), operating systems (Linux), and Kubernetes #LI-ZB1 ## Description This hybrid role combines the hands-on responsibilities of a Technical Support Engineer within a SaaS (Software as a Service) environment with a growing focus on Site Reliability Engineering (SRE)., Proactive Monitoring & Uptime Assurance * Monitor key SaaS application metrics, logs, and alerts to proactively identify and prevent service disruptions * Support a 24/7 operational model to ensure high availability and performance of the SaaS environment Customer-Centric Incident Response * Serve as a primary point of contact for technical customer inquiries and issues * Partner with customers to understand requirements, troubleshoot effectively, and provide clear, concise communication * Triage, escalate, and own technical incidents through resolution Development Team Collaboration * Analyze metrics, logs, and incident reports to provide actionable insights to engineering teams * Identify recurring patterns and opportunities for automation and continuous improvement * Partner cross-functionally to improve platform reliability and performance, * Design and implement automation to streamline incident response and improve system reliability * Introduce and mature SRE best practices, including enhanced monitoring, alerting, and self-healing capabilities * Contribute to building scalable, resilient infrastructure to support AI inference workloads