> Markdown version of [/jobs/ext/1964199-infrastructure-support-engineer-site-reliability-engineer-sre-cloud-operations-engineer](https://www.wearedevelopers.com/jobs/ext/1964199-infrastructure-support-engineer-site-reliability-engineer-sre-cloud-operations-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Infrastructure Support Engineer / Site Reliability Engineer (SRE) / Cloud Operations Engineer - **Company:** Wrymark, Inc. - **Location:** San Jose, CA, United States - **Contract:** Permanent contract - **Skills:** Cloud Computing, Domain Name System (DNS), Monitoring of Systems, Routing, Prometheus, TCP/IP, Virtual Machines, Datadog, Load Balancing, Grafana, Reliability of Systems, Firewalls (Computer Science), Splunk, New Relic (SaaS), Dynatrace - **Published:** August 7, 2026 - **Apply:** https://www.dice.com/job-detail/69504fc4-e8fc-4e01-b953-f15d59ade144 ## About the Role * Overall 8+ years of experience * Core platform Engineering * Incidents, Alerts * SRE Mindset * Customer handling skills * Infra/Cloud basic concepts * Network * Storage, 1. SRE Mindset * Understanding of reliability, availability, and performance. * Focus on automation and reducing manual efforts. * Experience with incident management and problem management. 2. Incident & Alert Management * Handling Sev1, Sev2, and Sev3 incidents. * Experience with monitoring tools such as: * Datadog * New Relic * Dynatrace * Splunk * Grafana * Prometheus * Performing RCA and post-incident reviews. 3. Infrastructure Fundamentals Strong understanding of: * Compute: Virtual Machines, CPU, Memory utilization * Storage: SAN, NAS, Disk management, Storage troubleshooting * Network: TCP/IP, DNS, Load Balancers, Firewalls, Routing basics 4. Cloud Basics ## Description * Monitor production environments. * Handle incidents, alerts, and outages. * Perform root cause analysis (RCA). * Troubleshoot infrastructure issues. * Coordinate with customers and internal teams during critical incidents. * Ensure system reliability and availability. ## Related Videos - [An Applied Introduction to eBPF with Go](https://www.wearedevelopers.com/videos/1075-an-applied-introduction-to-ebpf-with-go) - [Our journey with Spring Boot in a microservice architecture](https://www.wearedevelopers.com/videos/511-our-journey-with-spring-boot-in-a-microservice-architecture) - [Creating a routing app with Google Maps API from scratch](https://www.wearedevelopers.com/videos/831-creating-a-routing-app-with-google-maps-api-from-scratch) - [SRE Methods In an Agency Environment](https://www.wearedevelopers.com/videos/348-sre-methods-in-an-agency-environment) - [Turning Container security up to 11 with Capabilities](https://www.wearedevelopers.com/videos/718-turning-container-security-up-to-11-with-capabilities) - [Leading with Reliability: Applying SRE Principles to Build Stronger Engineering Organizations](https://www.wearedevelopers.com/videos/100185-leading-with-reliability-applying-sre-principles-to-build-stronger-engineering-organizations) ## Related Articles - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [What Are The Top Skills Required For Azure Developers?](https://www.wearedevelopers.com/magazine/77-what-are-the-top-skills-required-for-azure-developers) - [Where To Find Software Engineering Jobs](https://www.wearedevelopers.com/magazine/396-where-to-find-software-engineering-jobs) - [What is Software Engineering?](https://www.wearedevelopers.com/magazine/289-what-is-software-engineering)