> Markdown version of [/jobs/ext/2559697-infosec-site-reliability-engineering](https://www.wearedevelopers.com/jobs/ext/2559697-infosec-site-reliability-engineering). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Infosec Site Reliability Engineering - **Company:** Accede Solutions Inc - **Location:** United States (Remote available) - **Experience:** Experienced - **Contract:** Temporary to permanent - **Skills:** Application Programming Interfaces (APIs), Amazon Web Services, Microsoft Azure, Bash Shell, Cyber Security, Decision Support Systems, DevOps, Fault Tolerance, Monitoring of Systems, Python (Programming Language), Reliability Engineering, Site Reliability Engineering Practices, Newrelic, Prometheus, Ruby, Runbook, Scripting, Grafana, Multi-Cloud, Reliability of Systems, Pagerduty, Servicenow - **Published:** August 6, 2026 - **Apply:** https://www.dice.com/job-detail/de9ed0b6-cb2a-49e6-b19d-5a89ca0d8f9e ## About the Role * 4-5 years of experience with SRE, DevOps, or Infrastructure engineering * Skills with multi-cloud environments (AWS, Azure) as well as on-prem integrations * Strong experience with monitoring and observability platforms (NewRelic, Prometheus, Grafana, etc.) * Experience with incident management and on-call systems (e.g. PagerDuty) * Familiarity with ITSM platforms, specifically ServiceNow and integration via API * Solid understanding of system reliability, performance, and scalability * Automation scripting * Runbook generation Desired Skills/Knowledge: * Experience working within or alongside InfoSec teams * Proficiency with scripting languages such as Python, Bash, Ruby etc. * Knowledge of security or security-adjacent tools, controls, and compliance frameworks * Experience implementing SLO's, SLAs, and error budgeting * Exposure to chaos engineering practices (specifically, SteadyBit as a tool) * Strong communication and collaboration skills, specifically documenting * Systems thinking and problem-solving mindset * A strong focus on automation and continuous improvement methodologies * Data-driven decision making * Ability to operate in ambiguous, greenfield environments * Understanding of Infrastructure as Code * Passion for the work and responsibility to the consumer * Understanding of public cloud platforms and services * Positive attitude with a strong desire to continuously learn and adapt This is a foundational role that will shape jhow reliability and operational excellence are embedded within our security organization. You will have the opportunity to build from the ground up, influence strategy, and create lasting impact across the enterprise. ## Description The InfoSec SRE is a pivotal role focused on engineering and advancing the maturity of the organization''s site reliability framework development, adoption, and integration. This position establishes the foundational framework for Site Reliability Engineering (SRE) design fabrics within InfoSec. Additionally, the role supports management of runbook repositories, observability, and related technologies, acting as a trusted advisor to peers and stakeholders across the organization., * Design and implement foundational SRE practices (SLIs/SLOs, error budgets, incident management) * Partner with InfoSec and engineering teams to define reliability standards and operating models * Establish and drive adoption of SRE principles across teams and provide guidance to onboarding organizations to the framework. * Design and implement incident response processes and escalation models * Integrate and optimize alerting and on-call workflows using PagerDuty * Develop and maintain operational runbooks and playbooks * Lead or support incident reviews and postmortems with a focus on continuous improvement * Integrate SRE workflows with enterprise platforms such as ServiceNow * Automate operational tasks, incident workflows, and reporting * Improve system resilience through automation and self-healing mechanisms * Design and execute chaos engineering experiments to validate system resilience * Identify failure modes and proactively address system weaknesses * Collaborate with engineering teams to improve fault tolerance and recovery strategies * Drive adoption of reliability best practices across the organization * Provide guidance and mentorship on SRE principles ## Related Videos - [SRE Methods In an Agency Environment](https://www.wearedevelopers.com/videos/348-sre-methods-in-an-agency-environment) - [Coffee with Developers: David Heinemeier Hansson](https://www.wearedevelopers.com/videos/875-coffee-with-developers-david-heinemeier-hansson) - [From DevOps to Scaled DevOps: How We’re Rebuilding Continuous Delivery as a Platform](https://www.wearedevelopers.com/videos/100018-from-devops-to-scaled-devops-how-we-re-rebuilding-continuous-delivery-as-a-platform) - [Technical Documentation - How Can I Write Them Better and Why Should I Care?](https://www.wearedevelopers.com/videos/681-technical-documentation-how-can-i-write-them-better-and-why-should-i-care) - [3 Key Steps for Optimizing DevOps Workflows](https://www.wearedevelopers.com/videos/962-3-key-steps-for-optimizing-devops-workflows) - [Leading with Reliability: Applying SRE Principles to Build Stronger Engineering Organizations](https://www.wearedevelopers.com/videos/100185-leading-with-reliability-applying-sre-principles-to-build-stronger-engineering-organizations) ## Related Articles - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Events like RSAC Get You CISOs. Developers Decide What Actually Gets Deployed.](https://www.wearedevelopers.com/magazine/693-events-like-rsac-get-you-cisos-developers-decide-what-actually-gets-deployed) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Why Upskilling And Reskilling is Important For Developers](https://www.wearedevelopers.com/magazine/428-why-upskilling-and-reskilling-is-important-for-developers) - [What Are The Top Skills Required For Azure Developers?](https://www.wearedevelopers.com/magazine/77-what-are-the-top-skills-required-for-azure-developers) - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again)