> Markdown version of [/jobs/ext/1338648-staff-site-reliability-engineer](https://www.wearedevelopers.com/jobs/ext/1338648-staff-site-reliability-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Staff Site Reliability Engineer - **Company:** DAT, LLC - **Location:** Portland, OR, United States - **Experience:** Expert - **Salary:** $148,000.0 - $193,000.0 - **Contract:** Permanent contract - **Skills:** Java (Programming Language), JavaScript (Programming Language), Amazon Web Services, C++ (Programming Language), Cloud Computing, Github, Python (Programming Language), Octopus Deploy, Reliability Engineering, Site Reliability Engineering Practices, Software Engineering, Datadog, Grafana, Kotlin, Kubernetes, Terraform, Legacy Systems, Golang - **Published:** July 18, 2026 - **Apply:** https://jobs.localjobnetwork.com/apply/add/87757400/1 ## About the Role * Strong leadership and mentoring abilities, especially with SRE or Platform Engineering/Infrastructure teams. * Total of 10+ years industry experience * 3+ years of software engineering experience (JavaScript, Python, Go, Java/Kotlin, C++, etc) * Extensive experience with modern observability tools (Datadog preferred). * Extensive experience with cloud platforms (preferably AWS). * Demonstrated success in leading large technical initiatives, including design, project management and gaining executive buy-in. * Proven experience modernizing legacy code and infrastructure. * Ability to work closely with peer teams, platform/software architects and management to drive key reliability improvements. * Deep understanding of cloud infrastructure, automation, and best practices for reliability. * Experience with our tools (Kubernetes, ArgoCD, Terraform, Github Actions) a plus. ## Description DAT is seeking an experienced Staff Site Reliability Engineer to help grow our SRE practices. In this role, you will be responsible for leading major technical initiatives and mentoring engineers to enhance their skills. You'll work closely with development teams, platform architects and management to achieve critical reliability goals and help scale our platform. What You'll Do * Collaborate with platform architects and management to ensure reliability targets are met. * Advise engineering teams on best practices for measuring reliability and uptime. * Assist and respond to critical engineering incidents * Lead and mentor SRE engineers to improve their engineering skills. * Provide technical guidance and best practices for use of cloud infrastructure and tooling. Be a driver for Infrastructure-as-Code within the platform. * Spearhead major reliability-focused initiatives and projects. * Help optimize our work to be customer-focused. Continually seek feedback from our customers on how we can improve. * Migrate legacy systems to modern, scalable cloud environments. * Help develop and drive a culture of continuous improvement with the Platform Engineering and Software Engineering groups. * Participate in an on-call rotation and occasionally act as Incident Commander. ## Related Videos - [SRE Methods In an Agency Environment](https://www.wearedevelopers.com/videos/348-sre-methods-in-an-agency-environment) - [Go with the Flow: Stop the Leaks Before Your Memory's a Waterfall!](https://www.wearedevelopers.com/videos/100073-go-with-the-flow-stop-the-leaks-before-your-memory-s-a-waterfall) - [Innovating Developer Tools with AI: Insights from GitHub Next](https://www.wearedevelopers.com/videos/1268-innovating-developer-tools-with-ai-insights-from-github-next) - [Kotlin Multiplatform - True power of native code reuse](https://www.wearedevelopers.com/videos/4-kotlin-multiplatform-true-power-of-native-code-reuse) - [Retooling and refactoring - an investment in people.](https://www.wearedevelopers.com/videos/371-retooling-and-refactoring-an-investment-in-people) - [Leading with Reliability: Applying SRE Principles to Build Stronger Engineering Organizations](https://www.wearedevelopers.com/videos/100185-leading-with-reliability-applying-sre-principles-to-build-stronger-engineering-organizations) ## Related Articles - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Find a Developer Job: 12 Best Job Sites For Developers](https://www.wearedevelopers.com/magazine/165-find-a-developer-job-12-best-job-sites-for-developers) - [The Best Software Developer Blogs to Read](https://www.wearedevelopers.com/magazine/156-the-best-software-developer-blogs-to-read)