> Markdown version of [/jobs/ext/2793634-site-reliability-engineer](https://www.wearedevelopers.com/jobs/ext/2793634-site-reliability-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Site Reliability Engineer - **Company:** Filevine, Inc. - **Location:** San Francisco, CA, United States (Remote available) - **Experience:** Experienced - **Salary:** $160,000.0 - $190,000.0 - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Amazon Web Services, Build Automation, Bash Shell, Cloud Computing, Cloud Engineering, Information Systems, Continuous Integration, DevOps, Distributed Systems, Identity and Access Management, Python (Programming Language), Reliability Engineering, Software Tools, Software Engineering, Data Logging, Pulumi, Scripting, Cloudformation, Kubernetes, Infrastructure Automation Frameworks, Information Technology, Deployment Automation, Terraform, Dynatrace - **Published:** September 8, 2026 - **Apply:** https://startup.jobs/site-reliability-engineer-filevine-9943334 ## About the Role * 4+ years of hands-on experience in software engineering, cloud infrastructure, platform engineering, DevOps, or related technical roles, including at least 2 years in a Site Reliability Engineering or reliability-focused role. * Working knowledge of distributed systems and how applications, infrastructure, and cloud services interact in production; demonstrated ability to troubleshoot production issues, perform root cause analysis, and drive long-term reliability improvements. * Proficiency with Python, Bash, or similar scripting languages; experience building production tooling, automation, or CI/CD pipelines and deployment automation. * Hands-on experience operating Kubernetes-based workloads and cloud infrastructure in AWS or a comparable platform, including compute, container orchestration, networking, IAM, object storage, and cloud-native monitoring. * Experience with Infrastructure as Code tools such as Terraform, Pulumi, or AWS CloudFormation, and familiarity with modern observability practices including monitoring, logging, alerting, distributed tracing, and incident response. * Experience using AI-assisted engineering tools to improve productivity, accelerate troubleshooting, or automate operational tasks; curiosity, ownership, and a passion for building reliable systems through continuous improvement. * Strong written and verbal communication skills, Bachelor's degree in Computer Science, Information Systems, or a related field, equivalent industry certifications, or comparable ## Description * Design, build, and maintain the monitoring, logging, distributed tracing, dashboards, and alerting that give teams meaningful visibility into production health. * Build automation, tooling, and CI/CD improvements that increase engineering efficiency, reduce toil, and support reliable deployments at scale. * Design, implement, and maintain reliable systems for building, deploying, testing, and operating Filevine products - proactively identifying and resolving reliability, performance, scalability, and security risks before they impact customers. * Participate in a shared 24/7 on-call rotation, using operational insights to drive automation and long-term reliability improvements; continuously improve runbooks, documentation, and engineering standards. * Take ownership of technical initiatives from design through implementation, develop deep expertise in critical areas of the Filevine platform, and communicate clearly with technical and business stakeholders. ## Related Videos - [Shipping Faster with Less: Render on Cloud Hosting, AI Workloads, and the Future of DevOps](https://www.wearedevelopers.com/videos/1894-shipping-faster-with-less-render-on-cloud-hosting-ai-workloads-and-the-future-of-devops) - [JavaScript? No. Java Scripts! - Scripting with Java](https://www.wearedevelopers.com/videos/2094-javascript-no-java-scripts-scripting-with-java) - [Why segmenting your infrastructure into tiers makes your infrastructure design better](https://www.wearedevelopers.com/videos/1960-why-segmenting-your-infrastructure-into-tiers-makes-your-infrastructure-design-better) - [From DevOps to Scaled DevOps: How We’re Rebuilding Continuous Delivery as a Platform](https://www.wearedevelopers.com/videos/100018-from-devops-to-scaled-devops-how-we-re-rebuilding-continuous-delivery-as-a-platform) - [Unleashing Potential Across Teams: The Power of Infrastructure as Code](https://www.wearedevelopers.com/videos/930-unleashing-potential-across-teams-the-power-of-infrastructure-as-code) - [My journey into DevOps world - How it all started!](https://www.wearedevelopers.com/videos/545-my-journey-into-devops-world-how-it-all-started) ## Related Articles - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Where To Find Software Engineering Jobs](https://www.wearedevelopers.com/magazine/396-where-to-find-software-engineering-jobs) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers)