> Markdown version of [/jobs/ext/3012184-site-reliability-engineer](https://www.wearedevelopers.com/jobs/ext/3012184-site-reliability-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Site Reliability Engineer - **Company:** AppFolio, Inc. - **Location:** San Diego, CA, United States - **Experience:** Expert - **Salary:** $138,400.0 - $173,000.0 - **Contract:** Permanent contract - **Skills:** Amazon Web Services, Amazon S3, Backup Devices, Configuration Management, Relational Databases, DDoS Mitigation, Linux, Disaster Recovery, Amazon DynamoDB, Web Servers, Python (Programming Language), MySQL, Ruby on Rails, Reliability Engineering, Ruby, Runbook, Pulumi, Load Balancing, Cloudformation, Kubernetes, Infrastructure Automation Frameworks, Performance Monitor, Route53, Terraform - **Published:** September 20, 2026 - **Apply:** https://www.themuse.com/jobs/appfolio/sr-site-reliability-engineer-observability-683910 ## About the Role Proven ability to diagnose and monitor performance and reliability issues across the stack: relational databases, web servers, networking, OS, containers, load balancers, etc. You'll chase down performance problems and uncover root causes of system failures. * Strong coding background: you've written code to perform critical tasks or in production. The exact language doesn't matter, though we give bonus points for Go, Ruby, or Python. * Expertise with Infrastructure as Code (Terraform, CloudFormation, Pulumi, etc.) * Expertise with Kubernetes or other container orchestration tooling and technology patterns * Experience with Amazon Web Services (commonly EKS, RDS Aurora, Lambda, S3, EBS, Route53, DynamoDB, and VPCs) * Bachelor's degree and at least 5 years industry experience, or equivalent work experience Ways To Stand Out From The Crowd * Mastery experience with some areas of our tech like Ruby on Rails, Kubernetes, MySQL, Linux, container orchestration, Networking, etc. ## Description AppFolio is more than a company. We're a community of dreamers, big thinkers, problem solvers, active listeners, and multipliers. At every opportunity, we set the pace while delivering innovation built to carry real estate into the future. One in which every experience feels effortless, yet meaningful. Where customers are empowered to take on any opportunity. We show up as one team, connected by our values to be a force for good. Because together, we have the power to create extraordinary outcomes for our customers, our communities, and ourselves. What You'll Do You will be helping them to build common infrastructure as well as help improve the reliability, quality of services and overall observability patterns. Along with your team, you'll ensure all aspects of our shared product spaces have a plan to address any opportunities for exception reporting, capacity planning, monitoring and alerting, backups, runbooks, configuration management, DDoS protection, infrastructure as code, and disaster recovery. You'll collaborate or embed with engineering teams, helping them to improve the reliability and quality of their services and infrastructure as we look to support the business success of AppFolio. You'll be a member of the team that provides reliable, scalable observability and infrastructure for key components of the AppFolio Real Estate Platform. You'll help build the future of reliable, critical services, and support rapid and sustainable growth of new features. What We're Looking For This is a great opportunity for someone who enjoys supporting teams, helping them become more self-sufficient, and building reliable, simple systems. This position will require on-call responsibilities. Lastly, you have strong communication skills and enjoy working on a team that values openness, integrity, ownership, and attention to detail. Must-Haves * Experience developing Service Level Indicators and Service Level Objectives for the above systems. * ## Related Videos - [Reliable scalability: How Amazon.com scales on AWS](https://www.wearedevelopers.com/videos/983-reliable-scalability-how-amazon-com-scales-on-aws) - [MySQL Protocol Features You Should Be Aware Of](https://www.wearedevelopers.com/videos/100267-mysql-protocol-features-you-should-be-aware-of) - [Docker network without Docker](https://www.wearedevelopers.com/videos/1418-docker-network-without-docker) - [Coffee with Developers: David Heinemeier Hansson](https://www.wearedevelopers.com/videos/875-coffee-with-developers-david-heinemeier-hansson) - [Shipping Faster with Less: Render on Cloud Hosting, AI Workloads, and the Future of DevOps](https://www.wearedevelopers.com/videos/1894-shipping-faster-with-less-render-on-cloud-hosting-ai-workloads-and-the-future-of-devops) - [Coding for Good: Achieving social change with an app](https://www.wearedevelopers.com/videos/1645-coding-for-good-achieving-social-change-with-an-app) ## Related Articles - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [How to Answer the Interview Question: “Why Do You Want to Be a Software Engineer?”](https://www.wearedevelopers.com/magazine/392-how-to-answer-the-interview-question-why-do-you-want-to-be-a-software-engineer) - [Why Upskilling And Reskilling is Important For Developers](https://www.wearedevelopers.com/magazine/428-why-upskilling-and-reskilling-is-important-for-developers) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs)