> Markdown version of [/jobs/ext/1796180-senior-site-reliability-engineer](https://www.wearedevelopers.com/jobs/ext/1796180-senior-site-reliability-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Senior Site Reliability Engineer - **Company:** ServiceTitan, Inc. - **Location:** Glendale, CA, United States - **Experience:** Expert - **Salary:** $147,600.0 - $221,400.0 - **Contract:** Permanent contract - **Skills:** Java (Programming Language), .NET Framework, Amazon Web Services, Software Applications, Microsoft Azure, C Sharp (Programming Language), Cloud Computing, Continuous Integration, Software Debugging, Distributed Revision Control, Distributed Systems, Elasticsearch, Machine Learning, Enterprise Messaging Systems, Windows PowerShell, Systems Development Life Cycle, Reliability Engineering, Logstash, Responsive Web Design, Software Engineering, Web Applications, Azure Service Bus, Software Organization, Datadog, Cloud Platform System, System Availability, Snowflake, Grafana, Reliability of Systems, Git, Kubernetes, Information Technology, Apache Kafka, Data Lakehouse, Api Gateway, Kibana, Amazon Simple Queue Service (SQS), Network Server, Serverless Computing, Visual Basic Language, Jenkins, Databricks - **Published:** July 17, 2026 - **Apply:** https://servicetitan.wd1.myworkdayjobs.com/ServiceTitan/job/Glendale-CA/Senior-Site-Reliability-Engineer_JR115077 ## About the Role responsive web application development, building distributed systems for scale, and a proven ability to deliver technical leadership and strong architectural thought process., * BS or MS in Computer Science (or equivalent diploma and/or certifications) with 8-10 years related experience. * Familiarity with Continuous Integration and Delivery including experience with tools such as Jenkins or Team City. * Strong expertise & experience with majority of Kubernetes, Functions/Serverless computing, Distributed messaging systems (Kafka, Event Hubs, SQS * etc), Data Lakehouse architectures (e.g. Snowflake, Databricks Delta) and API gateways will be crucial to the success of this role. * Administration and building automation for Azure, AWS or other public cloud technology. * Hands-on experience with a Distributed Version Control System such as Git * Advanced knowledge of at least of the following programming languages: C#, Visual Basic, PowerShell, Java. * Experience scripting provisioning of servers, applications, and/or infrastructure in a production environment at scale. * Knowledge of software development best practices, SDLC and experience deploying high availability systems and software. * Experience with troubleshooting distributed web applications in a production environment. * Log / Metric collection and analysis tools (e.g. Elasticsearch-Logstash-Kibana, DataDog, Grafana) * Administration and building automation for Azure, AWS or other public cloud technology About You: You're someone who enjoys being directly accountable for the reliability of a business-critical, large-scale enterprise system. You're comfortable guiding and making decisions with limited information and are capable of operating within the trade-offs present when solving for immediate needs versus solving with bigger scale solutions. You might be considered a subject matter expert in Cloud Infrastructure & Systems Reliability and you feel rewarded by working to develop an operability culture in a quickly growing and changing environment. You're comfortable owning a wide and diverse set of problem areas and are willing to go out of your lane to affect change. You may have developed one or more metrics, log aggregation or performance analysis systems in your ## Description We make a huge impact on thousands of companies in the U.S. and abroad by enabling them to be more efficient and effective at running their business. Many of the features we offer - in particular, Machine Learning and AI-driven scheduling and dispatch automation - are light-years ahead of what currently exists on the market, and we love to hold this position. We are quality minded, use the most modern tools on .NET platform, have an amazing culture, love to solve complex problems, and embrace learning, exploration and growth. If you share the same values, we'd like to talk to you! Our Site Reliability and Infrastructure Engineering team is an investment by Cloud to make "big-hammer", impactful changes to ServiceTitan that help us constantly run better, faster, and cheaper. This team centralizes the concerns of developing and providing measurements and guidance, so every engineer is able to improve availability and efficiency in their area of the ServiceTitan cloud. Infrastructure Engineering improves ServiceTitan's customer experiences by ever-increasing availability and performance; reclaiming time spent by our engineers diagnosing issues or configuring software; reducing the total cost of owning and operating products and services. We have a cultural foundation built on diversity, inclusion and innovation and we want you and your ideas to thrive at ServiceTitan. Come join us. What You'll Do: * Design, develop, test, troubleshoot, debug, optimize, scale, perform the capacity planning, deploy, maintain, and improve software applications, * driving the delivery of high-quality value and features to ServiceTitan's customers. * Coding and Automation of Applications on Cloud Platform * Work collaboratively across the company to design, communicate and further assist with adoption of best practices in architecture and implementation. * Actively participate in research, development, support, management, and other company initiatives designing solutions to optimally address current * and future business requirements and infrastructure plans. * Collaborate with Product Engineering teams to plan and deploy product releases * Work with Engineering leadership to build scalable infrastructure & shared services that meet the requirements and need of the platform and * application teams * Define non-functional requirements as part of the product lifecycle to influence the new designs, standards, and methods for scalable, highly available * distributed systems * Contribute to product development / engineering as needed to ensure Quality of Service of Highly Available services * Resolution of product/service defects or design changes, infrastructure changes, or operational Changes * Establish strong relationships with company's leadership to ensure the use of technologies are well. ## Related Videos - [How a Small Team Shrank a Microsoft Monorepo by 94%](https://www.wearedevelopers.com/videos/1236-how-a-small-team-shrank-a-microsoft-monorepo-by-94) - [Debug a Kubernetes Operator](https://www.wearedevelopers.com/videos/487-debug-a-kubernetes-operator) - [Debugging in the Dark](https://www.wearedevelopers.com/videos/1658-debugging-in-the-dark) - [DevOps at Netflix](https://www.wearedevelopers.com/videos/270-devops-at-netflix) - [Remote Driving on Plant Grounds with State-of-the-Art Cloud Technologies](https://www.wearedevelopers.com/videos/251-remote-driving-on-plant-grounds-with-state-of-the-art-cloud-technologies) - [Git for Code Reviews](https://www.wearedevelopers.com/videos/429-git-for-code-reviews) ## Related Articles - [What Are The Top Skills Required For Azure Developers?](https://www.wearedevelopers.com/magazine/77-what-are-the-top-skills-required-for-azure-developers) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Now is the time for industrialized software development](https://www.wearedevelopers.com/magazine/601-now-is-the-time-for-industrialized-software-development) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs)