> Markdown version of [/jobs/ext/550083-site-reliability-engineer](https://www.wearedevelopers.com/jobs/ext/550083-site-reliability-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Site Reliability Engineer - **Company:** TD Ameritrade - **Location:** Austin, TX, United States - **Experience:** Expert - **Salary:** $126,000.0 - $140,000.0 - **Contract:** Permanent contract - **Skills:** Cloud Computing, Cloud Foundry, Computer Programming, Computer Engineering, DevOps, Fault Tolerance, Python (Programming Language), PostgreSQL, RabbitMQ, Reliability Engineering, Software Engineering, Data Streaming, Aerospike, Google Cloud, Cloud Platform System, Mttr, Cloudformation, Build Management, Information Technology, Apache Kafka, Terraform - **Published:** June 10, 2026 - **Apply:** https://www.schwabjobs.com/job/austin/site-reliability-engineer/33727/96223583952 ## About the Role * Bachelor's degree in Computer Engineering, Computer Science, or related field * 6+ years of software development and site reliability engineering experience supporting production applications in cloud environments such as Pivotal Cloud Foundry (PCF) or Google Cloud Platform (GCP) * 4+ years of DevOps engineering leadership experience focused on automation, tooling, and improving production operations * 2+ years of technical leadership experience guiding engineering teams and driving operational efficiencies * 2+ years of experience implementing and maturing operational best practices, including SLOs, SLIs, error budgets, monitoring, capacity planning, and incident management processes * Proficiency in programming and automation using tools such as Python, CloudFormation, or Terraform to build infrastructure-as-code solutions * Strong knowledge of database technologies (SQL, Aerospike, Postgres) * Experience working with messaging and streaming platforms such as RabbitMQ and Kafka Preferred Qualifications: * 4+ years of advanced technical leadership experience supporting highly skilled engineering teams * Demonstrated ability to influence development teams to design and build cloud-native systems that are scalable, maintainable, and resilient from initial deployment onward ## Description As a Sr Specialist - Site Reliability Engineer (SRE) within Client Data Technology, you will play a critical role in ensuring the availability, performance, and resiliency of highly visible cloud-based platforms and applications. In this role, you will influence how systems are designed, built, and operated, driving measurable improvements in reliability and scalability while advancing modern SRE practices across the organization. You will partner closely with engineering and platform teams to define and implement sustainable operating models, enabling consistent, repeatable, and high-performing systems at scale. Your impact will include identifying and executing opportunities to enhance service health and telemetry, shaping and delivering forward-looking resiliency and availability roadmaps, and leading the adoption of cloud-native technologies aligned with established SRE standards. Through strong collaboration and technical leadership, you will promote a proactive, "shift-left" approach that embeds reliability, fault tolerance, and performance into the development lifecycle from the start. This role requires a balance of strategic thinking and hands-on problem-solving to optimize systems, reduce operational toil, and improve key metrics such as MTTD and MTTR, ultimately ensuring a seamless and reliable experience for clients. ## Related Videos - [Inside Bitpanda's Tech Stack: Scaling a European Fintech Leader - Markus Dorner](https://www.wearedevelopers.com/videos/1979-inside-bitpanda-s-tech-stack-scaling-a-european-fintech-leader-markus-dorner) - [Beyond Kafka & RabbitMQ: Why NATS is the Future of Microservices Messaging](https://www.wearedevelopers.com/videos/1646-beyond-kafka-rabbitmq-why-nats-is-the-future-of-microservices-messaging) - [From DevOps to Scaled DevOps: How We’re Rebuilding Continuous Delivery as a Platform](https://www.wearedevelopers.com/videos/100018-from-devops-to-scaled-devops-how-we-re-rebuilding-continuous-delivery-as-a-platform) - [What Developers Get Wrong About Application Quality](https://www.wearedevelopers.com/videos/233-what-developers-get-wrong-about-application-quality) - [Platform Engineering vs. DevOps Why not both?](https://www.wearedevelopers.com/videos/885-platform-engineering-vs-devops-why-not-both) - [Shipping Faster with Less: Render on Cloud Hosting, AI Workloads, and the Future of DevOps](https://www.wearedevelopers.com/videos/1894-shipping-faster-with-less-render-on-cloud-hosting-ai-workloads-and-the-future-of-devops) ## Related Articles - [The Best X (Twitter) Accounts for Developers](https://www.wearedevelopers.com/magazine/294-the-best-x-twitter-accounts-for-developers) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [The Best Software Developer Blogs to Read](https://www.wearedevelopers.com/magazine/156-the-best-software-developer-blogs-to-read) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Find a Developer Job: 12 Best Job Sites For Developers](https://www.wearedevelopers.com/magazine/165-find-a-developer-job-12-best-job-sites-for-developers) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs)