> Markdown version of [/jobs/ext/155030-senior-site-reliability-engineer](https://www.wearedevelopers.com/jobs/ext/155030-senior-site-reliability-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Senior Site Reliability Engineer - **Company:** Loadup Technologies - **Location:** Alpharetta, GA, United States - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** Query Performance, Java (Programming Language), Amazon Web Services, Amazon Elastic Compute Cloud, Amazon S3, Software as a Service, Databases, Relational Databases, DevOps, Fault Tolerance, Identity and Access Management, Key Management, PostgreSQL, Network Segmentation, Node.Js, Performance Tuning, Reliability Engineering, Data Logging, Database Optimization, Spring-boot, Reliability of Systems, Infrastructure as Code (IaC), Backend, Cloudformation, Amazon Relational Database Service, Kubernetes, Infrastructure Automation Frameworks, Route53, Cloudwatch, Api Gateway, Terraform, Microservices - **Published:** May 24, 2026 - **Apply:** https://www.indeed.com/viewjob?jk=4aa6e5d4c8040aea ## About the Role Do you have experience in Tooling?, * Experience: Bachelor's degree in CS/Engineering OR 8+ years of practical SRE/DevOps experience provisioning large-scale, multi-tenant SaaS applications. * AWS Mastery: Extensive hands-on experience with services including EC2, ECS/EKS, RDS, S3, CloudWatch, IAM, Route 53, and API Gateway. * IaC & Automation: Deep, proven experience with Infrastructure as Code (IaC) tools like Terraform. * Container Orchestration: Hands-on experience with orchestration platforms supporting backend microservices (Java/Spring Boot, Node.js). * Database Reliability: Strong knowledge of PostgreSQL optimization, connection pooling, and query performance tuning. * Production Support: Proven experience handling production support, on-call rotations, and running blameless post-mortems., * An Ownership-Driven Engineer who values long-term system reliability and architectural resilience over short-term hotfixes. * An Elite Communicator capable of translating complex backend infrastructure architecture into clear concepts for non-technical stakeholders. * A Security Champion who naturally designs with least-privilege access and data isolation patterns in mind. ## Description As the Senior Site Reliability Engineer (SRE), you will be critical in ensuring the reliability, scalability, and security of LoadUp's next-generation platform-a modern, cloud-native microservices architecture built for enterprise scale. You will work closely with the CTO, Director of Architecture, and engineering teams to build, optimize, and maintain a world-class AWS infrastructure to support our mission-critical, multi-tenant enterprise systems., * Infrastructure Management: Design, provision, and manage a scalable AWS infrastructure supporting our cloud-native microservices platform. * Reliability & Observability: Build and maintain frameworks for logging, metrics, tracing, and dashboards to ensure deep visibility into system health. * Incident Leadership: Lead incident management and post-mortem processes, driving root cause analysis and long-term remediation. * Security & Compliance: Enforce infrastructure security best practices (IAM policies, secrets management, network segmentation) and support SOC 2 requirements. * Infrastructure as Code (IaC): Develop and maintain IaC using tools such as Terraform or AWS CloudFormation to ensure consistent, repeatable provisioning. * Database Optimization: Manage and optimize relational databases (PostgreSQL) at scale, including replication, failover, and performance tuning. * Capacity & Efficiency: Guide architectural reviews for resilience and fault tolerance while driving capacity planning and cloud cost optimization. ## Related Videos - [SRE Methods In an Agency Environment](https://www.wearedevelopers.com/videos/348-sre-methods-in-an-agency-environment) - [Stop using Node.js like in 2020! What changed and what you can do today with Node.js](https://www.wearedevelopers.com/videos/100011-stop-using-node-js-like-in-2020-what-changed-and-what-you-can-do-today-with-node-js) - [From DevOps to Scaled DevOps: How We’re Rebuilding Continuous Delivery as a Platform](https://www.wearedevelopers.com/videos/100018-from-devops-to-scaled-devops-how-we-re-rebuilding-continuous-delivery-as-a-platform) - [Reliable scalability: How Amazon.com scales on AWS](https://www.wearedevelopers.com/videos/983-reliable-scalability-how-amazon-com-scales-on-aws) - [Stop Using Node.js Like It’s 2020! - Alfonso Graziano](https://www.wearedevelopers.com/videos/1863-stop-using-node-js-like-it-s-2020-alfonso-graziano) - [DevOps Maturity Check – a way to balance autonomy and alignment](https://www.wearedevelopers.com/videos/58-devops-maturity-check-a-way-to-balance-autonomy-and-alignment) ## Related Articles - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Find a Developer Job: 12 Best Job Sites For Developers](https://www.wearedevelopers.com/magazine/165-find-a-developer-job-12-best-job-sites-for-developers) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Dev Digest 121 - AI goes offline](https://www.wearedevelopers.com/magazine/456-dev-digest-121-ai-goes-offline)