> Markdown version of [/jobs/ext/2992572-site-reliability-engineer](https://www.wearedevelopers.com/jobs/ext/2992572-site-reliability-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Site Reliability Engineer - **Company:** Altium LLC - **Location:** Cambridge, UK - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** .NET Framework, Altium Designer, Amazon Web Services, Application Performance Management, Software as a Service, Cloud Computing, Continuous Integration, Relational Databases, DevOps, Github, PostgreSQL, MySQL, Networking Basics, Octopus Deploy, Systems Development Life Cycle, Reliability Engineering, Ansible, Newrelic, Software Engineering, Software Systems, Data Logging, Cloud Platform System, Grafana, Reliability of Systems, Infrastructure as Code (IaC), Gitlab, Kubernetes, Terraform, Pagerduty, Jenkins, Microservices - **Published:** September 19, 2026 - **Apply:** https://www.careerjet.co.uk/jobad/gb7e0bd0fa6abb5cb395df5fa6842c8253 ## About the Role * 6+ years in SRE, DevOps or related role in a large-scale environment * Software development experience (ideally working with and as a .NET developer) * Strong understanding of SDLC, microservice and HA architecture * Observability - NewRelic, ELK, Grafana, PagerDuty, OTEL or similar * Experience with Kubernetes clusters in production setting, AWS, IOC * Experience with operational tasks * Knowledge of CI-CD tooling Jenkins, Gitlab, GitHub, ArgoCD or similar * Knowledge of IaaC Terraform, Ansible * Basic knowledge of networking fundamentals * Experience with relational databases (mysql, postgres) as a plus ## Description Senior Site Reliability Engineer ensures the reliability, availability, and performance of large-scale software systems through a blend of software engineering and systems administration. Key responsibilities involve automating operational tasks, improving observability, and contributing to incident management, while also collaborating with development and technology teams to build more reliable and scalable applications. Join Altium as a Senior Site Reliability Engineer to ensure the reliability and performance of the Altium Cloud Platforms., * Understanding how an Altium Cloud Platform works * Pioneer improvements in observability, including logging, monitoring, and application performance management (APM), ensuring system reliability and proactive issue detection. * Develop and implement reliability frameworks and patterns that standardize and elevate the resilience of our SaaS products across multiple regions and environments. * Cultivate a shared responsibility model where the SRE team collaborates with and educates engineering teams on reliability best practices. * Contribute to incident response and management, ensuring rapid resolution, clear stakeholder communication, and post-incident analysis for continuous improvement. * Participate in system design consulting, platform management, infrastructure upgrades and capacity planning. * Partner closely with engineering and development teams to enhance product stability, observability, and manageability through best practices in reliability engineering. * Partner closely with DevOps/Operations, drive automation initiatives, promote Infrastructure as Code (IaC), and streamline deployment processes to improve operational efficiency and scalability. * Champion Service-Oriented Organization (SOO) principles to ensure accountability and clarity in service ownership. ## Related Videos - [WeAreDevelopers LIVE - Modern DevOps for IoT Devices and More](https://www.wearedevelopers.com/videos/1805-wearedevelopers-live-modern-devops-for-iot-devices-and-more) - [SRE Methods In an Agency Environment](https://www.wearedevelopers.com/videos/348-sre-methods-in-an-agency-environment) - [From DevOps to Scaled DevOps: How We’re Rebuilding Continuous Delivery as a Platform](https://www.wearedevelopers.com/videos/100018-from-devops-to-scaled-devops-how-we-re-rebuilding-continuous-delivery-as-a-platform) - [MySQL Protocol Features You Should Be Aware Of](https://www.wearedevelopers.com/videos/100267-mysql-protocol-features-you-should-be-aware-of) - [Enabling automated 1-click customer deployments with built-in quality and security](https://www.wearedevelopers.com/videos/83-enabling-automated-1-click-customer-deployments-with-built-in-quality-and-security) - [Leading with Reliability: Applying SRE Principles to Build Stronger Engineering Organizations](https://www.wearedevelopers.com/videos/100185-leading-with-reliability-applying-sre-principles-to-build-stronger-engineering-organizations) ## Related Articles - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Where To Find Software Engineering Jobs](https://www.wearedevelopers.com/magazine/396-where-to-find-software-engineering-jobs) - [Find a Developer Job: 12 Best Job Sites For Developers](https://www.wearedevelopers.com/magazine/165-find-a-developer-job-12-best-job-sites-for-developers)