> Markdown version of [/jobs/ext/298098-site-reliability-engineering-lead](https://www.wearedevelopers.com/jobs/ext/298098-site-reliability-engineering-lead). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Site Reliability Engineering Lead - **Company:** Intigriti - **Location:** Antwerpen, Belgium - **Experience:** Expert - **Salary:** €83,200.0 - **Contract:** Permanent contract - **Skills:** Amazon Web Services, Backup Devices, Cloud Computing, Cloud Computing Security, Cyber Security, Computer Engineering, Disaster Recovery, Domain Name System (DNS), Elasticsearch, Identity and Access Management, Intrusion Detection and Prevention, Python (Programming Language), Key Management, PostgreSQL, Linux System Administration, MongoDB, RabbitMQ, Redis, Power BI, TypeScript, Software Vulnerability Management, Data Logging, Transport Layer Security, Load Balancing, Cloud Platform System, Istio, Cloudformation, Kubernetes, Infrastructure Automation Frameworks, Information Technology, Terraform, Security Orchestration, Automation & Response, Golang - **Published:** June 18, 2026 - **Apply:** https://intigriti.jobs.personio.com/job/2675617 ## About the Role * Proven experience leading and mentoring a team of SREs or operations professionals. Ability to inspire and motivate team members to achieve high performance. * Experience driving technical initiatives and delivering organizational improvements. Tech Skills * Bachelor's degree or equivalent experience in Computer Science, Computer Engineering, or Information Technology. * Strong Linux systems administration and troubleshooting experience. * Extensive experience with AWS or comparable cloud platforms. * Strong experience with Kubernetes and containerised workloads. * Experience designing and operating secure cloud environments. * Experience with Infrastructure as Code (Terraform, CDK, CloudFormation or equivalent). * Strong understanding of networking, DNS, load balancing, TLS, and cloud networking concepts. * Experience with CI/CD pipelines and software delivery practices. * Experience with monitoring, observability, and logging platforms. * Experience with identity and access management (IAM) and secrets management. * Experience writing software or automation in at least one programming language (TypeScript, Go, Python, or equivalent). Soft Skills * Fluent in English, professional proficiency of Dutch is a plus. * Proven technical skills with a 'can-do' & 'fix-it' mentality. * Willingness to share knowledge and experience with others, as well as be open to new ideas. * Strong critical thinking, troubleshooting and problem-solving skills with attention to detail. * Stress resistance with a clear focus on the resolution in an incident context. * Open to flexible working hours and willingness to take part in a 24x7 support organization. Nice to haves * Experience with Redis, RabbitMQ, PostgreSQL, MongoDB, MinIO, Elasticsearch, Graylog, cert-manager, or PowerBI. * Experience with service meshes such as Istio. * Experience operating multi-account cloud environments. * Experience working in cybersecurity or security-focused organizations. * Experience with incident response and security exercises. ## Description As the Site Reliability Engineering (SRE) Lead, your mission is to ensure the reliability, security, scalability, and performance of Intigriti's cloud platform and supporting systems. You are responsible for designing, implementing, and continuously improving the infrastructure, tooling, and operational practices that enable our engineering teams to deliver reliable and secure products to our customers. As a cybersecurity company, security is a core responsibility of the role. You will work closely with Engineering, Security, and Product teams to establish secure-by-default platform standards, maintain a strong security posture, and ensure our cloud environment remains resilient against evolving threats. Intigriti embraces strong mentorship and professional development as part of its culture. As SRE Lead, you are responsible for supporting the growth of the SRE team through coaching, mentorship, regular one-on-one meetings, and technical leadership. You proactively improve the reliability and security of our platform through automation, platform engineering, infrastructure modernization, and security initiatives. You also coordinate operational and security incident response activities, drive continuous improvement, and ensure lessons learned are translated into lasting improvements across the organization. What you'll be doing Technical Leadership * Provide technical leadership to the SRE team and act as a trusted advisor to engineering teams across the business. * Define and maintain platform engineering, cloud infrastructure, reliability, and security standards. * Drive architectural decisions that improve scalability, reliability, maintainability, and security. * Foster a culture of ownership, continuous improvement, collaboration, and operational excellence. * Collaborate with Engineering, Security, Product, and Leadership teams to align platform initiatives with business objectives. Platform Engineering & Automation * Drive the development of automation, self-service tooling, and infrastructure-as-code practices. * Reduce operational overhead through automation and process improvement. * Build and maintain reusable platform capabilities that enable engineering teams to operate efficiently and securely. * Lead cloud platform modernization initiatives and continuously improve platform reliability and developer experience. * Ensure infrastructure solutions remain cost-effective and operationally efficient. Reliability & Operational Excellence * Lead the response and coordination of production incidents and major service disruptions. * Drive root cause analysis and ensure identified improvements are implemented. * Establish and maintain monitoring, alerting, observability, and operational health practices across the platform. * Develop and maintain operational runbooks, recovery procedures, and platform documentation. * Lead disaster recovery, backup, resilience testing, and business continuity initiatives. * Ensure critical systems can be restored and operated effectively during failure scenarios. Platform Security * Design and maintain secure-by-default cloud and platform architectures. * Establish and enforce security best practices across cloud infrastructure, Kubernetes environments, networking, and identity management. * Partner closely with the Security team to strengthen Intigriti's overall security posture. * Drive infrastructure hardening, vulnerability remediation, secrets management, access control, and security automation initiatives. * Support the implementation and operation of security tooling, including endpoint security, cloud security controls, logging, and detection capabilities. * Participate in security architecture reviews and risk assessments for new platform initiatives. * Help establish foundational cloud security practices and security guardrails across the organization. Security Readiness & Governance * Support security incident response activities and coordinate platform remediation efforts. * Contribute to tabletop exercises, disaster recovery exercises, and security readiness initiatives. * Collaborate with Security, Compliance, and Engineering teams to support ISO27001, SOC2, customer security reviews, and audit activities. * Help maintain evidence, documentation, and technical controls required to meet security and compliance obligations. * Partner with internal stakeholders to continuously improve Intigriti's security posture and operational maturity. Mentorship & People Development * Mentor and coach SRE team members, supporting their professional growth and development. * Conduct regular one-on-one meetings and career development discussions. * Support hiring, onboarding, and performance review activities. * Encourage knowledge sharing and continuous learning across the team. * Stay informed on emerging technologies, reliability practices, and security trends, sharing relevant insights with the organization. ## Related Videos - [Reducing LLM Calls with Vector Search Patterns - Raphael De Lio (Redis)](https://www.wearedevelopers.com/videos/1714-reducing-llm-calls-with-vector-search-patterns-raphael-de-lio-redis) - [Rate-limiting using eBPF and Istio: How to protect your SaaS customers from themselves](https://www.wearedevelopers.com/videos/100220-rate-limiting-using-ebpf-and-istio-how-to-protect-your-saas-customers-from-themselves) - [Go with the Flow: Stop the Leaks Before Your Memory's a Waterfall!](https://www.wearedevelopers.com/videos/100073-go-with-the-flow-stop-the-leaks-before-your-memory-s-a-waterfall) - [SRE Methods In an Agency Environment](https://www.wearedevelopers.com/videos/348-sre-methods-in-an-agency-environment) - [Retooling and refactoring - an investment in people.](https://www.wearedevelopers.com/videos/371-retooling-and-refactoring-an-investment-in-people) - [Get started with securing your cloud-native Java microservices applications](https://www.wearedevelopers.com/videos/123-get-started-with-securing-your-cloud-native-java-microservices-applications) ## Related Articles - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Best Companies in the Netherlands: Top 25 Companies in 2023 ](https://www.wearedevelopers.com/magazine/193-best-companies-in-the-netherlands-top-25-companies-in-2023) - [Find a Developer Job: 12 Best Job Sites For Developers](https://www.wearedevelopers.com/magazine/165-find-a-developer-job-12-best-job-sites-for-developers) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [Where To Find Software Engineering Jobs](https://www.wearedevelopers.com/magazine/396-where-to-find-software-engineering-jobs)