> Markdown version of [/jobs/ext/1513396-platform-engineer](https://www.wearedevelopers.com/jobs/ext/1513396-platform-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Platform Engineer - **Company:** We ARE Recruitment Group - **Location:** Brussel, Belgium - **Contract:** Permanent contract - **Skills:** Application Programming Interfaces (APIs), Computing Platforms, User Authentication, Backup Devices, Bash Shell, Cluster Analysis, Computer Networks, Continuous Delivery, Continuous Integration, Software Debugging, Noise Reduction, Domain Name System (DNS), Monitoring of Systems, PostgreSQL, MongoDB, Object-Oriented Software Development, Open Source Technology, Role-Based Access Control, Redis, Prometheus, Reverse Proxy, TCP/IP, Management of Software Versions, Software Vulnerability Management, YAML, Cloud Platform System, Autoscaling, Grafana, Git, Kubernetes, Infrastructure Automation Frameworks, Hashicorp, Azure AKS, Api Gateway, Oracle Cloud Infrastructure, Software Version Control, Dynatrace, Docker, Vulnerability Analysis, Artifactory - **Published:** July 4, 2026 - **Apply:** https://www.adzuna.be/details/5787453848 ## About the Role Technical skills Kubernetes (Deep Production Expertise) Multi-cluster architecture & lifecycle management RBAC & least-privilege design Network policies & traffic segmentation Stateful workloads & storage strategy (CSI, PV/PVC) Autoscaling (HPA/VPA) & resource tuning Pod Security Standards Admission controllers Performance & reliability troubleshooting Cluster-level debugging (networking, DNS, scheduling, OOM, crash loops) GitOps & Continuous Delivery, HashiCorp Vault for dynamic secrets with CSI integration Image vulnerability scanning integration Supply chain security awareness TLS & certificate lifecycle management RBAC governance Observability & Reliability OpenTelemetry (metrics, logs, traces) Prometheus or VictoriaMetrics (recording rules, HA setup) Loki (log aggregation & LogQL) Tempo (distributed tracing) Grafana (advanced dashboards & alerting) SLI/SLO design & error budget thinking Alert noise reduction strategy Networking (Advanced) TCP/IP & DNS fundamentals TLS & mTLS concepts Kubernetes Services, Ingress & Reverse Proxy concepts East-west vs north-south traffic API routing & traffic management Network Policies implementation Automation Advanced Bash scripting Infrastructure automation mindset Nice to Have Kong API Gateway (api routing, plugins, authentication, rate limiting) Redis (operational knowledge: deployment, persistence, clustering, backups) PostgreSQL (migrations, backups, HA basics, Kubernetes deployment patterns) MongoDB (replica sets, backups, Kubernetes deployment patterns) Kargo on top ## Description The Platform Engineer in the Consumer Centricity Platform Operations team is responsible for the reliable, secure, and stable operation of the organization's high-availability cloud platform, built on Kubernetes and composed of multiple in-house platform components. The role focuses on platform lifecycle management, day-2 operations, incident response, and operational excellence, ensuring that customer-facing Web UIs and APIs remain available, performant, and secure 24/7. The Platform Engineer acts as a technical custodian of the platform, providing a stable foundation on which service teams can safely deploy and operate their workloads. Primary Objectives Maintain platform availability and reliability in accordance with SLOs/SLAs Ensure operational readiness of all environments (DEV / TEST / ACC / PROD) Provide 24/7 operational coverage for critical platform services (via on-call) Ensure the platform is observable, secure, well-controlled and documented Execute platform changes, upgrades, and maintenance in a predictable and low-risk manner, Troubleshoot Kubernetes-related failures: Pod lifecycle issues, networking problems, resource starvation Controlled rollouts with rollback plans Reliability & 24/7 Incident Response Participate in the 24/7 on-call rotation for critical services (incident responder) Lead or contribute to: Incident triage and mitigation Root Cause Analysis (RCA) Post-incident action tracking and follow-up Maintain and improve runbooks and operational procedures Observability & Monitoring Operate (and use) the open-source observability platform Ensure effective observability across the platform: Metrics, logs, and distributed traces Actionable alerts Reduced false positives Support incident analysis through correlation and telemetry inspection Change, Release & Maintenance Management Plan and execute platform changes Follow structured change management practices Stakeholder communication Ensure platform changes are documented and auditable Security & Compliance (Operational Focus) Operate platform security controls: RBAC, network boundaries, secret mgmt. Apply security updates and patches to platform components Support vulnerability remediation efforts Provide operational evidence for audits and security reviews Automation & Operational Improvement Automate repetitive operational tasks where appropriate Reduce operational risk through standardization and documented procedures Platform as Code approach (GitOps), Sync waves & hooks Drift detection & reconciliation Multi-environment promotion workflows Git-based deployment strategy with version management Declarative platform design with PR-driven changes YAML-based CI/CD pipelines with Harness.io Secure secret handling in CI/CD (with HashiCorp) Packaging & Configuration Helm (advanced chart authoring) Reusable library charts OCI-based registries Values layering strategy Kustomize overlays for multi-environment isolation and strategic patches Container & Artifact Management Docker (secure multi-stage builds, optimization) Harbor (RBAC, replication, vulnerability scanning) JFrog Artifactory (Docker & Helm registry management) Artifact versioning & promotion strategy Secrets & Security ## Related Videos - [One Platform Could Not Fit Them All](https://www.wearedevelopers.com/videos/1919-one-platform-could-not-fit-them-all) - [How a Small Team Shrank a Microsoft Monorepo by 94%](https://www.wearedevelopers.com/videos/1236-how-a-small-team-shrank-a-microsoft-monorepo-by-94) - [CI/CD with Github Actions](https://www.wearedevelopers.com/videos/856-ci-cd-with-github-actions) - [Reducing LLM Calls with Vector Search Patterns - Raphael De Lio (Redis)](https://www.wearedevelopers.com/videos/1714-reducing-llm-calls-with-vector-search-patterns-raphael-de-lio-redis) - [How I saved 200K/yr in direct costs writing 0 code lines in K8s](https://www.wearedevelopers.com/videos/1055-how-i-saved-200k-yr-in-direct-costs-writing-0-code-lines-in-k8s) - [Git for Code Reviews](https://www.wearedevelopers.com/videos/429-git-for-code-reviews) ## Related Articles - [Learning Kubernetes made easy with KubeCampus](https://www.wearedevelopers.com/magazine/348-learning-kubernetes-made-easy-with-kubecampus) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [Dev Digest 121 - AI goes offline](https://www.wearedevelopers.com/magazine/456-dev-digest-121-ai-goes-offline) - [Dev Digest 131 - AI'm not sure about OSS](https://www.wearedevelopers.com/magazine/472-dev-digest-131-ai-m-not-sure-about-oss) - [Dev Digest 138 - Are you secure about this?](https://www.wearedevelopers.com/magazine/486-dev-digest-138-are-you-secure-about-this) - [Dev Digest 119 - ❤️ === ❤️](https://www.wearedevelopers.com/magazine/454-dev-digest-119)