> Markdown version of [/jobs/ext/1866566-platform-operations-engineer](https://www.wearedevelopers.com/jobs/ext/1866566-platform-operations-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Platform Operations Engineer - **Company:** Tamarind Intelligence - **Location:** Barcelona, Spain - **Contract:** Temporary contract - **Skills:** Agile Methodology, Airflow, Command-Line Interface, Computer Networks, Data Infrastructure, Software Debugging, Linux, Domain Name System (DNS), Apache Hive, Log Analysis, Networking Basics, Open Source Technology, OpenShift, Reliability Engineering, Prometheus, TCP/IP, Data Logging, Load Balancing, DevOps Tools - Open-source, Grafana, Mttr, Containerization, Kubernetes, Terraform, Docker - **Published:** August 2, 2026 - **Apply:** https://www.jobleads.com/es/job/e6ed0454a20b6ec30fcff281ad54eed6d ## About the Role * Experience: Minimum 5 years of proven experience in Systems Operations, L2/L3 Application Support, Site Reliability Engineering (SRE), or Infrastructure Operations. * Kubernetes & Docker (Core): Deep hands-on experience operating, debugging, and troubleshooting containerized workloads in Kubernetes and Docker environments (e.g., pod lifecycle, ingress/service issues, resource limits, volume mounts). * Troubleshooting Mindset: Exceptional diagnostic skills. You enjoy reading logs, analyzing metrics, tracing application failure modes, and asking the right questions to solve user-reported problems. * Linux Fundamentals: Confident command-line usage, high-level handling of Linux systems, process management, log analysis, and system recovery. * Operational DevOps Tooling: Practical experience managing deployments and platform states using tools like ArgoCD, Helm, or Terraform from an operations perspective. * Observability & Open Source Stack: Hands-on experience using monitoring and logging stacks (e.g., Prometheus, Grafana) to diagnose issues. Familiarity with applications like Airflow, Trino, Hive, or OpenShift is a strong plus. * Networking Basics: Practical understanding of network concepts (DNS, TCP/IP, load balancers, network policies) to troubleshoot connection or traffic routing issues. * Mindset: Strong user-empathy, autonomy to take ownership of production issues, and a supportive team player attitude focused on knowledge sharing. * Communication: Good English skills (written and spoken) to collaborate effectively with platform users and international teams. ## Description * Platform Operations & Stability: Take full responsibility for running, maintaining, and maintaining the operational health of our Kubernetes-based cloud Big Data platform. * Troubleshooting & L2/L3 Incident Response: Act as the primary point of contact for platform users (data engineers, analysts, internal teams). You will dive deep into application logs, investigate container failures, analyze pod behavior, and diagnose performance bottlenecks or process crashes. * User Application Operations: Support and maintain internal user applications running on top of Docker and Kubernetes. Ensure deployed services are healthy, scalable, and resilient. * Root Cause Analysis (RCA): Apply an investigation-first mindset to identify why platform or application failures occur, preventing recurring incidents rather than just applying temporary fixes. * Agile Collaboration & Support: Work in 2-week sprints, addressing operational tickets, incidents, and platform health tasks while keeping team members and internal stakeholders aligned. * Continuous Operational Improvement: Continuously refine operational runbooks, monitoring setups, and incident response definitions to improve mean time to resolution (MTTR). ## Related Videos - [Docker network without Docker](https://www.wearedevelopers.com/videos/1418-docker-network-without-docker) - [My journey into DevOps world - How it all started!](https://www.wearedevelopers.com/videos/545-my-journey-into-devops-world-how-it-all-started) - [What Developers Get Wrong About Application Quality](https://www.wearedevelopers.com/videos/233-what-developers-get-wrong-about-application-quality) - [Docker Compose: Rediscovered](https://www.wearedevelopers.com/videos/1978-docker-compose-rediscovered) - [Docker exec without Docker](https://www.wearedevelopers.com/videos/1094-docker-exec-without-docker) - [Docker build without Docker](https://www.wearedevelopers.com/videos/100114-docker-build-without-docker) ## Related Articles - [DevOps Engineer Salary [2023]](https://www.wearedevelopers.com/magazine/203-devops-engineer-salary-2023) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Learning Kubernetes made easy with KubeCampus](https://www.wearedevelopers.com/magazine/348-learning-kubernetes-made-easy-with-kubecampus) - [MLops – Deploying, Maintaining And Evolving Machine Learning Models in Production](https://www.wearedevelopers.com/magazine/115-mlops-deploying-maintaining-and-evolving-machine-learning-models-in-production) - [The Best X (Twitter) Accounts for Developers](https://www.wearedevelopers.com/magazine/294-the-best-x-twitter-accounts-for-developers) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers)