Sovereign Engineering Platform Sre

T-Systems Iberia
Granada, Spain
2 days ago
Apply on www.buscojobs.com.es
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience required
5 years minimum
Working hours
Regular working hours
Languages
English, German

Tech stack

Artificial Intelligence Audit Trail Cloud Computing Computer Networks Continuous Integration DevOps Disaster Recovery Python (Programming Language) Network Segmentation Octopus Deploy Systems Development Life Cycle Role-Based Access Control
+12 more
Reliability Engineering Ansible Prometheus Shell Script Vault (Revision Control System) Data Processing Okta Grafana Gitlab Kubernetes Terraform Jenkins

Job description

Full-time Company Description T?Systems is part of the Deutsche Telekom Group, with around ** employees worldwide.We create technology with purpose to generate a positive impact on society.We are looking for curious talent, eager to learn, take on challenges, and contribute ideas that transform our customers’ experience.We trust people: we offer autonomy, continuous support, and a collaborative environment where you can grow without limits.We are one global team, guided by respect, integrity, and a passion for doing better every day.Job Description Key responsibilities Build andoperateKubernetes environments that host AI engineering tools, internal model gateways, retrieval components, workflow services, CI/CD runners, and documentation services.ImplementGitOpsand Infrastructure as Code patterns for reproducible provisioning, configuration, policy enforcement, platform upgrades, and disaster recovery readiness.Manage private registries, package mirrors, secrets, identity integration, network segmentation, storage classes, backup routines, and controlled connectivity models.Provide observability for engineering workloads, including metrics, logs, traces, GPU and CPUutilization, service health, cost signals, and operational runbooks.Work with software, security, and architecture teams to ensure the platform supports AI-assisted SDLC workflows without creating uncontrolled data exposure or audit gaps.Examples of market tools, models, and platform components expected Platform tooling such as Kubernetes, Helm, Terraform, Ansible,ArgoCD,Crossplane, GitLab runners, Jenkins agents, private registries, and internal package mirrors.AI platform components such asvLLM,Ollama, OpenAI-compatible gateways,Qdrantor similar vector stores, OpenWebUI, Continue-compatible endpoints, and workflow services.Observability and operations stacks such as Prometheus, Grafana, Loki,OpenTelemetry, ELK/OpenSearch,Alertmanager, SRE runbooks, and incident management tooling.Security and governance components such as Vault,Keycloak, network policies, RBAC, admission controls, image scanning, SBOM tooling, and audit logging.Infrastructure awareness covering GPU-backed nodes, CPU-only fallback, storage performance, network isolation, proxy patterns,on-premiseenvironments, and dedicated landing zones.Qualifications 5+ years in SRE, platform engineering, DevOps, cloud infrastructure, or operations roles with strong Kubernetes and Linuxexpertise.Proven experience building and operating production-grade engineering platforms withGitOps, Infrastructure as Code, observability, and operational runbooks.Hands-on skills in Terraform, Ansible, Helm, Python or shell scripting, CI/CD runners, private registries, and secure configuration management.Good understanding of networking, storage, secrets, access control, monitoring, backup, disaster recovery, and operational hardening in high-security environments.Comfortable supporting AI-enabled engineering workloads in sovereignty-driven contexts where isolation, controlled data handling, reliability, and auditability are mandatory.Additional Information Whatdoweofferyou?T-Social: socialinitiatives(sports,community,health, …).Hybridworkmodel(remote/on-site).Flexibleworkinghours.Growth&development Weeklylanguageclasses(English & German).InternationalMentoringSessions&ExperienceDays.Flexiblecompensationplan (healthinsurance,mealvouchers,childcare,transport).Socialfund.Wellbeing& time off 26+workingdaysofvacationperyear.And many more advantages of being part of T-Systems!#J-*****-Ljbffr

Requirements

Qualifications 5+ years in SRE, platform engineering, DevOps, cloud infrastructure, or operations roles with strong Kubernetes and Linuxexpertise. Proven experience building and operating production-grade engineering platforms withGitOps, Infrastructure as Code, observability, and operational runbooks. Hands-on skills in Terraform, Ansible, Helm, Python or shell scripting, CI/CD runners, private registries, and secure configuration management. Good understanding of networking, storage, secrets, access control, monitoring, backup, disaster recovery, and operational hardening in high-security environments. Comfortable supporting AI-enabled engineering workloads in sovereignty-driven contexts where isolation, controlled data handling, reliability, and auditability are mandatory.

Benefits & conditions

Flexibleworkinghours. Growth&development Weeklylanguageclasses(English & German). InternationalMentoringSessions&ExperienceDays. Flexiblecompensationplan (healthinsurance,mealvouchers,childcare,transport). Socialfund. Wellbeing& time off 26+workingdaysofvacationperyear. And many more advantages of being part of T-Systems! #J-*****-Ljbffr

About the company

Full-time Company Description T?Systems is part of the Deutsche Telekom Group, with around ** employees worldwide. We create technology with purpose to generate a positive impact on society. We are looking for curious talent, eager to learn, take on challenges, and contribute ideas that transform our customers’ experience. We trust people: we offer autonomy, continuous support, and a collaborative environment where you can grow without limits. We are one global team, guided by respect, integrity, and a passion for doing better every day. Job Description Key responsibilities Build andoperateKubernetes environments that host AI engineering tools, internal model gateways, retrieval components, workflow services, CI/CD runners, and documentation services. ImplementGitOpsand Infrastructure as Code patterns for reproducible provisioning, configuration, policy enforcement, platform upgrades, and disaster recovery readiness. Manage private registries, package mirrors, secrets, identity integration, network segmentation, storage classes, backup routines, and controlled connectivity models. Provide observability for engineering workloads, including metrics, logs, traces, GPU and CPUutilization, service health, cost signals, and operational runbooks. Work with software, security, and architecture teams to ensure the platform supports AI-assisted SDLC workflows without creating uncontrolled data exposure or audit gaps. Examples of market tools, models, and platform components expected Platform tooling such as Kubernetes, Helm, Terraform, Ansible,ArgoCD,Crossplane, GitLab runners, Jenkins agents, private registries, and internal package mirrors. AI platform components such asvLLM,Ollama, OpenAI-compatible gateways,Qdrantor similar vector stores, OpenWebUI, Continue-compatible endpoints, and workflow services. Observability and operations stacks such as Prometheus, Grafana, Loki,OpenTelemetry, ELK/OpenSearch,Alertmanager, SRE runbooks, and incident management tooling. Security and governance components such as Vault,Keycloak, network policies, RBAC, admission controls, image scanning, SBOM tooling, and audit logging.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.buscojobs.com.es
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

1:34 min

Pivoting careers into specialized platform engineering roles

Xavier Portilla Edo · LIVE

6:14 min

Structuring CI/CD pipelines with integrated security and quality checks

Christoph Ruggenthaler · LIVE

2:17 min

Mapping the maturity roadmap for scaled devops adoption

Dominik Krichbaum Dominik Krichbaum · World Congress 2026 Europe

2:33 min

Introduction to security advocacy and automation testing

Chris Heilmann +2 · LIVE

4:54 min

Implementing geographic salary tiers for compensation equity and fairness

Rudi Bauer Rudi Bauer +1 · Cappuccino with HR

3:18 min

Scaling global network engineering through DevOps culture

Stuart Clark · LIVE

Videos

See all

Related articles

See all