(Senior) Site Reliability Engineer in Berlin or Konstanz
- Discuss this with your agent
- Open in Claude
- Open in ChatGPT
Role details
Tech stack
+20 more
Job description
Experteer Overview In this role you will design, build, and operate KNIME’s next-generation cloud platform with a focus on reliability, security, and cost-efficiency. You will collaborate with multiple development teams to bring KNIME products into production-ready environments at scale. You will automate deployments, manage infrastructure as code, and drive operational standards across teams. Your work will emphasize observability, on-call readiness, and cost-aware optimizations, shaping a scalable SaaS platform. Pay / Benefits * Automate deployment and operations of large-scale SaaS systems using code * Build infrastructure as code with Helm, Terraform, CloudFormation, and Azure ARM * Participate in on-call rotations, triage incidents, perform root cause analysis and resolution * Set and communicate deployment standards for reliability, scalability, traceability, and monitoring * Instrument deployed systems for performance, reliability, and cost effectiveness * Collaborate with product and engineering teams to drive adoption of reliability standards * Lead planning with product teams and manage dependency risk across teams * Embed with product and engineering to ensure consistent adoption of operational standards Tasks * Cloud certifications on AWS, Kubernetes, Linux or similar technologies * Strong cloud experience with AWS or Azure (ideally both) and knowledge of VPC, IAM, EKS, ECR, EC2, S3, RDS, CloudWatch and equivalent Azure services * Experience deploying software in Kubernetes and understanding of Kubernetes patterns * Scripting in Python and Shell; knowledge of Go and Java is a plus * Linux system knowledge; networking expertise including security, routing, load balancers, firewalls * Knowledge of telemetry practices: metrics, logging, tracing in large distributed systems * Knowledge of OAuth/OIDC providers like Keycloak * Knowledge of PostgreSQL or similar relational databases * Ability to work independently and in a cross-cultural, distributed team * Strong communication skills across an organization Key requirements * Hybrid working * Flexible hours * Subsidised sports or yoga courses * Physiotherapy * Flu shots at select locations * Learning opportunities
Requirements
Helm, and engineering teams to drive adoption of reliability standards * Lead planning with product teams and manage dependency risk across teams * Embed with product and engineering to ensure consistent adoption of operational standards Tasks * Cloud certifications on AWS, Kubernetes, Linux or similar technologies * Strong cloud experience with AWS or Azure (ideally both) and knowledge of VPC, IAM, EKS, ECR, EC2, S3, RDS, CloudWatch and equivalent Azure services * Experience deploying software in Kubernetes and understanding of Kubernetes patterns * Scripting in Python and Shell; knowledge of Go and Java is a plus * Linux system knowledge; networking expertise including security, routing, load balancers, firewalls * Knowledge of telemetry practices: metrics, logging, tracing in large distributed systems * Knowledge of OAuth/OIDC providers like Keycloak * Knowledge of PostgreSQL or similar relational databases * Ability to work independently and in a cross-cultural, distributed team * Strong communication skills across an organization Key requirements * Hybrid working * Flexible hours * Subsidised sports or yoga courses * Physiotherapy * Flu shots at select locations * Learning opportunities
Apply for this position
This job is hosted externally. Click below to view the full posting and apply.
Apply on eu.experteer.comGood distractions
Talks and stories from around this role — technically off-topic, practically not.
Moments
Explore playlistsVideos
See allRelated articles
See all
Finding Jobs in Germany
Backend Developer Salary in Germany [2023]
How to Find Tech Jobs in Berlin
The Most Popular IT Jobs on the Market