(Senior) DevOps Engineer
- Discuss this with your agent
- Open in Claude
- Open in ChatGPT
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Role details
Tech stack
+8 more
Job description
- Define and own SLOs and SLIs for the platform and manage error budgets against them
- Carry on-call, act as incident commander, and run blameless post-incident reviews that produce real follow-up
- Run production readiness and capacity planning ahead of demand, not after the page fires
Platform & Infrastructure
- Run and harden the Kubernetes platform (Helm, GitOps, service mesh) and the cloud underneath it (Terraform, multi-region)
- Own observability: metrics, logs, and distributed tracing, so problems surface before users feel them
Automation & Efficiency
- Eliminate toil through automation, self-healing systems, and automated remediation
- Drive cost visibility and FinOps practice across cloud and LLM spend, * Flexible Work Models: Full-time position on-site in Stuttgart, Hamburg or Munich (3 days per week) with flexible working hours.
- Benefits: Deutschland-Ticket, Wellpass fitness membership, access to the latest AI tools, and regular company and team off-sites.
- Top Team: International team with exceptional talents.
- High Growth Potential: Steep learning curve in a fast-growing AI startup. A high level of personal responsibility and the freedom to actively shape processes.
- Top Equipment: MacBook, iPhone, headset, and all the tools you need to perform at your best.
Blockbrain is an equal opportunity employer. We celebrate diversity and are committed to an inclusive work environment.
Requirements
- Professional Experience: 5+ years in DevOps, SRE, or platform engineering, ideally operating production SaaS at scale. Hands-on experience running Kubernetes in production is essential.
- Communication: Clear and calm under pressure. Can coordinate an incident and write a post-mortem others learn from. Works in English; German is a plus.
- Tech Affinity: Treats infrastructure as code and operations as a software discipline. Genuinely enjoys automating manual work away.
- Solution Orientation: Measures success in incidents that did not happen. Fixes root causes, not symptoms.
- Organizational Talent: Plans capacity and reliability work ahead of demand and balances on-call, project work, and toil reduction.
- Education: Degree in computer science or a related field, or equivalent hands-on experience. We care about what you can operate, not the certificate.
- Hard Skills: Kubernetes, Terraform / IaC, CI/CD (e.g. GitHub Actions), observability (Prometheus, Grafana, distributed tracing), a major cloud (AWS, Azure, or GCP), scripting (Python, Go, or TypeScript), secrets management, and policy as code.
- Soft Skills & Mindset: Strong ownership, blameless culture, a bias toward automation, calm in incidents, and a security-by-default mindset.
Benefits & conditions
€80.000 - €120.000 EUR
About the company
For mid-sized and enterprise companies in DACH, knowledge is the last real competitive lever. Yet knowledge gets stuck in tools and SharePoint folders - or disappears when experts leave. While IT is still planning, employees are already using ChatGPT and the like - without governance, and sensitive data is leaking out.
With the Knowledge Bots platform, Blockbrain creates what companies truly need: AI-powered knowledge management that is quick to implement, flexibly scales, and operates in a compliant manner. Teams use our GenAI building blocks to build tailored AI assistants, agents, and workflows in minutes - just like Lego.
- Series A funded with strong growth momentum (10x product usage & 5x revenue in 2025)
- Enterprise clients such as Roland Berger, Bosch, IONOS, and Harting from industries including manufacturing, finance, and legal - sectors with the highest security requirements
- ISO 27001 certified, EU AI Act ready. Made in Germany., At Blockbrain, we don’t just talk about AI - we use it every day. In this role, you will:
- Use coding agents to build automation, write infrastructure code, and reason through failure modes faster
- Automate operational toil and incident workflows with AI in the loop
- Collaborate with the product team to give real-world feedback on Blockbrain’s own tools from an operator’s perspective
- Stay curious about emerging AI capabilities and apply them to platform and reliability work
We’re not looking for AI experts - we’re looking for people who are genuinely open to working with AI as a daily co-pilot.
Apply for this position
This job is hosted externally. Click below to view the full posting and apply.
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Good distractions
Talks and stories from around this role — technically off-topic, practically not.
Moments
Explore playlistsVideos
See allRelated articles
See all
Dev Digest 121 - AI goes offline
Fullstack developer salary in Germany [2023]
The Biggest German Tech Companies
Top-Paying Tech Jobs (with Salaries)