Back to jobs
New

(Senior) DevOps Engineer (m/f/d)

Stuttgart, Hamburg, Munich

About Blockbrain

For mid-sized and enterprise companies in DACH, knowledge is the last real competitive lever. Yet knowledge gets stuck in tools and SharePoint folders — or disappears when experts leave. While IT is still planning, employees are already using ChatGPT and the like — without governance, and sensitive data is leaking out.

With the Knowledge Bots platform, Blockbrain creates what companies truly need: AI-powered knowledge management that is quick to implement, flexibly scales, and operates in a compliant manner. Teams use our GenAI building blocks to build tailored AI assistants, agents, and workflows in minutes — just like Lego.

  • Series A funded with strong growth momentum (10x product usage & 5x revenue in 2025)

  • Enterprise clients such as Roland Berger, Bosch, IONOS, and Harting from industries including manufacturing, finance, and legal — sectors with the highest security requirements

  • ISO 27001 certified, EU AI Act ready. Made in Germany.


Role & Impact

You join the Platform team and own the reliability of the systems our AI agents run on. This is a reliability engineering role: you treat operations as a software problem, so where others run a manual procedure, you write the automation that makes it unnecessary. You define what "healthy" means in numbers, measure it, and hold the line on it in production. When something breaks, you bring the system back and then make sure it cannot break the same way twice.

Reliability & Operations

  • Define and own SLOs and SLIs for the platform and manage error budgets against them

  • Carry on-call, act as incident commander, and run blameless post-incident reviews that produce real follow-up

  • Run production readiness and capacity planning ahead of demand, not after the page fires

Platform & Infrastructure

  • Run and harden the Kubernetes platform (Helm, GitOps, service mesh) and the cloud underneath it (Terraform, multi-region)

  • Own observability: metrics, logs, and distributed tracing, so problems surface before users feel them

Automation & Efficiency

  • Eliminate toil through automation, self-healing systems, and automated remediation

  • Drive cost visibility and FinOps practice across cloud and LLM spend

Security

  • Bake security into the platform: least-privilege access, secrets management, policy as code, and vulnerability management

 

Expected AI Skills

At Blockbrain, we don't just talk about AI — we use it every day. In this role, you will:

  • Use coding agents to build automation, write infrastructure code, and reason through failure modes faster

  • Automate operational toil and incident workflows with AI in the loop

  • Collaborate with the product team to give real-world feedback on Blockbrain's own tools from an operator's perspective

  • Stay curious about emerging AI capabilities and apply them to platform and reliability work

We're not looking for AI experts — we're looking for people who are genuinely open to working with AI as a daily co-pilot.

Your Profile

  • Professional Experience: 5+ years in DevOps, SRE, or platform engineering, ideally operating production SaaS at scale. Hands-on experience running Kubernetes in production is essential.

  • Communication: Clear and calm under pressure. Can coordinate an incident and write a post-mortem others learn from. Works in English; German is a plus.

  • Tech Affinity: Treats infrastructure as code and operations as a software discipline. Genuinely enjoys automating manual work away.

  • Solution Orientation: Measures success in incidents that did not happen. Fixes root causes, not symptoms.

  • Organizational Talent: Plans capacity and reliability work ahead of demand and balances on-call, project work, and toil reduction.

  • Education: Degree in computer science or a related field, or equivalent hands-on experience. We care about what you can operate, not the certificate.

  • Hard Skills: Kubernetes, Terraform / IaC, CI/CD (e.g. GitHub Actions), observability (Prometheus, Grafana, distributed tracing), a major cloud (AWS, Azure, or GCP), scripting (Python, Go, or TypeScript), secrets management, and policy as code.

  • Soft Skills & Mindset: Strong ownership, blameless culture, a bias toward automation, calm in incidents, and a security-by-default mindset.

 

What Blockbrain Offers

  • Flexible Work Models: Full-time position on-site in Stuttgart, Hamburg or Munich (3 days per week) with flexible working hours.

  • Benefits: Deutschland-Ticket, Wellpass fitness membership, access to the latest AI tools, and regular company and team off-sites.

  • Top Team: International team with exceptional talents.

  • High Growth Potential: Steep learning curve in a fast-growing AI startup. A high level of personal responsibility and the freedom to actively shape processes.

  • Top Equipment: MacBook, iPhone, headset, and all the tools you need to perform at your best.

Blockbrain is an equal opportunity employer. We celebrate diversity and are committed to an inclusive work environment.

Compensation

€80.000 - €120.000 EUR

Apply for this job

*

indicates a required field

Phone
Resume/CV*

Accepted file types: pdf, doc, docx, txt, rtf


Select...
Select...