📢 New: get today's jobs on our WhatsApp Channel
Jobiglo

No results.

This job is no longer available

This job expired on 06/10/2026. It no longer accepts applications.

Team Lead - Site Reliability Engineering (all genders)

FactFinder · Berlin

Senior 🇬🇧 English
Kubernetes Harvester Longhorn VLAN load balancing ingress GitOps Argo CD Flux custom Kubernetes operator HPA VPA KEDA cluster autoscaler capacity planning AI tools

Job description

About the role

FACT-Finder is looking for a Team Lead to own the reliability, scalability and cost of its on‑premise and cloud hosting environments. You will drive the migration to a Kubernetes‑based platform on Harvester and lead a four‑person hosting team.

Key responsibilities

  • Maintain operational health of hosting across Frankfurt and Stockholm data‑centers and cloud, covering availability, performance and incident management.
  • Design and implement the Kubernetes‑on‑Harvester platform, including cluster topology, storage (Longhorn), networking (VLAN, load balancing, ingress), backup and disaster‑recovery.
  • Build production‑grade K8s platform features: lifecycle management, upgrades, RBAC, secrets, GitOps (Argo CD / Flux), observability and policy guardrails.
  • Develop and operate the NG Search Operator and auto‑scaling solutions (HPA, VPA, KEDA, cluster autoscaler).
  • Define the hybrid on‑prem/cloud model, manage workload placement, burst capacity, latency and cost while keeping the architecture portable.
  • Own capacity planning, cost management and turn cost into a controlled lever.
  • Lead, mentor and grow the hosting team, set technical direction and foster a culture of ownership.
  • Integrate AI tools into daily operations for diagnosis, automation, monitoring and insight.

Required profile

  • Strong background in infrastructure or platform engineering across on‑premise and cloud environments.
  • Proven people‑leadership experience with excellent communication and stakeholder‑management skills.
  • Hands‑on experience migrating bare‑metal or classic VMs to a Kubernetes‑based platform, including stateful workloads and storage migration.
  • Fluent English; German is a plus.

Required skills

  • Kubernetes (cluster lifecycle, upgrades, networking, storage, RBAC, observability)
  • Harvester or comparable HCI/virtualisation platforms (e.g., KubeVirt, vSphere/ESXi, OpenStack)
  • GitOps tools – Argo CD, Flux
  • Longhorn storage
  • VLAN, load balancing, ingress controllers
  • Custom Kubernetes operators / CRDs
  • Auto‑scaling primitives – HPA, VPA, KEDA, cluster autoscaler
  • Capacity planning and cost management for hybrid on‑prem/cloud setups
  • AI‑assisted operational tools

Questions fréquentes

Le salaire n'est pas communiqué publiquement par le recruteur. Vous pouvez postuler et négocier directement avec FactFinder.
Cliquez sur "Postuler maintenant" en haut de la page. Vous pouvez importer votre CV en 1 clic — Jobiglo extrait automatiquement vos informations et postule pour vous.

Why are you reporting this job?

Thank you for your report. We will review this job.

A question about this job?

Ask it here: you will get the full job summary by e-mail, right away.

💬 Chat with us on Telegram

Published 1 month ago

29 views · 0 interested

Boost your chances

Upload your CV — we will match you with relevant openings.

Analyzing your CV...

FactFinder

Berlin