Sign up to save this job, get alerts, and apply with an optimized CV.

Team Lead - Site Reliability Engineering (all genders)

Full Time Lead

Job description

Introduction

FACT-Finder develops product discovery technology for eCommerce and is used by leading online shops in Europe with the products Next Generation and Infinity. We are currently consistently modernizing our hosting towards Kubernetes on Harvester – as an on-prem hybrid with the option to scale fully into the cloud in the medium term. As Team Lead Site Reliability Engineering (all genders), you will be responsible for the reliability, scalability, and costs of our hosting environments, drive this transformation end-to-end, and lead the team that implements it.

Your Responsibilities

  • You will be responsible for the operational health of our hosting across on-premise (Frankfurt, Stockholm) and cloud – availability, performance, incident management.
  • You will actively drive the modernization towards Kubernetes on Harvester: cluster topology, storage (Longhorn), networking (VLAN, load balancing, ingress), backup, and disaster recovery.
  • You will build a production-ready k8s platform: lifecycle, upgrades, RBAC, secrets, GitOps (Argo CD / Flux), observability, and policy guardrails.
  • You will design the NG Search Operator (Custom Kubernetes Operator) and solve auto-scaling (HPA, VPA, KEDA, Cluster Autoscaler) for the current architecture.
  • You will concretely define our on-prem hybrid model: which workloads run where, how we burst into the cloud, how we keep latency and costs under control – and keep the architecture portable enough for a later cloud-only step.
  • You will be responsible for capacity planning and hosting costs and make costs a specifically controllable lever.
  • You will lead and develop our current 4-person hosting team in terms of expertise and discipline, be responsible for performance, and shape the technical standards and ownership culture.
  • You will make AI a permanent component of our operations: diagnosis, automation, monitoring, and insight.

Your Profile

  • Solid background in infrastructure or platform engineering across on-premise and cloud.
  • Hands-on depth with Kubernetes in production: cluster lifecycle, upgrades, networking, storage, RBAC, observability, GitOps delivery.
  • Demonstrably strong leadership experience, excellent communication, and stakeholder management.
  • Ideally, practical experience with Harvester or comparable HCI/virtualization platforms (KubeVirt, vSphere/ESXi, OpenStack).
  • Experience with a real migration from bare metal / classic VMs to a k8s-based platform – including stateful workloads, storage migration, cutover, and rollback.
  • Proficient in using Kubernetes Operators (Custom Controllers / CRDs), ideally for stateful systems such as search, databases, or streaming.
  • Solid understanding of auto-scaling primitives (HPA, VPA, Cluster Autoscaler, KEDA) and their interaction with capacity planning.
  • Experience with on-prem hybrid architectures and responsibility for the reliability, capacity, and costs of productive systems.
  • Hands-on fluency in using AI tools in operational business.
  • Very good English skills; German advantageous.

THE JOY OF WORKING WITH US

  • Impact from day one: Your work directly impacts the revenues of leading eCommerce brands in Europe.
  • Leadership role with scope for design: You will lead a well-rehearsed team and shape our platform during a crucial phase of our transformation.
  • Modern Tech Stack: Kubernetes, Harvester, GitOps, auto-scaling, and an exciting path towards the cloud – with room to build things new and right.
  • AI-first Mindset: We use AI not as a buzzword, but as a permanent component of our daily work.
  • Ownership & Growth: Clear responsibility, short decision-making paths, and the opportunity to actively shape your role.
  • Flexible Working: Hybrid work model with a focus on results.
  • Strong Team: Experienced engineers, open feedback culture, and an environment where reliability is taken seriously as an engineering discipline.
  • Attractive Benefits: Competitive salary, modern equipment, training budget, and regular team events.

Location

​Berlin, Munich, Pforzheim or Stockholm (hybrid)

Sign up to apply

Create a free account to apply for this job and get access to:

  • AI-powered CV optimization for this specific job
  • Save jobs and create custom alerts
  • See your CV match score for each job

Company information

Company
FactFinder
Location
Berlin
Germany
Posted
1 week ago

Find similar jobs

Explore more opportunities like this one.

Interested in this position?

Create your free account and tailor your CV to match this job.