Jobs / FACT-Finder Holding GmbH

Senior Site Reliability Engineer / SRE – Kubernetes & Hybrid Cloud (m/f/d)

FACT-Finder Holding GmbH · Berlin, BE, Deutschland
Berlin, BE, DeutschlandHybrid
Remuneration
Not specified
Location
Berlin, BE, Deutschland
Visa sponsorship
Not specified

Job summary

The Senior Site Reliability Engineer role is centered on developing and maintaining a private cloud platform utilizing Kubernetes and Harvester, while working closely with development teams. Responsibilities include defining SLOs, leading incident responses, automating processes with GitOps, and enhancing observability across various stacks. The position demands hands-on experience with Kubernetes in production and a strong focus on automation.

Qualifications

  • Kubernetes in production experience
  • Lived SRE practice: SLOs, error budgets, incident management
  • Hands-on experience with GitOps or similar automation
  • Solid observability skills: metrics, logs, traces, alerting
  • Strong automation instinct

Responsibilities

  • Define and own SLOs, SLIs and error budgets
  • Lead incident response end-to-end
  • Eliminate toil through automation and GitOps
  • Help build custom Kubernetes operator
  • Plan capacity, performance and cost

Skills

Argo CDCephFluxGrafanaK3sKubernetesOpenStackPrometheusvSphere

Relocation

No