Jobs / FACT-Finder Holding GmbH
Senior Site Reliability Engineer / SRE – Kubernetes & Hybrid Cloud (m/f/d)
FACT-Finder Holding GmbH · Berlin, BE, Deutschland
Berlin, BE, DeutschlandHybrid
Remuneration
Not specified
Location
Berlin, BE, Deutschland
Visa sponsorship
Not specified
Job summary
The Senior Site Reliability Engineer role is centered on developing and maintaining a private cloud platform utilizing Kubernetes and Harvester, while working closely with development teams. Responsibilities include defining SLOs, leading incident responses, automating processes with GitOps, and enhancing observability across various stacks. The position demands hands-on experience with Kubernetes in production and a strong focus on automation.
Qualifications
- Kubernetes in production experience
- Lived SRE practice: SLOs, error budgets, incident management
- Hands-on experience with GitOps or similar automation
- Solid observability skills: metrics, logs, traces, alerting
- Strong automation instinct
Responsibilities
- Define and own SLOs, SLIs and error budgets
- Lead incident response end-to-end
- Eliminate toil through automation and GitOps
- Help build custom Kubernetes operator
- Plan capacity, performance and cost
Skills
Argo CDCephFluxGrafanaK3sKubernetesOpenStackPrometheusvSphere
Relocation
No