Jobs / Helloprint B.V.
Senior Site Reliability Engineer (SRE & AI Platform Operations)
Helloprint B.V. · Rotterdam, ZH, Netherlands
Rotterdam, ZH, NetherlandsRemote
Remuneration
Not specified
Location
Rotterdam, ZH, Netherlands
Visa sponsorship
Not specified
Job summary
The Senior Site Reliability Engineer (SRE) role is responsible for ensuring production reliability, distributed observability, deployment safety, cost optimization, and AI runtime infrastructure across high-velocity microservices and cloud workloads. Key responsibilities include defining Service Level Objectives (SLOs), expanding telemetry with Google Cloud tools, and leading incident management efforts to enhance system resilience and performance.
Qualifications
- Experience operating high-traffic distributed production systems
- Strong troubleshooting skills in Linux and containerized environments
- Hands-on experience with Google Cloud and Terraform
- Ability to optimize cloud and AI runtime costs
- Experience configuring Sentry and Google Cloud Monitoring
- Proficiency in Python or TypeScript/JavaScript
- Familiarity with LLM integrations and background task orchestration
- Pragmatic builder with technical agency
Responsibilities
- Define, track, and enforce Service Level Objectives (SLOs) and error-budget policies
- Expand telemetry and monitoring with Google Cloud tools
- Evolve CI/CD pipelines with automated rollbacks and health gates
- Architect and scale AI runtime infrastructure
- Drive continuous FinOps practices across cloud workloads
- Lead incident response and post-mortems
- Drive capacity forecasting and disaster recovery validations
- Own infrastructure workflows using Terraform and Google Cloud Run
- Build internal tooling and self-service deployment primitives
Skills
Cloud RunGCPGitHubGitHub ActionsIAMJavaScriptLinuxPHPPythonRedisSentryCloud MonitoringTerraformTypeScript
Relocation
No