Senior Site Reliability Engineer (SRE & AI Platform Operations)
Senior Site Reliability Engineer (SRE & AI Platform Operations)
Introduction
The challenge also extends into AI: you will manage the operational side of LLM integrations, semantic pipelines and AI runtime costs. Working hands-on with Google Cloud, you will give engineering teams the tools and safeguards to release quickly and confidently. You will have the freedom to make meaningful technical improvements and the responsibility to see them through.
Learn more about the company
About the company
About the role
Benefits
- Gross annual salary of €65,000–€95,000, including 8% holiday allowance.
- Full technical ownership of reliability and AI platform operations.
- Autonomy to implement meaningful technical improvements.
- Development opportunities within your role and across HelloPrint.
- An international team of more than 100 colleagues from over 20 nationalities.
- Direct impact on the development of an AI-driven organisation.
- 24/7 access to the HelloFit gym.
- Urban Sports Club discount.
- Company events.
- Company-sponsored healthy meals every day.
Required Skills
- Proven experience operating and scaling high-traffic distributed production systems.
- Strong troubleshooting skills across Linux, containerised runtimes, relational and NoSQL databases, Redis queues and cloud networks.
- Extensive hands-on experience with Google Cloud Platform, Cloud Run and Terraform.
- Experience with SLOs, SLIs, error budgets, monitoring and distributed tracing.
- Demonstrated experience with cloud cost governance, FinOps and infrastructure optimisation.
- Practical experience with Sentry and Google Cloud Monitoring.
- Strong programming or scripting skills in Python or TypeScript/JavaScript.
- Practical familiarity with modern PHP in a Laravel environment.
- Affinity with operationalising LLM integrations, embeddings, vector workflows, external APIs or background task orchestration.
- Availability to work full-time from the Rotterdam office, five days per week.
Key Responsibilities
- Define and enforce SLOs, SLIs and error-budget policies for critical services and customer journeys.
- Improve distributed tracing, telemetry and automated diagnostics with Sentry and Google Cloud Monitoring.
- Strengthen CI/CD through canary releases, automated rollbacks and health gates in GitHub Actions.
- Design, monitor and scale runtime infrastructure for AI agents, semantic pipelines and background automation.
- Manage cloud, model and token costs, alongside latency, rate limits and provider availability.
- Lead incident response and blameless post-mortems, turning findings into structural improvements.
- Improve capacity planning, dependency isolation, load testing and disaster-recovery validation.
- Manage infrastructure as code using Terraform, Cloud Run, IAM and Secret Manager.
- Build internal tools, runbooks and self-service deployment solutions that reduce repetitive operational work.
Take ownership of platform reliability
Shape how HelloPrint approaches observability, deployment safety, incident management and cloud costs. You will have responsibility across the complete reliability domain.
Bring cloud engineering and AI together
Combine established SRE practices with the operational challenges of LLMs, embeddings and AI workloads. Your decisions will influence performance, availability and running costs.
Help shape the next platform
HelloPrint is rebuilding its frontend, pricing, content and product engines. You will create the technical foundations that help these systems run reliably as they evolve.
Staff Site Reliability Engineer
Platform Engineering Lead
Head of Platform Engineering
Application process
Get to know each other and learn more about our company.
A more in-depth interview with the hiring manager to explore how you can make an impact within the team.
A follow-up with senior team members or decision-makers.
Office locations





“Hire top SaaS talent. Faster.”
We cover the full SaaS spectrum: sales, marketing, customer success, operations, tech, and leadership positions. From entry-level to executive.
