About the role

Senior Site Reliability Engineer (SRE & AI Platform Operations)

Introduction

How do you keep an international e-commerce platform reliable while its core technology is being rebuilt? As Senior Site Reliability Engineer at HelloPrint, you will take ownership of the infrastructure, monitoring and deployment practices that make this possible. Your work will support essential customer journeys, including checkout and payments, as well as catalogue pipelines and routing.

The challenge also extends into AI: you will manage the operational side of LLM integrations, semantic pipelines and AI runtime costs. Working hands-on with Google Cloud, you will give engineering teams the tools and safeguards to release quickly and confidently. You will have the freedom to make meaningful technical improvements and the responsibility to see them through.

Learn more about the company

About the company

Culture and values
Office
Colleagues
Language

About the role

Benefits

  • Gross annual salary of €65,000–€95,000, including 8% holiday allowance.
  • Full technical ownership of reliability and AI platform operations.
  • Autonomy to implement meaningful technical improvements.
  • Development opportunities within your role and across HelloPrint.
  • An international team of more than 100 colleagues from over 20 nationalities.
  • Direct impact on the development of an AI-driven organisation.
  • 24/7 access to the HelloFit gym.
  • Urban Sports Club discount.
  • Company events.
  • Company-sponsored healthy meals every day.

Required Skills

  • Proven experience operating and scaling high-traffic distributed production systems.
  • Strong troubleshooting skills across Linux, containerised runtimes, relational and NoSQL databases, Redis queues and cloud networks.
  • Extensive hands-on experience with Google Cloud Platform, Cloud Run and Terraform.
  • Experience with SLOs, SLIs, error budgets, monitoring and distributed tracing.
  • Demonstrated experience with cloud cost governance, FinOps and infrastructure optimisation.
  • Practical experience with Sentry and Google Cloud Monitoring.
  • Strong programming or scripting skills in Python or TypeScript/JavaScript.
  • Practical familiarity with modern PHP in a Laravel environment.
  • Affinity with operationalising LLM integrations, embeddings, vector workflows, external APIs or background task orchestration.
  • Availability to work full-time from the Rotterdam office, five days per week.

Key Responsibilities

  • Define and enforce SLOs, SLIs and error-budget policies for critical services and customer journeys.
  • Improve distributed tracing, telemetry and automated diagnostics with Sentry and Google Cloud Monitoring.
  • Strengthen CI/CD through canary releases, automated rollbacks and health gates in GitHub Actions.
  • Design, monitor and scale runtime infrastructure for AI agents, semantic pipelines and background automation.
  • Manage cloud, model and token costs, alongside latency, rate limits and provider availability.
  • Lead incident response and blameless post-mortems, turning findings into structural improvements.
  • Improve capacity planning, dependency isolation, load testing and disaster-recovery validation.
  • Manage infrastructure as code using Terraform, Cloud Run, IAM and Secret Manager.
  • Build internal tools, runbooks and self-service deployment solutions that reduce repetitive operational work.

Why Work at HelloPrint

Take ownership of platform reliability

Shape how HelloPrint approaches observability, deployment safety, incident management and cloud costs. You will have responsibility across the complete reliability domain.

Bring cloud engineering and AI together

Combine established SRE practices with the operational challenges of LLMs, embeddings and AI workloads. Your decisions will influence performance, availability and running costs.

Help shape the next platform

HelloPrint is rebuilding its frontend, pricing, content and product engines. You will create the technical foundations that help these systems run reliably as they evolve.

Your Growth Path at HelloPrint

Staff Site Reliability Engineer

Platform Engineering Lead

Head of Platform Engineering

Application process

First Interview

Get to know each other and learn more about our company.

Second Interview

A more in-depth interview with the hiring manager to explore how you can make an impact within the team.

Third Interview

A follow-up with senior team members or decision-makers.

Thom Bakker

Is here to help you

Apply Now

Office locations

Address
Schiedamse Vest 89, 3012 BG , Netherlands

Details
Job Type
Full-time
Location
Salary
5414 - €7917

“Hire top SaaS talent. Faster.”

We cover the full SaaS spectrum: sales, marketing, customer success, operations, tech, and leadership positions. From entry-level to executive.

Start now