Senior Site Reliability Engineer
Clera · remote
What this role requires
One requirement, read out of the advert rather than guessed from the job title:
Also mentioned, not required: Kubernetes, Terraform, Python, Bash, Go, Prometheus, Grafana, OpenTelemetry. Worth having, but their absence is not what gets a CV filtered out.
See how often each of these is required across open DevOps and platform roles in Europe.
Free, no card. Tells you which of them your CV evidences and which it only implies.
Job description
About the Role
We are a well-funded AI/ML company operating at the intersection of geospatial intelligence and infrastructure analytics. Our engineering team is distributed across Europe and North America, and we're looking for a Senior Site Reliability Engineer to take ownership of our cloud infrastructure and elevate our DevOps and reliability practices.
In this role, you'll evolve our Google Cloud Platform (GCP) infrastructure, mature our observability platform, drive incident management processes, and partner closely with Product & Engineering teams to ship reliable, high-quality software. You'll be a key voice in championing SLOs, error budgets, and DORA metrics across the organisation.
What You'll Do
Design, evolve, and scale our cloud infrastructure on GCP.
Build tooling and automation that promote team autonomy and reduce toil.
Read the full description (21 more sections)Show less
Advance our observability platform, improving mean time to recovery (MTTR) and system visibility.
Build transparency into infrastructure costs and drive cost optimisation initiatives.
Champion reliability best practices including SLOs/SLIs, error budgets, and post-incident reviews.
Lead on-call rotations and incident management, fostering a blameless culture.
Help engineering teams leverage GCP effectively and govern usage at scale.
What We're Looking For
Required:
3+ years of Site Reliability Engineering or production SRE experience.
Strong proficiency with Google Cloud Platform (GCP) , including cost optimisation and governance.
Nice to Have / Additional Skills:
Hands-on experience with Kubernetes for cluster and workload management.
Infrastructure as Code experience — Terraform , Deployment Manager, or similar.
Scripting and automation skills in Python , Bash , or Go .
Strong observability stack experience: Prometheus , Grafana , OpenTelemetry , logging, and tracing.
Proven ability to define and implement SLOs/SLIs and error budgets.
Experience with incident management, post-incident reviews, and on-call rotations.
Location
This is a fully remote role, open to candidates based in the EU, UK, or North America . The primary hub is in the Netherlands . Please note that visa sponsorship is not available for this position.
Compensation & Benefits
Compensation details were not provided for this role. Our team spans multiple countries and we offer competitive, location-adjusted packages. Further details will be discussed during the interview process.
Find Jobs in Germany on Arbeitnow
You will apply. Then you will hear nothing.
And no one will tell you what was wrong. See it before you send: your ATS score, every weak line, and the fix for each.
- 1
Drop in your CV
One PDF, thirty seconds. No card.
- 2
See what is wrong with it
Every weak passage, quoted from your own CV, with the line to replace it.
- 3
Apply where you fit
Every European role ranked against what your CV actually says.
- Free, no card
- Your CV file is deleted after parsing
- Or skip the upload — build your profile by hand
- Refreshed every 6 hours