/#jobs
#Platform#About Us#Employers#AI Jobs
Get Started
Delft
6 days ago
Senior Site Reliability Engineer logo

Senior Site Reliability Engineer

TOPdesk

Delft
6 days ago
Apply

Senior Site Reliability Engineer

TOPdesk is seeking a Senior Site Reliability Engineer to own the reliability of its Azure SaaS estate, setting SLOs, eliminating toil, and building self-healing automation. The role involves leading incident response, consolidating observability, and embedding AI-native practices. Candidates need 5+ years of SRE/DevOps experience with Azure, Kubernetes, Terraform, and strong observability skills.

AI-enabledOn-siteFull-timeSeniorAzureKubernetes

Senior Site Reliability Engineer

TOPdesk is seeking a Senior Site Reliability Engineer to own the reliability of its Azure SaaS estate, setting SLOs, eliminating toil, and building self-healing automation. The role involves leading incident response, consolidating observability, and embedding AI-native practices. Candidates need 5+ years of SRE/DevOps experience with Azure, Kubernetes, Terraform, and strong observability skills.

Apply
AI-enabledOn-siteFull-timeSeniorAzure

Salary

Not specified

Work Location

Delft, South Holland, Netherlands, NL

Work Model

On-site

Experience Required

5 years

Employment Type

Full-time

Experience Level

Senior

Core Qualifications

Technical (Must-have)
AzureKubernetesHelmTerraformCI/CDGrafanaPrometheusVictoriaMetricsPythonLinuxPuppetAnsibleSLOsError budgetsIncident response
Soft Skills
CommunicationCollaborationProblem-solvingLeadershipMentoringTransparencyOpen feedbackWork-life balance

Preferred Qualifications

Technical (Nice-to-have)
Progressive deliveryAzure Cloud Adoption FrameworkClaude CodeEU data residency

Key Responsibilities

  • •Define and own SLOs and error budgets across the Azure estate
  • •Eliminate toil and implement self-healing automation
  • •Standardise observability (metrics, alerting, tracing) across datacenters
  • •Lead incident response and blameless postmortems
  • •Harden CI/CD and progressive delivery (canaries, safe rollouts, automated rollback)
  • •Model capacity and load-test critical paths
  • •Bring AI-native reliability (agents, bounded automation) into detection and remediation
  • •Maintain runbooks linked to automation candidates
  • •Own capacity and cost forecasting across the multi-cloud estate
  • •Collaborate proactively with product teams during design phases
Site Reliability EngineeringAzureKubernetesTerraformObservabilitySLOsAutomationAI-nativeSaaSSenior
/#jobs

Your gateway to a successful career. Show your growth. Be ready for your next step. Capture and seize the best opportunities.

  • Data
  • FAQ
  • Articles
  • AI Jobs
  • Platform
  • Employers
  • About Us
  • Legal
© 2026/#jobsAll rights reserved.

For queries/support, email jobs.support@slashhash.ai