Netherlands
1 week ago
Engineering Manager, SRE logo

Engineering Manager, SRE

Jobgether

Engineering Manager, SRE

Engineering Manager, SRE needed to lead a Site Reliability Engineering team for a globally distributed technology platform. Requires hands-on technical depth in Kubernetes, AWS, PostgreSQL, CI/CD, observability, and infrastructure as code, plus proven people leadership experience. Fully remote, asynchronous environment with a salary range of USD $75,450–$169,700.

AI-enabledHybridFull-timeSeniorKubernetesAWS

Salary

Not specified

Work Location

Netherlands, NL

Work Model

Hybrid; environment is fully remote and asynchronous

Employment Type

Full-time

Experience Level

Associate

Core Qualifications

Technical (Must-have)
KubernetesAWSPostgreSQLCI/CDObservabilityInfrastructure as CodeTerraformGitLab CIGitHub ActionsJenkinsDockerShell scriptingDNSTLSAI infrastructure
Soft Skills
PrioritizationWritten communicationDocumentationRelationship-buildingCollaborationConflict resolutionStakeholder managementCoachingJudgmentAccountabilityAdaptabilityCuriosity

Preferred Qualifications

Technical (Nice-to-have)
ElixirJavaClojureNode.jsPythonOpenTelemetryDistributed tracingHoneycombAuroraLinux systems administrationSecurityFinOpsCloud cost management

Key Responsibilities

  • Lead and develop a Site Reliability Engineering team, owning the full career lifecycle of direct reports including onboarding, feedback, performance management, progression, coaching, and hiring.
  • Establish a clear team direction and priorities aligned with broader company goals, balancing operational commitments with project delivery and protecting the team’s focus.
  • Serve as the team's spokesperson across engineering and with senior leadership, communicating priorities, progress, risks, and technical challenges clearly.
  • Own SRE delivery goals, deciding what the team commits to, how work is prioritized, and how operational responsibilities are managed.
  • Design and maintain effective support rotations and on-call processes while strengthening incident response practices.
  • Provide technical leadership across Kubernetes, AWS, PostgreSQL, DNS and TLS, CI infrastructure, and the broader infrastructure platform.
  • Guide the development of reliability practices including SLOs, error budgets, observability, incident response, and post-incident improvements.
  • Partner closely with Security on infrastructure threats, patching, controls, audits, and compliance obligations.
  • Manage relationships with infrastructure and platform vendors, including renewals and commercial discussions with support from senior leadership.
  • Remain hands-on enough to review technical work, challenge architectural decisions, participate credibly in incidents, and identify emerging reliability issues before they escalate.
  • Build strong relationships across engineering and encourage teams to bring operational and reliability challenges forward early.
  • Continuously improve team health, collaboration, conflict resolution, and retrospective practices.
Engineering ManagerSRESite Reliability EngineeringKubernetesAWSRemoteFull-timeInternet Marketplace PlatformsInfrastructureLeadership