Functiebeschrijving
logo NN Group
Verwijder uit favorieten Voeg toe aan favorieten Voeg toe aan favorieten logo NN Group
Taal NL EN
Bedrijf logo NN Group
Verwijder uit favorieten Voeg toe aan favorieten Voeg toe aan favorieten
The Engineering Experience & Platforms domain aims to make life easier for engineers across NN. To strengthen reliability and improve observability across key platforms, NN is looking for an interim Site Reliability Engineer to design, implement and embed a scalable observability foundation.
What you are going to do
The assignment focuses on setting up an OpenTelemetry collector layer for sources including Azure Databricks, AWS Kubernetes, Azure Kubernetes, AWS and Azure. You will define and implement SLIs, SLOs and SLAs for the Portable Stack domain and AI domain, and enable teams to use the new observability capabilities effectively.
What we offer you
Our people are the driving force behind our organisation. We value the knowledge and expertise you bring. We believe that your temporary commitment can take our organisation to a higher level. We offer you: - Competitive hourly rate depending on your knowledge and experience - Project starts 1st of August 2026 and has a duration of 6 to 9 months - Hybrid way of working, partly from home and partly from the office - International working environment with loads of knowledge sharing
Who you are
You are an experienced Site Reliability Engineer with strong production experience in Kubernetes and containerized workloads. You have hands-on cloud engineering experience in Azure and/or AWS, including infrastructure as code and GitOps-based deployments. You bring deep observability expertise across metrics, logs and traces, using tools such as Grafana, Prometheus, Loki, Tempo or similar. You are comfortable defining and managing SLOs, SLIs and error budgets, and you have a structured approach to incident management, root cause analysis and reliability improvements. Automation ...