Site Reliability Engineer

Hoofddorp, North Holland, Netherlands

6000 - 8500

32 - 40 hours

Tech & Digital, EN

Apply Now

Keep our global cloud platform reliable, scalable, and resilient.

Do you enjoy solving complex production challenges before they turn into incidents? Are you the kind of engineer who automates repetitive tasks, improves reliability through engineering, and believes that every outage is an opportunity to build a better system?

We're looking for a Site Reliability Engineer to join the Global Platform Team at HeadFirst x Impellam Group. In this role, you'll help build and operate the cloud platform that powers our global Workforce-as-a-Service (WaaS) ecosystem, working alongside Cloud Engineers, Platform Engineers, Data Engineers, and AI specialists to improve platform resilience, reduce operational overhead, and ensure that our Azure-based engineering environment remains reliable, scalable, and resilient.

Your impact

As a Site Reliability Engineer, your focus is simple: keep our platforms healthy, reliable, and easy to operate. You'll help build the engineering foundations behind our Headless Data Architecture (HDA), which runs on Azure and Databricks, as well as the Custom Apps Infrastructure (CA) that powers integrations, internal applications, and operational workflows across our international organization.

Instead of spending your days reacting to incidents, you'll focus on preventing them through automation, observability, and reliability engineering. You'll reduce operational burden, improve platform resilience, and build systems that scale, recover automatically whenever possible, and provide engineering teams across Cloud, Data, and AI with the visibility they need to run production workloads with confidence.

What you will do

  • Improve the reliability, availability, and performance of our Azure platform and production environments;

  • Build and improve monitoring, logging, and alerting using Grafana, OpenTelemetry, Azure Monitor, and Log Analytics;

  • Automate operational tasks and eliminate repetitive manual work using Infrastructure as Code and scripting;

  • Design self-healing capabilities and automated remediation to reduce incidents and improve recovery times;

  • Investigate production incidents, perform root cause analyses, and implement long-term improvements;

  • Define, measure, and improve Service Level Indicators (SLIs) and Service Level Objectives (SLOs);

  • Optimize platform performance, scalability, and operational efficiency;

  • Work closely with Cloud Engineers to improve platform architecture, resilience, and security;

  • Support Data and AI teams by improving the reliability of Azure Databricks environments;

  • Drive engineering best practices in observability, automation, and operational excellence;

  • Continuously look for opportunities to reduce operational complexity and improve developer productivity.

About the role

As part of the Global Platform Team, you'll work alongside engineers in Cloud, Data, and AI to improve the reliability of our Azure-based platform. Using technologies such as Kubernetes, Terraform, Databricks, GitHub Actions, Grafana, and OpenTelemetry, you'll help ensure that our global Workforce-as-a-Service ecosystem remains reliable, scalable, and resilient.

About HeadFirst Group x Impellam Group

HeadFirst Group x Impellam Group is one of Europe's leading providers of workforce and talent solutions. Operating across multiple countries, we are transforming into a cloud-native, AI-powered organization that connects people, technology, and data through a modern digital platform. The Global Platform Team is at the heart of that transformation, enabling engineering teams across Cloud, Data, AI, and Software Engineering to build and operate scalable solutions for the future.

Interested?

If you're passionate about building reliable, scalable cloud platforms and enjoy solving complex engineering challenges, we'd love to hear from you. Apply today and let's find out how you can make an impact as part of our Global Platform Team.

Here's what we offer you

Salary that matches your experience
You will receive a salary range that matches your knowledge and experience. We reward you fairly for your commitment and development.
Vacation pay and days off
You will receive 8.33% vacation pay and are entitled to 27 vacation days per year based on full-time employment. This gives you plenty of time to recharge your batteries.
Hybrid working
We support hybrid working and apply a 60/40 split, whereby full-time employees spend three days at the office and two days working from home.
Premium-freepension
We pay your entire pension premium. This means you accrue pension without having to contribute to it yourself. As a result, you are left with a higher net salary.
Performancebonus
Depending on your performance and that of HeadFirst Group, you may receive an annual bonus of one or two months' salary.
Monthly extras
You will receive a monthly allowance for internet, a vitality budget, a contribution towards lunch, and support in setting up your home office.

What we expect from you

You're passionate about building reliable systems and solving operational challenges through engineering rather than manual intervention. You enjoy understanding how distributed systems behave, thrive in cloud-native environments, and are always looking for ways to improve automation, resilience, and observability. You stay calm under pressure, take ownership of problems, and enjoy collaborating with others to continuously improve the platform's reliability.

Ideally, you should also have:

  • 4+ years of experience as a Site Reliability Engineer, Platform Engineer, DevOps Engineer, or Cloud Engineer;

  • Strong hands-on experience with Microsoft Azure;

  • Experience with Infrastructure as Code using Terraform;

  • Experience building and maintaining CI/CD pipelines using GitHub Actions or Azure DevOps;

  • Experience with observability tools such as Grafana, OpenTelemetry, Azure Monitor, or Log Analytics;

  • Strong scripting skills using Python, Bash, or similar languages;

  • Experience supporting distributed cloud platforms in production;

  • Experience with incident management, root cause analysis, and post-incident improvements;

  • Familiarity with GitOps principles and modern deployment practices;

  • Experience with Azure Databricks is a strong advantage;

  • Experience with SnapLogic or similar integration platforms is a plus.

We know there's no such thing as the perfect candidate. If this role excites you but you don't meet every single requirement, we'd still love to hear from you. We're just as interested in your potential, mindset, and ambition as we are in your experience.

Any questions?

Feel free to ask, I'm happy to help!

Britt van Heffen

Corporate Recruiter

Application Process

Submit your application

We will contact youwithin 48 hours.

First interview

Assessment

Second interview

Job offer

Welcome to HeadFirst Group!

Your benefits at HeadFirst Group

27 vacation days

AND the possibility of adding or selling days. We work hard, but don't forget to relax.

Moments of celebration

With more than 450 colleagues, we always have something to to celebrate. As one team, we celebrate birthdays, anniversaries and other successes!

OpenUp & HeadFirst Group Academy

Work on your mental health with access to the platform OpenUp. We learn every day, which is why we offer you free trainings and courses to keep developing yourself.

Bonus 

Achieve your goals and see this reflected in the form of a bonus? That's possible with us! Hard work is rewarded.

Hybrid working

Work at one of our 4 locations, at home, or at any wherever you feel comfortable. Of course you will receive a mobility and home working allowance from us.

Our headquarters

Take advantage of the free sports facilities, fine workstations and enjoy a delicious lunch buffet. A formidable competitor of home!

👑 We are a Great Place to Work!

Ready to shake up the next world of work?

Meet a colleague

Erix Santman

Product Manager

Diederick van Tellingen

UX Designer

Mattijs Wassenburg

Managing Director

Christine Koekkoek

Product Owner

Place your bid on Striive

https://login.striive.com/

For this assignment you need to place a bid on Striive. Striive is the largest assignment platform in the Benelux where more than 20,000 assignments are published annually.