Job Description

We are partnering with an organisation that is looking to appoint a Head of Site Reliability Engineering (SRE) to establish and lead its SRE function.


This is an excellent opportunity for a highly technical leader who enjoys building teams, defining best practices, and driving reliability, scalability, and operational excellence across modern cloud platforms. We are looking for someone who can remain hands on while developing a high performing SRE capability.


Key Responsibilities

- Build and lead the Site Reliability Engineering function, defining the vision, operating model, and best practices

- Develop SRE principles, standards, and processes to improve platform reliability, availability, and performance

- Lead the adoption of cloud native technologies while supporting both modern cloud applications and legacy workloads

- Partner with engineering, infrastructure, and product teams to improve system resilience, scalability, and operational efficiency

- Drive automation across monitoring, deployment, incident management, and platform operations

- Define and measure SLIs, SLOs, and error budgets to improve service reliability

- Lead incident management, root cause analysis, and continuous improvement initiatives

- Mentor and develop engineers, creating a strong engineering culture focused on reliability and operational excellence


Requirements

- 10+ years of experience in Site Reliability Engineering, Platform Engineering, DevOps, or Cloud Engineering

- Experience building or scaling an SRE function within a large enterprise or technology organisation

- Strong hands-on experience with Google Cloud Platform (GCP) and cloud native technologies

- Proven experience supporting both cloud native applications and legacy enterprise workloads

- Strong understanding of Kubernetes, containers, infrastructure automation, CI/CD, observability, and platform engineering

- Experience with monitoring and observability tools such as Prometheus, Grafana, Datadog, or similar

- Strong scripting or programming experience using languages such as Python, Go, or Bash


Job Details

Role Level: Mid-Level Work Type: Full-Time
Country: United Arab Emirates City: Abu Dhabi
Company Website: http://www.markwilliams.ae Job Function: DevOps & QA
Company Industry/
Sector:
Holding Companies

What We Offer


About the Company

Searching, interviewing and hiring are all part of the professional life. The TALENTMATE Portal idea is to fill and help professionals doing one of them by bringing together the requisites under One Roof. Whether you're hunting for your Next Job Opportunity or Looking for Potential Employers, we're here to lend you a Helping Hand.

Report

Disclaimer: talentmate.com is only a platform to bring jobseekers & employers together. Applicants are advised to research the bonafides of the prospective employer independently. We do NOT endorse any requests for money payments and strictly advice against sharing personal or bank related information. We also recommend you visit Security Advice for more information. If you suspect any fraud or malpractice, email us at abuse@talentmate.com.


Recent Jobs
View More Jobs
Talentmate Instagram Talentmate Facebook Talentmate YouTube Talentmate LinkedIn