- Home
- Jobs
- Client of Mark Williams
- Head of Site Reliability Engineering (SRE)
Head of Site Reliability Engineering (SRE)
Job Overview
We are partnering with an organisation that is looking to appoint a Head of Site Reliability Engineering (SRE) to establish and lead its SRE function. This is an excellent opportunity for a highly technical leader who enjoys building teams, defining best practices, and driving reliability, scalability, and operational excellence across modern cloud platforms. We are looking for someone who can remain hands on while developing a high performing SRE capability. Key Responsibilities - Build and lead the Site Reliability Engineering function, defining the vision, operating model, and best practices - Develop SRE principles, standards, and processes to improve platform reliability, availability, and performance - Lead the adoption of cloud native technologies while supporting both modern cloud applications and legacy workloads - Partner with engineering, infrastructure, and product teams to improve system resilience, scalability, and operational efficiency - Drive automation across monitoring, deployment, incident management, and platform operations - Define and measure SLIs, SLOs, and error budgets to improve service reliability - Lead incident management, root cause analysis, and continuous improvement initiatives - Mentor and develop engineers, creating a strong engineering culture focused on reliability and operational excellence Desired Candidate Profile 10+ years of experience in Site Reliability Engineering, Platform Engineering, DevOps, or Cloud Engineering Experience building or scaling an SRE function within a large enterprise or technology organisation Strong hands on experience with Google Cloud Platform (GCP) and cloud native technologies Proven experience supporting both cloud native applications and legacy enterprise workloads Strong understanding of Kubernetes, containers, infrastructure automation, CI/CD, observability, and platform engineering Experience with monitoring and observability tools such as Prometheus, Grafana, Datadog, or similar Strong scripting or programming experience using languages such as Python, Go, or Bash
Ready to apply?
You are viewing this role on JobSphere AI. Applications are completed on the original employer / source website.
Apply NowOpens the employer's site in a new tab
- CompanyClient of Mark Williams
- LocationDubai, UAE
- CategoryDevOps
- SourceNaukrigulf
- Listed2 months ago
Related DevOps jobs
Full Stack Developer
We seek a highly analytical and problem-solving-oriented Senior Software Engineer with a strong foundation in object-oriented programming (OOP) and scalable…
Operation Executive
Coordinate and optimize daily operational activities to enhance efficiency, ensuring that all tasks align with company goals and objectives. Implement process…
HR Executive
Manage the end-to-end recruitment process, including job postings, candidate screening, interview scheduling, and onboarding. Coordinate employee onboarding…
Mechanic (Automotive Workshop)
Diagnose and repair complex automotive issues, from engine performance to electrical systems, ensuring customer vehicles are returned to optimal working…