1. Home
  2. Jobs
  3. SITA
  4. Site Reliability Engineer/Expert/Specialist (DevOps)
Cybersecurity

Site Reliability Engineer/Expert/Specialist (DevOps)

SITA
Amman, Jordan Listed 1 month ago via Naukrigulf
kubernetes ci/cd devops security

Job Overview

Company Industry IT - Software Services
Department / Functional Area IT Software
Keywords Site Reliability Engineer/Expert/Specialist (DevOps)

& TEAM The Site Reliability Engineer is responsible for the proactive support of products to ensure high product performance, with a continuous focus on improvement. The role involves identifying and resolving the root causes of operational incidents, implementing solutions to enhance stability, and preventing recurrence. The Site Reliability Engineer manages the creation and maintenance of the event catalogue to trigger events and develops both manual remediation approaches and automated workflows to address alerts. Additionally, they oversee the deployment of IT services and solutions, ensuring seamless integration with minimal disruption. WHAT YOU LL DO Design, build, and maintain support systems to ensure high availability, scalability, and performance of critical infrastructure. Lead incident response and root cause analysis for system failures, including problem investigations and coordination with relevant teams. Implement and manage automation for system provisioning, deployment, self-healing, and performance monitoring to increase operational efficiency. Establish and monitor SLIs/SLOs, proactively identify performance issues, and drive continuous improvements in service reliability. Collaborate with development and operations teams to embed reliability best practices and evolve toward zero-downtime architecture. Manage and optimize an event catalog, including event definitions, thresholds, remediation actions, and relevance across products. Develop event response protocols, provide training, and ensure efficient handling of incidents across teams. Drive post-incident reviews and feedback loops to enhance event definitions and service reliability. Oversee quality and readiness of deployments, ensuring clear processes, assigned responsibilities, and minimal disruption. Maintain deployment schedules and conduct risk assessments to ensure operational stability and deployment readiness. Coordinate and execute deployment plans, manage resources, and incorporate feedback for continuous process improvement. Manage CI/CD pipelines and infrastructure as code, ensuring seamless integration between development and operations. Support and evolve DevOps practices, automating operational tasks and maintaining tools to drive ongoing efficiency.

Desired Candidate Profile

Who You Are: Bachelor's degree in computer science, Information Technology, Engineering, or a related field. Several years of experience in IT operations, service management, or infrastructure management, including roles such as Site Reliability Engineer, Problem Manager, or DevOps Manager. Proven experience in managing high-availability systems and ensuring operational reliability. Extensive experience in root cause analysis (RCA), incident management, and developing permanent solutions for recurring service disruptions. Hands-on experience with CI/CD pipelines, automation, system performance monitoring, and the implementation of infrastructure as code. Strong background in collaborating with cross-functional teams (development, operations, engineering, etc.) to improve operational processes and service delivery. Experience in managing deployments, risk assessments, and optimizing event and problem management processes. Familiarity with cloud technologies, containerization, and scalable architecture, including experience with zero-downtime deployment strategies. CompTIA Security+ or Certified Kubernetes Administrator (CKA). Functional

Skills

Collaboration Stakeholder Management Service Design Communication Problem Solving Incident Management Change Management Innovation

Technical Skills

Cloud Infrastructure Automation & AI Operations Monitoring & Diagnostics Deployment Programming & Scripting Languages

Ready to apply?

You are viewing this role on JobSphere AI. Applications are completed on the original employer / source website.

Apply Now

Opens the employer's site in a new tab

  • CompanySITA
  • LocationAmman, Jordan
  • CategoryCybersecurity
  • SourceNaukrigulf
  • Listed1 month ago

Related Cybersecurity jobs

More Cybersecurity