1. Home
  2. Jobs
  3. Client of Salt
  4. Senior DevOps / Site Reliability Engineer
FullStack

Senior DevOps / Site Reliability Engineer

Client of Salt
Abu Dhabi, UAE Listed 1h ago via Naukrigulf
python docker kubernetes aws azure gcp terraform github actions ci/cd devops linux llm security

Job Description Roles & Responsibilities We are seeking a hands-on Senior DevOps / Site Reliability Engineer to build and operate the delivery and runtime foundations for a portfolio of modern enterprise applications, workflow platforms and AI-enabled products. This is not primarily an infrastructure-administration role. You will work directly with software engineers to create reliable, secure and automated paths from code to production. You will own deployment automation, runtime reliability, observability, infrastructure-as-code and operational readiness across applications that integrate with critical enterprise systems. What You Will Own Platform & Infrastructure Build and maintain cloud infrastructure using Infrastructure as Code. Design secure, repeatable environments across development, test, staging and production. Manage containerized workloads and Kubernetes-based deployments where appropriate. Define standard application deployment patterns for backend, frontend and AI services. Implement secure secrets and configuration management. Support network, identity and connectivity requirements for enterprise integrations. CI/CD & Developer Productivity Build automated CI/CD pipelines. Standardize build, test, security scanning and deployment processes. Automate environment provisioning and configuration. Reduce manual deployment steps and production configuration drift. Work closely with engineering teams to improve release frequency and reliability. Reliability & Observability Establish logging, metrics, tracing and alerting. Define service-level indicators and operational thresholds. Build dashboards for system health and application performance. Implement incident-response and production-support practices. Design for graceful degradation, retries, failover and recovery. Lead root-cause analysis of production incidents. Security & Operational Controls Implement least-privilege access and secure deployment patterns. Support auditability of infrastructure and production changes. Integrate security checks into delivery pipelines. Work with security and infrastructure teams to meet enterprise control requirements. Resilience Support business continuity and disaster-recovery design. Define backup, restore and recovery procedures. Test operational recovery rather than relying solely on documented plans. Desired Candidate Profile 6+ years in DevOps, SRE, platform engineering or cloud infrastructure. Strong production experience with Azure, AWS or GCP; Azure strongly preferred. Docker and Kubernetes. Infrastructure as Code using Terraform, Bicep, Pulumi or equivalent. CI/CD using Azure DevOps, GitHub Actions, GitLab CI or similar. Strong Linux and networking fundamentals. Observability tooling and distributed-system troubleshooting. Secure secrets, identity and access-management patterns. Production incident-management experience. Scripting/programming capability in Python, Go, Bash or equivalent. Strong Advantage Azure Kubernetes Service. Azure Service Bus, API Management, Key Vault and related Azure services. Enterprise integration platforms. SAP-connected environments. AI/LLM application deployment. Regulated or government environments. High-availability and disaster-recovery architecture. Company Industry RecruitmentPlacement FirmExecutive Search Department / Functional Area Engineering Keywords Senior DevOps / Site Reliability Engineer Get real-time job updates only on our App

Ready to apply?

You are viewing this role on JobSphere AI. Applications are completed on the original employer / source website.

Apply Now

Opens the employer's site in a new tab

  • CompanyClient of Salt
  • LocationAbu Dhabi, UAE
  • CategoryFullStack
  • SourceNaukrigulf
  • Listed1h ago

Related FullStack jobs

More FullStack