- Home
- Jobs
- Dicetek LLC
- Site Reliability Engineer (SRE) - Azure focus
Site Reliability Engineer (SRE) - Azure focus
Job Description Roles & Responsibilities We are seeking an experienced Site Reliability Engineer (SRE) to ensure the reliability, availability, performance, and scalability of critical applications and infrastructure. The ideal candidate will have strong experience in monitoring, automation, cloud technologies, incident management, and DevOps practices. Key Responsibilities Monitor and maintain the availability, performance, and reliability of applications and infrastructure. Design and implement effective monitoring, logging, and alerting solutions. Manage and troubleshoot production incidents and perform root cause analysis. Automate operational and deployment processes to improve system reliability and efficiency. Work closely with development, infrastructure, and DevOps teams to improve application performance and resilience. Implement and maintain CI/CD pipelines and Infrastructure as Code (IaC). Support containerized environments and cloud-based infrastructure. Develop scripts and automation tools to reduce manual operational activities. Implement observability solutions, including monitoring, logging, and distributed tracing. Ensure proper documentation of operational procedures, incidents, and system configurations. Desired Candidate Profile Monitoring: Prometheus, Grafana, Zabbix Logging: ELK/Elastic Stack, Splunk APM: Dynatrace, AppDynamics, New Relic Cloud: AWS, Microsoft Azure, or GCP Containers: Docker, Kubernetes, OpenShift CI/CD: Jenkins, GitLab CI/CD, Azure DevOps Infrastructure as Code: Terraform, Ansible Version Control: Git, GitHub, GitLab Incident Management: ServiceNow, PagerDuty Distributed Tracing: OpenTelemetry, Jaeger Scripting: Bash, Python, PowerShell Databases: PostgreSQL, Oracle, SQL Server Web & APIs: IIS, Nginx, Apache, REST APIs Qualifications & Experience Bachelor's degree in Computer Science, Information Technology, or a related field. Proven experience as a Site Reliability Engineer, DevOps Engineer, or Production Support Engineer. Strong experience in cloud infrastructure, automation, monitoring, and incident management. Hands-on experience with Kubernetes and containerized environments. Strong troubleshooting and root cause analysis skills. Experience working in highly available, large-scale production environments. Excellent communication and collaboration skills. Skills cloud infrastructure monitoring incident managemen Employment Type Full-time Company Industry RecruitmentPlacement FirmExecutive Search Department / Functional Area Engineering Keywords AzureAutomationInfrastructure As CodeCloud Operations EngineerAzure SRELead SRETechnical Operations EngineerSystems Engineer Azure Get real-time job updates only on our App
Ready to apply?
You are viewing this role on JobSphere AI. Applications are completed on the original employer / source website.
Apply NowOpens the employer's site in a new tab
- CompanyDicetek LLC
- LocationDubai, UAE
- CategoryDevOps
- SourceNaukrigulf
- Listed1 month ago
Related DevOps jobs
Cashier -Billing Officer
Job Purpose :Responsible for patient billing, cash collection, insurance verification, claim submission, and ensuring accurate financial transactions within…
Managed Services Delivery Manager
Location: Riyadh, Saudi Arabia As a Service Delivery Manager (Team Manager) within the DES team, you will lead a team of 8+ engineers with diverse technical…
Refrigeration Technician
Install, maintain, and repair refrigeration and cooling systems used in bakery operations, including chillers, freezers, cold rooms, proofers, and…
Civil Foreman
Civil Foreman Knowledge & Qualification: IIT + Minium 5 years gulf experience in similar position in Oil & Gas civil construction project Role &…