Job matched to your search
Site Reliability Engineer (SRE)
Dicetek LLC · Dubai
Dubai · On-siteFull-TimePosted Sep 5, 2026
Free · Join 5,000+ job seekers using Qarera
How well do you match this role?
Tap the skills you already have — then see your real match score, what’s missing, and your resume fixed for this job.
↑ tap the skills you have
Loading sign-in…
Free · no credit card · 30 secondsJob description
Job Purpose We are seeking an experienced Site Reliability Engineer (SRE) to ensure the reliability, availability, performance, and scalability of critical applications and infrastructure. The ideal candidate will have strong experience in monitoring, automation, cloud technologies, incident management, and DevOps practices.
Key Responsibilities
- Monitor and maintain the availability, performance, and reliability of applications and infrastructure.
- Design and implement effective monitoring, logging, and alerting solutions.
- Manage and troubleshoot production incidents and perform root cause analysis.
- Automate operational and deployment processes to improve system reliability and efficiency.
- Work closely with development, infrastructure, and DevOps teams to improve application performance and resilience.
- Implement and maintain CI/CD pipelines and Infrastructure as Code (IaC).
- Support containerized environments and cloud-based infrastructure.
- Develop scripts and automation tools to reduce manual operational activities.
- Implement observability solutions, including monitoring, logging, and distributed tracing.
- Ensure proper documentation of operational procedures, incidents, and system configurations.
Required Technical Skills
- Monitoring: Prometheus, Grafana, Zabbix
- Logging: ELK/Elastic Stack, Splunk
- APM: Dynatrace, AppDynamics, New Relic
- Cloud: AWS, Microsoft Azure, or GCP
- Containers: Docker, Kubernetes, OpenShift
- CI/CD: Jenkins, GitLab CI/CD, Azure DevOps
- Infrastructure as Code: Terraform, Ansible
- Version Control: Git, GitHub, GitLab
- Incident Management: ServiceNow, PagerDuty
- Distributed Tracing: OpenTelemetry, Jaeger
- Scripting: Bash, Python, PowerShell
- Databases: PostgreSQL, Oracle, SQL Server
- Web & APIs: IIS, Nginx, Apache, REST APIs
Qualifications & Experience
- Bachelor's degree in Computer Science, Information Technology, or a related field.
- Proven experience as a Site Reliability Engineer, DevOps Engineer, or Production Support Engineer.
- Strong experience in cloud infrastructure, automation, monitoring, and incident management.
- Hands-on experience with Kubernetes and containerized environments.
- Strong troubleshooting and root cause analysis skills.
- Experience working in highly available, large-scale production environments.
- Excellent communication and collaboration skills.
More jobs in Dubai
Sr Account Executive
F5 · Dubai
Senior Backend Developer (Node.js+API Banking+Playwright)
ValueLabs · Dubai
Technical Account Manager, Enterprise Support - MENAT
Amazon Web Services (AWS) · Dubai
Senior Technical Account Manager, Microsegmentation
Akamai Technologies · Dubai
Senior Software Test Manager
Dyson · Dubai
Solution Engineer
AvePoint · Dubai
Associate Principal Hardware & Software Test Engineer
Dyson · Dubai
Card and Payments Consultant ( PowerCard system)
Dicetek LLC · Dubai
Browse related jobs
Don’t just read the job — see if you’ll get it.
Get your match score, a resume tailored to this exact role, and jobs like it — free.
Check my fit for this job