Overview
Senior DevOps Engineer Jobs in Dubai, United Arab Emirates at Acveti
Title: Senior DevOps Engineer
Company: Acveti
Location: Dubai, United Arab Emirates
Acveti is seeking an experienced Senior DevOps / Site Reliability Engineer (SRE) to support and optimize a high-traffic marketplace platform. This role will focus on enhancing platform reliability, improving performance, strengthening observability, reducing infrastructure costs, and driving operational excellence across cloud-native environments.
Key Responsibilitie
- sManage, optimize, and scale cloud infrastructure hosted on Microsoft Azure
- .Design, deploy, and maintain containerized workloads using Azure Kubernetes Service (AKS)
- .Administer and optimize Cloudflare services for security, performance, caching, and traffic management
- .Perform in-depth analysis and tuning of PostgreSQL databases to improve performance and scalability
- .Manage and troubleshoot distributed systems leveraging Kafka and Redis
- .Build and enhance observability frameworks using Azure Monitor, Application Insights, New Relic, Grafana, and Prometheus
- .Develop and maintain robust CI/CD pipelines and release engineering processes
- .Lead infrastructure cost optimization (FinOps) initiatives, identifying opportunities to improve efficiency while maintaining performance and reliability
- .Drive incident response activities, root cause analysis (RCA), and post-incident remediation efforts
- .Identify and resolve recurring performance bottlenecks, latency issues, and HTTP 500 errors across the platform
- .Collaborate closely with engineering, product, and operations teams to improve platform stability and scalability
.
Required Skills & Experien
- ceStrong hands-on experience with Microsoft Azure and cloud-native architecture
- s.Deep expertise in Kubernetes (AKS) administration and troubleshootin
- g.Experience managing and optimizing Cloudflare environment
- s.Proven track record in PostgreSQL performance tuning and database optimizatio
- n.Strong knowledge of Kafka, Redis, and distributed system
- s.Expertise in monitoring, logging, and observability tools including Azure Monitor, Application Insights, New Relic, Grafana, and Prometheu
- s.Experience designing and maintaining modern CI/CD pipeline
- s.Practical experience with FinOps, cloud cost management, and infrastructure optimizatio
- n.Strong understanding of incident management, production support, and root cause analysis methodologie
- s.Excellent troubleshooting skills with the ability to diagnose complex performance and reliability issues in production environment
s.
Preferred Qualificati
- onsExperience supporting high-volume marketplace, e-commerce, or SaaS platfor
- ms.Strong scripting and automation skills using PowerShell, Bash, Python, or similar technologi
- es.Experience implementing SRE best practices, SLIs, SLOs, and error budge
- ts.Excellent communication and stakeholder management skil
ls.