Site Reliability Engineer - Lead at Avrioc Technologies
dubai, Abu Dhabi, United Arab Emirates -
Full Time


Start Date

Immediate

Expiry Date

18 Nov, 26

Salary

0.0

Posted On

20 Aug, 26

Experience

0 year(s) or above

Remote Job

Yes

Telecommute

Yes

Sponsor Visa

Yes

Skills

Industry

Information Technology & Services

Description

Key Responsibilities:

• Architect and deploy scalable, highly available cloud infrastructure

• Lead SRE best practices to ensure reliability, performance, and scalability

• Optimize CI/CD pipelines (Jenkins, Argo CD or similar) for seamless deployments

• Define and track SLOs & SLIs to maintain uptime and service health

• Build robust observability frameworks (Elastic Stack, Prometheus, Grafana, Dynatrace, New Relic)

• Manage Kubernetes clusters and Helm charts for efficient orchestration

• Implement auto-healing systems and proactive monitoring

• Drive chaos engineering and resilience testing (Chaos Mesh, Litmus, AWS FIS)

• Collaborate with engineering and product teams to embed reliability into development

• Maintain clear infrastructure and incident documentation


What We’re Looking For:

• 8+ years of experience in DevOps/SRE, including leadership in enterprise environments

• Hands-on experience with AWS, GCP, or Azure

• Strong expertise in Infrastructure as Code (Terraform, CloudFormation, Ansible)

• Proven experience in CI/CD, monitoring, and incident response

• Deep knowledge of observability tools and practices

• Strong Kubernetes and Helm experience at scale

• Experience with databases like MySQL, Cassandra, etc.

• Proficiency in Python, Bash, or Go

• Experience in BCP/DR planning and capacity management


How To Apply:

Incase you would like to apply to this job directly from the source, please click here

Responsibilities
Loading...