HamburgerMenu
hirist

Job Description

AblyWorks is looking for an experienced DevOps / Site Reliability Engineer (SRE) to build, automate, and maintain scalable cloud infrastructure and deployment pipelines. The ideal candidate will have strong expertise in cloud platforms, Kubernetes, CI/CD, Infrastructure as Code, monitoring, and production support, with a focus on improving system reliability, performance, and security.

Responsibilities :

- Design, implement, and manage cloud infrastructure on AWS, Azure, or GCP.

- Develop and maintain CI/CD pipelines for automated application deployments.

- Automate infrastructure provisioning using Terraform, Ansible, or CloudFormation.

- Deploy and manage containerised applications using Docker and Kubernetes.

- Monitor infrastructure and applications using Prometheus, Grafana, ELK, CloudWatch, Datadog, or similar tools.

- Troubleshoot production issues, perform root cause analysis, and implement preventive measures.

- Ensure high availability, scalability, security, and performance of production environments.

- Collaborate with development teams to improve deployment processes and operational efficiency.

- Implement logging, monitoring, backup, disaster recovery, and security best practices.

- Participate in on-call support and production release activities.

Requirements :

- 6+ years of experience in DevOps, SRE, or Cloud Engineering.

- Strong hands-on experience with AWS, Azure, or GCP.

- Expertise in Docker, Kubernetes, and container orchestration.

- Experience with CI/CD tools such as Jenkins, GitHub Actions, GitLab CI/CD, or Azure DevOps.

- Hands-on experience with Terraform, Ansible, or CloudFormation.

- Strong Linux administration and scripting skills (Bash/Python).

- Experience with monitoring and observability tools such as Prometheus, Grafana, ELK, Datadog, or CloudWatch.

- Good understanding of networking, load balancing, DNS, SSL/TLS, and security best practices.

- Strong troubleshooting, analytical, and problem-solving skills.

Preferred Skills :

- Experience with microservices architecture.

- Knowledge of Helm, ArgoCD, or GitOps practices.

- Exposure to SRE concepts including SLIs, SLOs, and Error Budgets.

- Cloud certifications (AWS/Azure/GCP) are an added advantage.

- Bachelor's degree in Computer Science, Information Technology, or a related field.

info-icon

Did you find something suspicious?

Similar jobs that you might be interested in

Loading chat...