Posted on: 01/09/2026
Role Overview :
We are looking for a hands-on DevOps Engineer who can manage production infrastructure, improve deployment reliability, automate operational processes, and work closely with Engineering teams.
The ideal candidate should have real production experience with AWS and Kubernetes/EKS, strong infrastructure-as-code skills using Terraform, and a solid understanding of CI/CD, monitoring, security, and cloud operations.
Key Responsibilities :
- Manage and maintain production infrastructure and workloads on AWS.
- Deploy, manage, and troubleshoot production applications running on Kubernetes/EKS.
- Design, maintain, and improve CI/CD pipelines for reliable and repeatable deployments.
- Build and maintain infrastructure using Terraform and infrastructure-as-code best practices.
- Automate repetitive operational tasks, deployments, infrastructure provisioning, and environment management.
- Monitor production systems, applications, infrastructure health, availability, and performance.
- Troubleshoot production incidents, identify root causes, and implement preventive solutions.
- Work closely with developers to improve application deployment, scalability, reliability, and observability.
- Manage cloud resources, networking, compute, storage, IAM, security, and access controls on AWS.
- Implement logging, monitoring, alerting, and operational dashboards.
- Support capacity planning, performance optimisation, backup, disaster recovery, and high availability.
- Maintain DevOps documentation, runbooks, deployment procedures, and operational standards.
Required Skills :
- Hands-on experience in DevOps with significant exposure to production environments.
- Strong hands-on experience with AWS cloud services and production infrastructure.
- Strong hands-on experience with Kubernetes and Amazon EKS.
- Strong proficiency in Terraform and infrastructure-as-code.
- Experience designing and managing CI/CD pipelines.
- Strong understanding of Linux systems, networking, DNS, load balancing, and troubleshooting.
- Experience with Docker and containerised application deployments.
- Experience with monitoring and logging tools and production observability.
- Strong understanding of cloud security, IAM, secrets management, and access controls.
- Good scripting/automation skills using Bash, Python, or similar languages.
- Strong troubleshooting, incident management, and root-cause analysis skills.
- Good communication and collaboration skills.
Must-Have Production Experience :
- 5+ years of actual hands-on DevOps experience - not only system administration or development exposure.
- Production experience managing AWS infrastructure and workloads.
- Production experience working with Kubernetes/EKS.
- Hands-on Terraform experience for infrastructure provisioning and management.
- Experience owning or actively contributing to production CI/CD pipelines.
- Experience troubleshooting live production issues and improving system reliability.
Good to Have :
- Experience with Helm and Kubernetes deployment patterns.
- Experience with GitHub Actions, GitLab CI, Jenkins, or similar CI/CD platforms.
- Experience with Prometheus, Grafana, ELK/OpenSearch, CloudWatch, or similar monitoring/logging platforms.
- Experience with AWS networking including VPC, subnets, security groups, load balancers, and Route 53.
- Experience with ArgoCD or GitOps practices.
- Experience implementing DevSecOps practices.
- Experience with high-availability and auto-scaling production architectures.
- Experience working in a fast-paced product or startup environment.
Ideal Candidate :
Someone who is highly hands-on, production-focused, and comfortable taking ownership of cloud infrastructure and deployment reliability.
The ideal candidate should be able to independently troubleshoot production issues, automate infrastructure and operational workflows, improve deployment processes, and work closely with engineering teams to build scalable and reliable systems.
Key Hiring Preference :
- Candidates with strong hands-on AWS + Kubernetes/EKS + Terraform experience and proven production DevOps ownership will be prioritised.
- Candidates should be able to clearly explain their actual production responsibilities, infrastructure ownership, deployment pipelines, incidents handled, and automation work.
Did you find something suspicious?
Posted by
Posted in
DevOps / SRE
Functional Area
DevOps / Cloud
Job Code
1667597