Posted on: 10/10/2026
Role : Senior DevOps Engineer
Role Overview :
We are looking for a hands-on Senior DevOps Engineer with 4+ years of experience to own and improve production infrastructure, cloud operations, automation, and deployment reliability. The ideal candidate will bring strong practical expertise in AWS, Kubernetes/EKS, Terraform, CI/CD, monitoring, and production troubleshooting.
Key Responsibilities :
- Manage, optimise, and maintain production infrastructure and workloads on AWS.
- Deploy, operate, and troubleshoot production applications running on Kubernetes/Amazon EKS.
- Design, maintain, and improve CI/CD pipelines for safe, reliable, and repeatable releases.
- Provision and manage infrastructure using Terraform and infrastructure-as-code best practices.
- Automate deployments, infrastructure provisioning, environment management, and repetitive operational tasks.
- Implement and maintain monitoring, logging, alerting, and observability for infrastructure and applications.
- Lead troubleshooting of production incidents, perform root-cause analysis, and implement preventive measures.
- Manage AWS networking, compute, storage, IAM, security, secrets, and access controls.
- Improve availability, scalability, performance, capacity planning, backups, and disaster recovery.
- Collaborate with developers and other stakeholders to improve deployment workflows and operational standards.
- Maintain technical documentation, runbooks, incident notes, and deployment procedures.
Required Skills:
- 4+ years of hands-on DevOps experience, including significant ownership of production environments.
- Strong practical experience with AWS services and cloud infrastructure.
- Hands-on experience with Kubernetes and Amazon EKS, and containerisation using Docker.
- Strong Terraform skills and experience managing infrastructure as code.
- Experience building, operating, and troubleshooting CI/CD pipelines.
- Good understanding of Linux, networking, DNS, load balancing, and production troubleshooting.
- Experience with monitoring and logging tools, alerting, and production observability.
- Understanding of cloud security, IAM, secrets management, and access controls.
- Scripting and automation skills using Bash, Python, or similar languages.
- Strong incident management, root-cause analysis, ownership, and collaboration skills.
Must-Have Production Experience:
- Hands-on ownership of AWS infrastructure and production workloads.
- Production experience with Kubernetes/EKS and Terraform-based provisioning.
- Experience owning or actively contributing to production CI/CD pipelines.
- Ability to troubleshoot live incidents and improve system reliability.
Good to Have:
- Experience with Kafka, including deployment, monitoring, or operational support.
- Exposure to data warehousing infrastructure, platforms, or related cloud data services.
- Experience with Helm and Kubernetes deployment patterns.
- Familiarity with GitHub Actions, GitLab CI, Jenkins, or similar CI/CD platforms.
- Experience with Prometheus, Grafana, ELK/OpenSearch, CloudWatch, or similar tools.
- Experience with AWS networking (VPC, subnets, security groups, load balancers, Route 53).
- Experience with ArgoCD/GitOps, DevSecOps practices, high availability, and auto-scaling architectures.
Ideal Candidate:
A highly hands-on, production-focused engineer who can independently take ownership of infrastructure, deployment reliability, automation, and incident resolution in a fast-paced product environment.
Key Hiring Preference:
- Candidates with strong hands-on AWS + Kubernetes/EKS + Terraform experience and proven production ownership will be prioritised.
- Candidates should be able to clearly explain their infrastructure responsibilities, CI/CD pipelines, incidents handled, and automation work.
Did you find something suspicious?
Posted by
Posted in
DevOps / SRE
Functional Area
DevOps / Cloud
Job Code
1677794