HamburgerMenu
hirist

Job Description

Job Description :

We are looking for a highly skilled DevOps Engineer with strong expertise in Kubernetes (K8s), Cloud Infrastructure, CI/CD Automation, Containerization, and Application Development. The ideal candidate should not only be capable of managing and optimizing cloud-native infrastructure but also possess the ability to understand existing application codebases and contribute to software development when required.

This role requires a blend of DevOps, Site Reliability Engineering (SRE), Platform Engineering, and Development skills to support application deployment, scalability, reliability, and performance across modern cloud environments.

Key Responsibilities :

1. Kubernetes Administration & Container Orchestration :

- Design, deploy, manage, and optimize Kubernetes clusters in production environments.

- Configure and manage :

1. Pods

2. Deployments

3. ReplicaSets

4. StatefulSets

5. DaemonSets

6. Services

7. Ingress Controllers

8. Namespaces

9. ConfigMaps

10. Secrets

- Perform cluster upgrades, patching, and maintenance activities.

- Implement auto-scaling using :

1. Horizontal Pod Autoscaler (HPA)

2. Vertical Pod Autoscaler (VPA)

3. Cluster Autoscaler

- Troubleshoot cluster, networking, and application deployment issues.

- Optimize Kubernetes resource utilization and performance.

2. DevOps & CI/CD Automation :

- Design and maintain CI/CD pipelines.

- Automate application build, testing, deployment, and release processes.

- Implement Infrastructure as Code (IaC) solutions.

- Improve deployment efficiency and release reliability.

- Support GitOps implementation and automation practices.

- Build deployment strategies including :

1. Blue-Green Deployment

2. Canary Deployment

3. Rolling Updates

3. Cloud Infrastructure Management :

- Deploy and manage cloud-native solutions on :

1. AWS

2. Azure

3. Google Cloud Platform (GCP)

- Manage :

1. Compute services

2. Networking

3. Storage

4. Security configurations

5. Monitoring solutions

- Optimize cloud resource utilization and cost management.

4. Infrastructure as Code (IaC) :

- Develop and maintain infrastructure automation using :

1. Terraform

2. Ansible

3. CloudFormation (preferred)

- Automate provisioning and configuration management.

- Maintain reusable infrastructure modules and templates.

5. Development & Code Contribution :

- Understand existing application architecture and source code.

- Participate in debugging and root cause analysis.

- Modify and enhance application code when required.

- Support development teams during deployment and production issues.

- Review code and contribute to feature enhancements.

- Work closely with software engineers to improve application reliability and deployment readiness.

6. Monitoring & Reliability Engineering :

- Implement observability solutions for infrastructure and applications.

- Configure :

1. Prometheus

2. Grafana

3. ELK Stack

4. OpenTelemetry

5. Splunk (preferred)

- Monitor :

1. Application health

2. Infrastructure performance

3. Cluster utilization

4. Security events

- Define SLIs, SLOs, and SLA metrics.

7. Security & Compliance :

- Implement Kubernetes security best practices.

- Manage :

1. RBAC

2. Network Policies

3. Secrets Management

4. Container Security

- Perform vulnerability scanning and remediation.

- Implement DevSecOps practices throughout CI/CD pipelines.

8. Incident Management & Troubleshooting :

- Investigate and resolve :

1. Production incidents

2. Deployment failures

3. Performance bottlenecks

4. Infrastructure outages

- Conduct Root Cause Analysis (RCA).

- Support high availability and disaster recovery initiatives.

Required Technical Skills :

1. Kubernetes (Mandatory) :

- Strong hands-on experience in :

1. Kubernetes Administration

2. Helm Charts

3. Kubernetes Networking

4. Ingress Controllers

5. Service Mesh (Istio preferred)

6. Cluster Management

7. Cluster Security

8. Troubleshooting Production Clusters

2. Containerization :

- Docker

- Container Image Management

- Container Registry Administration

3. CI/CD Tools :

- Experience with :

1. Jenkins

2. GitHub Actions

3. GitLab CI/CD

4. Azure DevOps Pipelines

5. ArgoCD

6. FluxCD

4. Infrastructure as Code :

- Terraform

- Ansible

- CloudFormation (Preferred)

5. Cloud Platforms :

- Strong experience in at least one :

1. Amazon Web Services (AWS)

2. Microsoft Azure

3. Google Cloud (GCP)

6. Programming / Development Skills (Mandatory) :

- Hands-on experience in at least one :

1. Java

2. Python

3. Go (Golang)

4. Node.js

5. C#

- Candidate should be able to :

1. Read and understand code.

2. Fix bugs.

3. Implement minor enhancements.

4. Support development teams during production issues.

7. Source Control :

- Git

- GitHub

- GitLab

- Bitbucket

8. Monitoring & Logging :

- Prometheus

- Grafana

- ELK Stack

- Splunk

- OpenTelemetry

Preferred Skills :

- Site Reliability Engineering (SRE)

- DevSecOps

- Service Mesh (Istio)

- Kafka Administration

- Redis

- NGINX

- API Gateway Management

- Microservices Architecture

- Linux Administration

Education :

- Bachelor's Degree in Computer Science, Information Technology, Engineering, or related discipline.

info-icon

Did you find something suspicious?

Similar jobs that you might be interested in

Loading chat...