HamburgerMenu
hirist

Senior DevOps Engineer - Site Reliability

HriZen
5 - 10 Years
Multiple Locations

Posted on: 19/05/2026

Job Description

Location : Gurgaon-Hybrid

Experience : 5-10 Years

Role Overview :

We are looking for a Senior DevOps Engineer to build, scale, and operate highly available infrastructure for our messaging platform. You will work closely with Engineering, SRE, and Security teams to ensure reliability, performance, and fast deployments.

Key Responsibilities :

Infrastructure & Cloud :

- Design, deploy, and maintain scalable infrastructure on AWS / GCP / Azure

- Manage compute, networking, storage, and security for high-traffic systems

- Optimize infrastructure for low latency and high availability

CI/CD & Automation :

- Build and maintain robust CI/CD pipelines

- Automate deployments, rollbacks, and environment provisioning

- Improve developer productivity through tooling and automation

Reliability & Observability :

- Ensure 99.9%+ uptime for messaging services

- Set up monitoring, logging, and alerting (Prometheus, Grafana, ELK, Datadog, etc.)

- Perform root cause analysis (RCA) and incident management

Scalability & Performance :

- Support real-time workloads (chat, presence, notifications)

- Tune systems for throughput, concurrency, and resilience

- Implement load balancing, autoscaling, and failover strategies

Security & Compliance :

- Implement security best practices (IAM, secrets management, encryption)

- Ensure secure deployments and access controls

- Support compliance and data protection requirements

Required Skills & Qualifications :

- 5+ years of experience in DevOps / SRE

- Strong experience with Linux, Bash, and scripting

- Expertise in Docker and Kubernetes

- Experience with Terraform / CloudFormation / IaC tools

- Hands-on experience with cloud platforms (AWS/GCP/Azure)

- Strong understanding of networking, DNS, load balancing

- Experience managing high-scale distributed systems

Good to Have :

- Experience with Messaging / Chat / Real-time systems

- Exposure to Kafka, RabbitMQ, Redis

- Experience with Erlang / Go / Java backend platforms

- Knowledge of Zero-downtime deployments

- SRE practices (SLIs, SLOs, error budgets)

Key Systems Youll Work With :

- Messaging servers & real-time gateways

- Databases & caches

- CI/CD pipelines

- Monitoring & alerting platforms

- Security and access management

If this opportunity excites you and matches your skillset, Please share the following details :

Updated Resume :
Total Experience :
Technologies worked on :
Current Company :
Current Designation :
Current CTC (Fixed + Variable+ stocks if any) :
Expected CTC (Please provide a number) :
Any offer in hand :
Notice Period to join :
Current location :
Reason for Change :
Ready for Work From Office
Availability for Hangout Interview :

info-icon

Did you find something suspicious?

Similar jobs that you might be interested in

Loading chat...