HamburgerMenu
hirist

MediaMint - Senior DevOps Engineer - Cloud Infrastructure

VUCHI MEDIA
5 - 10 Years
Hyderabad

Posted on: 28/09/2026

Job Description

Role Overview:

As a DevOps Engineer (On-call) at MediaMint AI Platform, you'll be the guardian of our platform's 24/7 reliability. You'll maintain critical infrastructure, respond to incidents, and continuously improve our deployment pipelines and observability systems. This role requires someone comfortable with on-call responsibilities who thrives under pressure and views incidents as learning opportunities. You'll work closely with engineering teams to ensure our AI agents run smoothly and our clients experience zero disruption.

Responsibilities:

- Participate in on-call rotation to respond to production incidents and ensure rapid resolution

- Maintain and optimize cloud infrastructure (AWS/GCP/Azure) for high availability and performance

- Automate deployment pipelines using CI/CD tools for fast, reliable releases

- Monitor system performance, set up alerts, and proactively address potential issues

- Handle escalations from support and engineering teams during critical outages

- Improve observability through logging, monitoring, and distributed tracing systems

- Implement disaster recovery procedures and conduct regular backup/restore testing

- Manage containerized workloads using Docker and Kubernetes

- Conduct capacity planning and cost optimization for cloud infrastructure

Must-Have Requirements:

- 7+ years of DevOps/SRE experience in production environments

- Strong experience with at least one major cloud platform (AWS/GCP/Azure)

- Hands-on experience with Docker and Kubernetes for container orchestration

- Proficiency with CI/CD tools (Jenkins, GitLab CI, GitHub Actions, CircleCI)

- Proven incident management skills with experience in on-call rotations

- Strong scripting skills in Python, Bash, or similar languages

- Comfortable working in on-call rotation (including nights and weekends as needed)

Nice-to-Have Requirements:

- Experience with SRE practices and error budgets

- Infrastructure as Code experience with Terraform or CloudFormation

- Hands-on experience with observability tools (Datadog, Grafana, Prometheus, New Relic)

- Experience with PagerDuty or similar incident management platforms

- Knowledge of service mesh technologies (Istio, Linkerd)

- Familiarity with chaos engineering and resilience testing

Additional Notes:

- Must be flexible to work in night shift

- Prefer local to Hyderabad or who can attend in-person interview.

info-icon

Did you find something suspicious?

Similar jobs that you might be interested in

Posted by

Recruiter

HR at VUCHI MEDIA

Last Active: NA as recruiter has posted this job through third party tool.

Job Views:  
20
Applications:  15
Recruiter Actions:  0

Posted in

DevOps / SRE

Functional Area

Site Reliability Engineering

Job Code

1675271

Loading chat...