HamburgerMenu
hirist

DevOps Consultant - Site Reliability

Helius Technologies
7 - 10 Years
Hyderabad

Posted on: 22/07/2026

Job Description

About the job :

At Helius Technologies, we help organisations deliver technology outcomes that work in the real world. Our teams support complex technology initiatives across Asia and the Middle East, particularly in regulated and enterprise environments where reliability, governance, and execution matter.


Our focus is simple : to be a trusted partner in technology transformation - delivering outcomes reliably and built to last. We are currently looking for a SRE Devops Professional to join our team in Hyderabad. In this role, you will work with engineering teams, stakeholders, and delivery leads to support the execution of technology initiatives. The role is suited to professionals who combine strong technical capability with a structured and accountable approach to delivery.


We value individuals who bring depth, accountability, and a strong commitment to quality. Please note role requires strictly working from our Hyderabad Office. WaveRock SEZ, 3rd Floor IT Park Road Number 2 Nanakramguda, Gachibowli, Telangana 500032, India. Immediate Starters Preferred.

About the Role :

We are seeking an experienced SRE / DevOps Engineer with strong hands-on expertise in Grafana, Prometheus, and modern observability stacks to help build and scale our monitoring, alerting, and incident-response capabilities. You will be responsible for ensuring the reliability, performance, and observability of production systems, working closely with development and infrastructure teams to reduce downtime, automate operations, and drive a culture of engineering excellence.

Key Responsibilities :

- Design, build, and maintain observability pipelines using Grafana, Prometheus, Loki, Tempo, and/or the ELK/EFK stack to provide end-to-end visibility across infrastructure and applications.

- Build and maintain Grafana dashboards for metrics, logs, and traces, tailored to different stakeholders (engineering, SRE, leadership/business).

- Define and implement SLIs, SLOs, and error budgets, and set up meaningful alerting (Alertmanager, PagerDuty, Opsgenie) to reduce noise and improve mean-time-to-detect (MTTD) and mean-time-to-resolve (MTTR).

- Own and improve CI/CD pipelines (GitHub Actions, GitLab CI, Jenkins, ArgoCD) to enable safe, frequent, and automated deployments.

- Manage and scale Kubernetes clusters (EKS/AKS/GKE or self-managed) and containerized workloads, ensuring high availability and resilience.

- Automate infrastructure provisioning and configuration using Infrastructure as Code (Terraform, Ansible, CloudFormation).

info-icon

Did you find something suspicious?

Similar jobs that you might be interested in

Loading chat...