HamburgerMenu
hirist

Senior Cloud Monitoring Engineer - Azure/Google Cloud Platform

BCForward
8 - 10 Years
Multiple Locations

Posted on: 17/09/2026

Job Description

Role Overview :


As a Senior Cloud Monitoring Engineer, you will serve as a technical authority for the observability and performance management of complex cloud ecosystems. You will work closely with cross-functional DevOps, SRE, and infrastructure teams to design, implement, and optimize monitoring strategies that ensure high availability and optimal performance for global clients. Your daily contributions will directly influence the reliability of production environments, enabling stakeholders to make data-driven decisions while minimizing downtime and operational friction.


Key Responsibilities :


- Architect and deploy robust monitoring and observability frameworks across multi-cloud environments to ensure end-to-end visibility into application and infrastructure health.

- Analyze performance metrics and log data to identify bottlenecks, proactively resolving incidents before they impact end-user experience.



- Collaborate with engineering teams to integrate automated alerting and incident response workflows, reducing mean time to resolution (MTTR) for critical production issues.



- Define and maintain Service Level Objectives (SLOs) and Service Level Indicators (SLIs) to align technical performance with business requirements and client expectations.



- Mentor junior engineers and provide technical guidance on best practices for cloud monitoring, capacity planning, and cost optimization strategies.


Required Skillset :


- Demonstrated expertise in managing enterprise-grade monitoring tools such as Datadog, New Relic, Dynatrace, or Prometheus/Grafana stacks within large-scale cloud environments.

- Advanced proficiency in cloud platforms (AWS, Azure, or GCP) with a deep understanding of cloud-native monitoring services and infrastructure-as-code principles.



- Strong scripting capabilities in Python, Go, or Bash to automate monitoring tasks and integrate disparate observability tools.



- Proven ability to communicate complex technical insights to non-technical stakeholders, fostering a culture of transparency and proactive problem-solving.



- Exceptional analytical skills to troubleshoot distributed systems and resolve complex performance degradation issues in high-traffic production environments.



- A Bachelor's degree in Computer Science, Information Technology, or a related field, combined with 8 to 10 years of professional experience in cloud infrastructure or site reliability engineering.



- Ability to thrive in a hybrid work environment in Hyderabad, demonstrating self-motivation and the capacity to lead technical initiatives independently.


info-icon

Did you find something suspicious?

Similar jobs that you might be interested in

Loading chat...