HamburgerMenu
hirist

Staff Site Reliability Engineer/Platform Engineer - CI/CD Pipeline

Scaling Theory Technologies
5 - 12 Years
rupee30-80 LPA
Bangalore

Posted on: 09/07/2026

Job Description

CRole Overview:

As a Staff Site Reliability Engineer, you will serve as a technical anchor for our infrastructure initiatives, bridging the gap between software development and system operations. You will spend your day architecting resilient cloud environments, automating complex deployment workflows, and driving the strategy for our multi-cloud footprint. Working closely with cross-functional product and engineering teams, you will ensure that our systems remain performant and fault-tolerant as we scale. Your work directly impacts the stability of our platform, ensuring seamless experiences for our users and setting the standard for operational excellence across the organization.

Key Responsibilities:

- Architect and maintain highly available, multi-cloud infrastructure to ensure consistent uptime and performance for global users.

- Design and implement automated CI/CD pipelines that accelerate release cycles while maintaining rigorous security and quality standards.

- Lead the migration and optimization of containerized workloads on Kubernetes to improve resource utilization and system scalability.

- Develop Infrastructure as Code (IaC) frameworks that allow engineering teams to provision resources reliably and repeatably.

- Establish comprehensive monitoring, logging, and alerting strategies to proactively identify and resolve system bottlenecks before they impact the end-user.

- Mentor junior engineers and foster a culture of reliability, incident response, and continuous improvement across the technical organization.

Required Skillset:

- Demonstrated expertise in managing large-scale, multi-cloud environments (AWS, GCP, or Azure) with a focus on cost-efficiency and performance.

- Deep technical proficiency in Kubernetes orchestration, including cluster management, service mesh integration, and security hardening.

- Proven ability to write and maintain production-grade Infrastructure as Code using tools like Terraform, Pulumi, or CloudFormation.

- Strong background in building and optimizing CI/CD pipelines using modern tooling to facilitate rapid, safe code deployments.

- Exceptional problem-solving skills with the ability to diagnose and resolve complex distributed system failures under pressure.

- Excellent communication skills, with the ability to articulate technical strategies to stakeholders and collaborate effectively within a hybrid work environment.

- A minimum of 5 to 12 years of relevant experience in platform engineering, DevOps, or site reliability roles.

- A Bachelor's or Master's degree in Computer Science, Engineering, or a related technical field from a premier institute.

info-icon

Did you find something suspicious?

Similar jobs that you might be interested in

Loading chat...