Posted on: 08/06/2026
Job Description :
Site Reliability Engineer (SRE) to build and maintain highly scalable, reliable, and secure cloud infrastructure. The ideal candidate will have strong expertise in automation, cloud platforms, monitoring, incident management, and DevOps practices to ensure high system availability and performance.
Responsibilities :
- Design, implement, and maintain scalable cloud infrastructure.
- Improve system reliability, availability, and performance.
- Automate deployment, monitoring, and operational processes.
- Manage production incidents and perform root cause analysis.
- Develop and maintain CI/CD pipelines.
- Monitor system health, logs, metrics, and alerts.
- Collaborate with Software Engineering teams to improve application reliability.
- Implement disaster recovery, backup, and security best practices.
- Create operational runbooks and documentation.
- Participate in on-call support rotations.
Skills :
- Strong experience with Linux/Unix Administration.
- Expertise in AWS, Azure, or Google Cloud Platform (GCP).
- Hands-on experience with Docker and Kubernetes.
- Proficiency in Python, Go, Bash, or Shell Scripting.
- Strong knowledge of Networking, DNS, TCP/IP, and Load Balancing.
Qualifications :
- BE/B.TECH
The job is for:
Did you find something suspicious?
Posted by
KBD TALENT FORGE INDIA PRIVATE LIMITED
Balkrushna Kadam at KBD Talent Forge India Pvt Ltd
Last Active: 13 Aug 2026
Posted in
DevOps / SRE
Functional Area
Site Reliability Engineering
Job Code
1642646