Posted on: 03/10/2026
Job Description :
We are seeking an experienced Platform Operations Engineer to join our team. The role is responsible for continuous monitoring of our platform, rapid incident response, and execution of operational procedures to ensure high availability and reliability of services.
Key Responsibilities :
- Monitor platform health 24x7 and acknowledge incoming alerts
- Perform firstlevel incident troubleshooting and investigation
- Coordinate incident recovery, escalation, and communication with stakeholders
- Execute approved diagnostic commands and healthcheck procedures
- Initiate GitOpsbased redeployments as required by operational policies
- Develop, maintain, and follow operational runbooks and procedures
Required Experience :
- Total professional experience of 510 years
- Minimum 5 years in IT operations, infrastructure, cloud or application support environments
Required Skills :
- Proficient with Linux commandline tools and basic system administration
- Strong networking knowledge including IP addressing, DNS, ports, and connectivity troubleshooting
- Familiarity with container technologies and core Kubernetes concepts
- Ability to follow and execute documented diagnostic and remediation steps
- Experience working in a 24x7 support model with incident coordination and escalation
Preferred Skills :
- (None specified)
Recruiter Summary :
- 510 years overall experience, with at least 5 years in IT operations or cloud support
- Role focuses on 24x7 platform monitoring, incident triage, and GitOps redeployments
- Requires solid Linux CLI skills and foundational networking knowledge
- Must understand containers and basic Kubernetes concepts
- Emphasis on following operational procedures and effective incident coordination
Did you find something suspicious?
Posted by
Rupali Chavan
Head at NeedsO'Need Recruitment Consultancy
Last Active: NA as recruiter has posted this job through third party tool.
Posted in
DevOps / SRE
Functional Area
IT Management / IT Support
Job Code
1676433