Posted on: 15/09/2026
Mandatory Skills:
- Strong hands-on experience in Core Linux Administration.
- Minimum 4 years of hands-on experience with PCS/Pacemaker Clustering on RHEL.
- Strong knowledge of RHEL (Red Hat Enterprise Linux) server administration.
- Experience working in an SRE (Site Reliability Engineering) environment.
- Excellent communication skills with the ability to interact directly with clients and stakeholders.
- Strong troubleshooting and problem-solving skills related to Linux servers and cluster environments.
Secondary Skills:
- VMware / virtualization technologies.
- AWS or any other Cloud Platform.
- Oracle / Oracle Database environment.
Key Responsibilities:
- Manage day-to-day Linux administration, maintenance, and production support activities.
- Monitor, maintain, and troubleshoot RHEL/Linux servers.
- Configure, monitor, maintain, and troubleshoot PCS/Pacemaker clusters.
- Handle cluster-related activities including high availability, failover/failback, fencing, and resource management.
- Monitor system health, availability, capacity, and performance.
- Troubleshoot Linux and cluster-related incidents and provide timely resolution.
- Perform root cause analysis (RCA) for recurring or critical incidents.
- Work closely with client stakeholders, support teams, and other technical teams.
- Ensure system reliability, availability, stability, and performance.
- Participate in incident, problem, and change management activities within an SRE/production support environment.
- Follow defined SLA requirements and operational procedures.
- Support rotational shifts as required.
Did you find something suspicious?