HamburgerMenu
hirist

Job Description

Role : Platform Engineer - Service Integration - BLR

Location : Bengaluru, India

Role Summary :

We are seeking a platform expert proficient in Puppet and Ansible ecosystems. You will be comfortable redesigning the entire internal platform integration, migrating from legacy Puppet 5 to Puppet 7 or later, and consolidating/redesigning multiple server configuration technologies, including Ansible, Puppet 7, and Puppet 5.

A key focus is delivering next-generation integration within Kubernetes (K8S) or VM environments (Nutanix, VMware, Kubevirt). Strong VMware knowledge and Kubernetes knowledge is a plus.

We are looking for someone with :

- Deep knowledge of the Puppet 5/7 ecosystems, covering both user and development experience.

- A passion for developing internal tools, building platform integrations, and evolving deployment pipelines for applications and services.


What You Will Do :

- Design, architect, and lead the Puppet ecosystem redesign, retiring the legacy Puppet 5 system

- Gather requirements and use cases to build solutions.

- Redesign the Puppet ecosystem and infrastructure to an enterprise-standard, at-scale solution.

- Work closely with the datacenter and virtual infrastructure teams to build platform integration.

- Implement the next-gen solution.

- Develop internal tools and workflow automations.

- Manage interrupts and lead incident response during critical incidents.

- Build Automation and Tools using Python and Shell Scripting

- Develop Python-based tools and automation to support self-service workflows.

- Automate repetitive operational tasks and improve existing deployment script efficiency.

- Integrate with internal and external APIs to orchestrate infrastructure workflows (compute, storage, network).

- Configuration Management, Infrastructure, and Monitoring as Code (IaC)

- Use tools such as Ansible, Terraform, Puppet, or similar tools to manage infrastructure declaratively.

- Maintain reusable playbooks/modules and templates for common infrastructure patterns.

- Enforce configuration standards, security baselines, and repeatable deployments across environments.

Monitoring, Observability, and Reliability :

- Implement and improve monitoring, alerting, and dashboards for infrastructure health (e.g., Prometheus, Grafana, ELK/Nagios, or similar tools).

- Define and track key metrics (availability, latency, capacity, error rates) and drive improvements based on data.

- Participate in incident response, perform root cause analysis, and implement long-term fixes and runbooks.

- Maintain the observability codebase and actively develop new monitors and alerts. Minimum Qualifications

- 5+ years of experience in Infrastructure/Platform Engineer or SRE/DevOps engineering roles. - 5+ years of hands-on Python or Go development for :

1. Automation scripts and tools

2. REST API integrations

3. Internal tooling development

- Strong experience with configuration management or IaC tools (e.g., Ansible, Terraform, Puppet).

- Strong Linux systems administration knowledge (e.g., Ubuntu, CentOS/RHEL).

- Strong experience with Kubernetes development and deployment.

- Strong knowledge of on-premises infrastructure and deployment.

- Experience with office tools such as Jira, Slack, Google Workspace, and others.

- Knowledge of monitoring and alerting on key storage metrics.

- Strong problem-solving skills, ownership mindset, and clear written and verbal communication. Plus Points

- Cloud experience with CI/CD and service management.

- SRE/DevOps knowledge, monitoring and alerting philosophy, and Incident Commander system knowledge.

- Exposure to CI/CD tooling and pipelines (e.g., Jenkins, GitHub Actions, ArgoCD, GitLab CI).

- Experience in globally distributed teams and "follow-the-sun" support models.

info-icon

Did you find something suspicious?

Similar jobs that you might be interested in

Loading chat...