Posted on: 29/04/2026
Role : Platform Engineer - Service Integration - BLR
Location : Bengaluru, India
Role Summary :
We are seeking a platform expert proficient in Puppet and Ansible ecosystems. You will be comfortable redesigning the entire internal platform integration, migrating from legacy Puppet 5 to Puppet 7 or later, and consolidating/redesigning multiple server configuration technologies, including Ansible, Puppet 7, and Puppet 5.
A key focus is delivering next-generation integration within Kubernetes (K8S) or VM environments (Nutanix, VMware, Kubevirt). Strong VMware knowledge and Kubernetes knowledge is a plus.
We are looking for someone with :
- Deep knowledge of the Puppet 5/7 ecosystems, covering both user and development experience.
- A passion for developing internal tools, building platform integrations, and evolving deployment pipelines for applications and services.
What You Will Do :
- Design, architect, and lead the Puppet ecosystem redesign, retiring the legacy Puppet 5 system
- Gather requirements and use cases to build solutions.
- Redesign the Puppet ecosystem and infrastructure to an enterprise-standard, at-scale solution.
- Work closely with the datacenter and virtual infrastructure teams to build platform integration.
- Implement the next-gen solution.
- Develop internal tools and workflow automations.
- Manage interrupts and lead incident response during critical incidents.
- Build Automation and Tools using Python and Shell Scripting
- Develop Python-based tools and automation to support self-service workflows.
- Automate repetitive operational tasks and improve existing deployment script efficiency.
- Integrate with internal and external APIs to orchestrate infrastructure workflows (compute, storage, network).
- Configuration Management, Infrastructure, and Monitoring as Code (IaC)
- Use tools such as Ansible, Terraform, Puppet, or similar tools to manage infrastructure declaratively.
- Maintain reusable playbooks/modules and templates for common infrastructure patterns.
- Enforce configuration standards, security baselines, and repeatable deployments across environments.
Monitoring, Observability, and Reliability :
- Implement and improve monitoring, alerting, and dashboards for infrastructure health (e.g., Prometheus, Grafana, ELK/Nagios, or similar tools).
- Define and track key metrics (availability, latency, capacity, error rates) and drive improvements based on data.
- Participate in incident response, perform root cause analysis, and implement long-term fixes and runbooks.
- Maintain the observability codebase and actively develop new monitors and alerts. Minimum Qualifications
- 5+ years of experience in Infrastructure/Platform Engineer or SRE/DevOps engineering roles. - 5+ years of hands-on Python or Go development for :
1. Automation scripts and tools
2. REST API integrations
3. Internal tooling development
- Strong experience with configuration management or IaC tools (e.g., Ansible, Terraform, Puppet).
- Strong Linux systems administration knowledge (e.g., Ubuntu, CentOS/RHEL).
- Strong experience with Kubernetes development and deployment.
- Strong knowledge of on-premises infrastructure and deployment.
- Experience with office tools such as Jira, Slack, Google Workspace, and others.
- Knowledge of monitoring and alerting on key storage metrics.
- Strong problem-solving skills, ownership mindset, and clear written and verbal communication. Plus Points
- Cloud experience with CI/CD and service management.
- SRE/DevOps knowledge, monitoring and alerting philosophy, and Incident Commander system knowledge.
- Exposure to CI/CD tooling and pipelines (e.g., Jenkins, GitHub Actions, ArgoCD, GitLab CI).
- Experience in globally distributed teams and "follow-the-sun" support models.
Did you find something suspicious?
Posted by
Posted in
DevOps / SRE
Functional Area
DevOps / Cloud
Job Code
1632106