Posted on: 01/09/2026
Sr Software Engineer - Infrastructure Platform (Python+Puppet)
We are seeking an experienced Platform Engineer to lead the modernization of our core infrastructure services. You will be responsible for redesigning our global Puppet ecosystem - migrating from legacy Puppet 5 to Puppet 7+ - while simultaneously scaling and advancing the Puppet infrastructure ecosystem to meet global enterprise needs.
Key Responsibilities:
- Redesign & Consolidate: Lead the architectural redesign of the Puppet ecosystem, migrating legacy Puppet 5 systems to Puppet 7+.
- Technology Consolidation: Consolidate multiple configuration management technologies (Ansible, Puppet) into a unified, enterprise-standard, and at-scale solution.
- IaC Implementation: Maintain reusable playbooks/modules and templates to enforce configuration standards, security baselines, and repeatable deployments across environments.
- PuppetDB & Hiera Optimization: Architect and manage PuppetDB for high availability and performance. Optimize complex Hiera hierarchies and data lookups to ensure efficient configuration management at scale.
- Development Workflows: Improve Puppet code development lifecycles by implementing robust testing frameworks (e.g., RSpec, Beaker) and streamlining CI/CD pipelines for configuration changes.
- Global Scalability: Design and maintain a geographically distributed Puppet architecture that supports thousands of nodes across global datacenters with minimal latency and high reliability.
Python Automation & Platform Integration:
- API-First Development: Develop enterprise-grade Python tools and APIs to orchestrate infrastructure workflows (compute, storage, network, cloud).
- Self-Service Workflows: Build automation to support self-service workflows, reducing manual operational tasks and improving deployment script efficiency.
- Orchestration: Integrate internal/external APIs to facilitate automated service provisioning and resource management.
Observability & Reliability:
- Monitoring: Implement and improve monitoring, alerting, and dashboards for infrastructure health (e.g., Prometheus, Grafana, ELK/Nagios).
- Incident Response: Participate in incident response, perform root cause analysis, and implement long-term fixes/runbooks.
Required Skills & Experience:
- Python Programming: 5+ years of enterprise software development experience, focusing on automation scripts, internal tooling, and REST API integrations.
- Puppet Expertise: Deep, hands-on knowledge of Puppet 5/7 ecosystems, including both user and development experience.
- Advanced Puppet Mastery: Expert-level experience in Puppet architecture, including advanced Hiera data modeling and complex custom module development.
- Configuration Management: Strong experience with IaC tools such as Ansible, Terraform, or Puppet.
- Infrastructure: Proficiency with Linux systems administration, bare-metal provisioning, and hypervisor deployment.
Plus Points:
- CI/CD Pipeline Experience: Exposure to tooling such as Jenkins, GitHub Actions, ArgoCD, or GitLab CI.
- SRE/DevOps: Background in SRE methodologies, observability, and incident management.
- Cloud & Kubernetes: Experience with large-scale internal platforms, Kubernetes deployment, and hybrid/multi-datacenter environments.
- Infrastructure as Product: Experience operating self-service platforms and developer enablement.
Did you find something suspicious?
Posted by
Posted in
DevOps / SRE
Functional Area
DevOps / Cloud
Job Code
1667790