Posted on: 30/09/2026
Key Responsibilities :
- Define and drive the architecture of large-scale, fault-tolerant distributed systems across multi-cloud environments.
- Design, build, and operate highly available and scalable platforms across AWS, Google Cloud, and/or Azure.
- Architect and manage enterprise-scale Kubernetes platforms, including complex multi-tenant environments.
- Establish cloud-native architecture standards, platform engineering practices, and infrastructure strategies.
- Drive large-scale, cross-team technical initiatives from architecture and design through production deployment.
- Influence technical direction across engineering teams without direct reporting authority.
- Develop multi-cloud and cross-cloud strategies covering scalability, reliability, security, cost optimization, and operational resilience.
- Identify emerging technologies and translate them into practical technology strategies and platform capabilities.
- Lead organizational AI platform strategy and GenAI adoption initiatives.
- Design intelligent platforms incorporating autonomous and predictive capabilities.
- Drive development and adoption of AI-powered engineering solutions such as developer agents, code-review agents, testing agents, and SRE/operations agents.
- Establish platform capabilities that improve developer productivity, system reliability, automation, and operational efficiency.
- Define technical roadmaps aligned with long-term business and technology objectives.
- Mentor senior engineers and architects and promote engineering best practices.
- Communicate complex technical concepts effectively to engineering teams, architects, and executive stakeholders.
Technical Expertise :
- Distributed Systems & Large-Scale Architecture
- AWS / Google Cloud / Azure
- Multi-Cloud & Cross-Cloud Architecture
- Kubernetes & Cloud-Native Platforms
- Containerization & Platform Engineering
- High Availability & Fault-Tolerant Systems
- Infrastructure & Platform Automation
- Generative AI / LLMs / AI Platforms
- AI Agents & Autonomous Engineering
- Developer Productivity Platforms
- SRE & Predictive Operations
- Cloud Security, Reliability & Scalability
Required Candidate Profile :
- 10+ years of experience in software and platform engineering.
- Significant experience in Staff Engineer, Principal Engineer, Platform Architect, or similar technical leadership roles.
- Deep expertise in designing, building, and operating large-scale distributed systems.
- Strong experience architecting multi-cloud platforms across AWS and Google Cloud and/or Azure.
- Expert-level knowledge of Kubernetes and the broader cloud-native ecosystem.
- Experience architecting and operating Kubernetes platforms at enterprise scale, including multi-tenant environments.
- Proven ability to lead complex, cross-functional technical initiatives from conception through production.
- Strong understanding of cloud infrastructure, scalability, resilience, security, and operational excellence.
- Experience defining multi-cloud architecture and infrastructure strategies.
- Strong understanding of Generative AI and experience driving AI platform or GenAI adoption initiatives.
- Excellent communication, stakeholder management, technical leadership, and influencing skills.
- Strategic mindset with the ability to anticipate future technology challenges and translate them into actionable technical roadmaps.
Preferred Candidate :
- Experience building enterprise-scale internal developer platforms or cloud platforms.
- Experience implementing AI-powered engineering and SRE solutions.
- Strong background in platform reliability, automation, observability, and DevOps/SRE.
- Experience working across multiple engineering teams and influencing architecture decisions at an organizational level.
Did you find something suspicious?
Posted by
Posted in
DevOps / SRE
Functional Area
DevOps / Cloud
Job Code
1676011