Posted on: 24/09/2026
About the role :
This SRE professional will have the opportunity to work for an AI-based Unicorn which is recognized as one of the fastest-growing companies on the Inc. 5000 list.
This role will provide you opportunity to deploy, maintain and support the o9 Digital Brain Platform across the world on AWS, AZURE, GCP & Samsung Cloud utilizing state of the art CI/CD tools.
This role will empower you to continuously challenge the status quo and implement the great ideas you may have to create value for o9 clients.
What you'll do for us :
In this position, you may find yourself doing pretty much everything; writing infrastructure tests, fixing automation workflows, improving CI/CD, building/deploying the latest code across environments, learning new DevOps tools, catching issues, providing solutions to users.
Has to be a good team player with good personal communication skills and willing to make sure the availability of the infrastructure 24-7 to our customers.
- Deploy, maintain and support o9 digital Brain SaaS environments on all major clouds
- Monitor availability and maintain system in good health
- Build software and systems to manage platform infrastructure and applications
- Improve reliability, quality, cost, and time-to-deploy, and time-to-upgrade
- Monitor, measure and optimize system performance
- Provide on-call support on rotation basis
- Ability and flexibility to work with teams globally, across the time zones
What you'll have :
- Primary Skill: Kubernetes, Jenkins, Cloud expertise, strong in operating system concepts, Linux and troubleshooting.
- Secondary Skill: Automation
- Education: Bachelor's degree in computer science, Software Engineering, Information Technology, Industrial Engineering, Engineering Management
- Cloud (at least one) and Kubernetes administration certification
- Experience: 6+ years of experience in a SRE role, deploying and maintaining applications, performance tuning, conducting application upgrades, patches, and supporting continuous integration and deployment tooling
- 2+ years of experience deploying and maintaining applications in any one of the clouds (AWS, AZURE, GCP)
- Experience with Dockers or similar and experience with Kubernetes or similar
- Experience supporting Hadoop or any other big data platform
- Skills: Ability to debug issues and solve problems
- Working knowledge with Jenkins, Ansible, Terraform, ArgoCD
- Administration of databases (MS SQL, Mongo)
- Working knowledge with Linux and Windows operating system
- You'll have to work on either one of the permanent Shift: 2:00 AM to 11:00 AM IST, Sunday to Thursday.
- Characteristics: Passion to learn and adapt to new technology
- We really value team spirit: Transparency and frequent communication is key. At o9, this is not limited by hierarchy, distance, or function
What we'll do for you :
- Competitive Salary
- Flexibility: Work in Hybrid mode and flexible working hours
- Get social: When we work from home, we play from home with fun after-work activities like Friday Socials. If you're in the office, feel free to join these events in person.
- Flat organization: With a very strong entrepreneurial culture (and no corporate politics).
- Support network: Work with a team you can learn from and every day.
- Diversity: We pride ourselves on our international working environment.
Did you find something suspicious?
Posted by
Posted in
DevOps / SRE
Functional Area
Site Reliability Engineering
Job Code
1674318