HamburgerMenu
hirist

Senior Cloud & DevOps Engineer - AI Platform Operations

Optegral Advisory Services
5 - 12 Years
Bhubaneshwar

Posted on: 26/06/2026

Job Description

Senior Cloud & DevOps Engineer AI Platform Operations

Job Description :

We are looking for a Senior Cloud & DevOps Engineer AI Platform Operations to join our client in the AI and decision intelligence space.

This role is ideal for a senior AWS cloud, DevOps, SRE, or platform engineering professional who can design, deploy, monitor, and support production-grade AI and data applications. The selected candidate will own cloud infrastructure, deployment automation, monitoring, incident response, and operational readiness for AI applications built using AWS, Claude/Bedrock, FastAPI, PostgreSQL, S3, and related technologies.

This is not a traditional systems administration role. The role requires hands-on cloud engineering, DevOps discipline, production support maturity, and the ability to manage junior systems engineers in an 18x7 support model.

Key responsibilities include :

- Design, deploy, and manage AWS infrastructure for AI, data, and application workloads.

- Set up and manage services such as EC2/ECS/App Runner, RDS PostgreSQL/Aurora PostgreSQL, S3, IAM, VPC, Secrets Manager, KMS, CloudWatch, and related AWS services.

- Support deployment of FastAPI, Streamlit, background workers, and AI application services.

- Configure secure access to Amazon Bedrock / Claude and related AI platform services.

- Build and maintain CI/CD pipelines for reliable application deployment.

- Implement monitoring, logging, alerting, backup, recovery, and incident-response processes.

- Create and maintain runbooks, SOPs, deployment checklists, and support documentation.

- Lead L1/L2 operational support and coordinate L3 escalation to AI engineers, data engineers, and AI scientists.

- Manage, guide, and review the work of junior systems engineers.

- Ensure infrastructure is secure, auditable, cost-aware, and suitable for financial-services workloads.

- Participate in client-facing technical discussions, status reviews, and incident reviews as required.

Requirements :

Required :

- 7+ years of experience in cloud engineering, DevOps, SRE, platform engineering, or production systems operations.

- Strong hands-on experience with AWS.

- Experience with core AWS services such as EC2, ECS/Fargate, App Runner, RDS PostgreSQL/Aurora PostgreSQL, S3, IAM, VPC, Security Groups, Secrets Manager, KMS, CloudWatch, and CloudTrail.

- Experience deploying and supporting Python-based applications, APIs, containers, and web services.

- Strong understanding of Docker, containerized deployments, CI/CD pipelines, and release management.

- Experience with Linux administration, shell scripting, networking basics, log analysis, and troubleshooting.

- Experience supporting production systems with monitoring, alerting, incident management, backup, recovery, and root-cause analysis.

- Working knowledge of PostgreSQL operations, connectivity, backup/restore, and performance monitoring.

- Ability to lead junior engineers and operate in an 18x7 support model.

- Strong documentation skills for run books, SOPs, incident reports, and deployment notes.

- Good communication skills and ability to work with engineering, data, AI, and client teams.

- Willingness to work from the Bhubaneswar office.

Preferred :

- Experience with Amazon Bedrock, Claude, OpenAI, or other enterprise LLM platforms.

- Experience supporting AI/ML, data science, analytics, or SaaS applications.

- Experience with FastAPI, Streamlit, Celery, Redis, or Python application stacks.

- Experience with Terraform, CloudFormation, or infrastructure-as-code tools.

- Experience with GitHub Actions, GitLab CI/CD, Jenkins, or similar tools.

- Exposure to LangGraph, LangChain, LLMOps, model-serving, or AI application observability.

- Experience with CloudWatch dashboards, centralized logging, alerting, and cost monitoring.

- Experience in fintech, banking, lending, regulated industries, or client-facing managed services.

- AWS certifications such as AWS Solutions Architect, AWS SysOps Administrator, AWS DevOps Engineer, or equivalent.

Benefits :

- Competitive compensation aligned with market standards, based on experience and capability.

- Opportunity to lead cloud and DevOps operations for a high-impact AI underwriting programme.

- Exposure to AWS, Claude/Bedrock, AI platform operations, financial-services technology, and production AI systems.

- Opportunity to manage and mentor junior systems engineers.

- Work on practical, governed AI systems rather than generic proof-of-concept demos.

- Collaborate with experienced AI, data science, data engineering, and platform professionals.

- Gain experience in 18x7 support operations, client-facing delivery, and AI platform reliability.

- Candidates from other cities are welcome to apply if they are open to relocating to Bhubaneswar.

info-icon

Did you find something suspicious?

Similar jobs that you might be interested in

Loading chat...