Posted on: 04/06/2026
Job Title : MLOps Platform Engineer
Location : Remote
Domain : Cloud Infrastructure, MLOps, Platform Engineering, AI Platforms
About ReXruiters :
At ReXruiters, we help companies connect with highly skilled technology professionals through a skills-first hiring approach. Our platform focuses on evaluating real-world capabilities, technical expertise, and project experience to ensure stronger hiring outcomes. We partner with innovative organizations across industries to identify top talent for critical technology and engineering roles.
About the Opportunity :
Our client is building next-generation digital platforms by combining cloud infrastructure, automation, analytics, and modern software delivery practices. They are investing heavily in AI-powered and agentic technologies and are seeking an experienced MLOps / Platform Engineer to help build and scale a reliable, secure, and cost-efficient platform environment.
The ideal candidate will own infrastructure deployment, platform operations, observability, runtime management, and cloud cost governance across the entire AI and agentic platform ecosystem.
Key Responsibilities :
- Build and maintain Infrastructure as Code (IaC) using Terraform and/or AWS CDK.
- Design and implement CI/CD pipelines using GitHub Actions, GitLab CI, AWS CodePipeline, or similar technologies.
- Configure and manage observability solutions using CloudWatch, OpenTelemetry, and related monitoring tools.
- Establish environment isolation strategies and manage prompt, model, and application versioning.
- Deploy and manage containerized workloads using Docker, ECS Fargate, Kubernetes, and EKS.
- Monitor cloud resource utilization and optimize platform costs through governance and budgeting controls.
- Ensure platform reliability, scalability, security, and operational excellence.
- Collaborate with engineering, AI, and product teams to support deployment and lifecycle management of AI-powered applications.
Required Qualifications :
- Bachelor's degree in Computer Science, Information Systems, Engineering, or a related field.
- 5+ years of hands-on experience in AWS platform operations and cloud infrastructure management.
- Strong expertise in Infrastructure as Code (Terraform and/or AWS CDK).
- Experience implementing and maintaining CI/CD pipelines in enterprise environments.
- Deep understanding of AWS services, architecture, and operational best practices.
- Strong experience with monitoring, logging, and observability frameworks.
- Hands-on experience with Docker and container orchestration platforms such as Kubernetes, EKS, or ECS.
- Demonstrated ability to optimize cloud environments for reliability, performance, and cost efficiency.
Preferred Qualifications :
- AWS Solutions Architect Associate or Professional certification.
- Kubernetes or CNCF certifications.
- Experience supporting AI, machine learning, LLM, or agentic application platforms.
- Familiarity with model deployment, versioning, and AI platform operations.
What We're Looking For :
- Strong problem-solving and troubleshooting capabilities.
- Ownership mindset with a focus on operational excellence.
- Ability to work in fast-paced, cloud-native environments.
- Passion for automation, scalability, and platform engineering best practices.
- Commitment to building secure, reliable, and highly available systems.
Benefits :
- Comprehensive health insurance coverage with dependent coverage options.
- Paid time off, holidays, and additional leave benefits.
- Flexible work arrangements supporting work-life balance.
- Ongoing learning and certification support.
- Employee wellness and well-being programs.
- Statutory and retirement benefits as per applicable regulations.
- Inclusive, collaborative, and growth-oriented work environment.
Did you find something suspicious?
Posted by
Posted in
DevOps / SRE
Functional Area
DevOps / Cloud
Job Code
1641522