Posted on: 08/10/2026
Job Description :
The Site Reliability Engineer is responsible for ensuring the availability, performance, and resilience of the organization's digital banking and financial services platforms. This role focuses on automating operational processes, defining and maintaining service-level objectives, and engineering systems that can withstand and recover from failure.
You will work closely with engineering, DevOps, QA, cybersecurity, and compliance teams to ensure platform reliability meets both technical and regulatory standards, while minimizing risk to production systems through proactive monitoring, incident response, and continuous improvement of the software delivery lifecycle.
How You'll Make an Impact :
Reliability Planning & Governance :
- Define and maintain service-level objectives (SLOs), error budgets, and reliability targets aligned with business goals and compliance deadlines.
- Oversee the end-to-end service lifecycle, from code integration to production deployment, with a focus on stability and risk reduction.
- Ensure all changes comply with relevant financial regulations.
- Conduct reliability risk and blast-radius assessments before production changes.
- Coordinate go/no-go decisions with engineering, QA, compliance, and operations stakeholders.
Execution & Delivery :
- Own build, test, and deployment pipelines across multiple environments (staging, UAT, production), ensuring changes are safe, repeatable, and observable.
- Design and maintain automated CI/CD pipelines and enforce version control policies (e.g., Git Flow) to reduce toil and human error.
- Engineer zero-downtime deployments and low-impact change strategies for high-availability systems.
- Develop and maintain rollback, failover, and disaster recovery runbooks for production incidents.
Compliance & Security Oversight :
- Collaborate with Information Security and Compliance teams to validate that infrastructure and deployment practices meet data protection and privacy standards.
- Maintain audit-ready documentation of change activity, incident timelines, and remediation records.
- Support internal and external audits with detailed operational and change history.
Continuous Improvement :
- Drive automation, standardization, and observability improvements across the production environment.
- Conduct post-incident reviews (blameless post-mortems) to identify systemic failures and prevent recurrence.
- Contribute to DevOps and SRE maturity initiatives across engineering teams.
Stakeholder Communication :
- Act as the central liaison between product, development, and compliance teams on production health and change risk.
- Communicate change scope, reliability risks, and incident status clearly to both technical and non-technical stakeholders.
- Provide regular reliability reporting, SLO performance metrics, and incident trends to senior management.
What You Bring :
Technical Proficiency :
- CI/CD tools.
- Cloud platforms (AWS, Azure).
- Containers and orchestration (Docker, Kubernetes).
- Scripting languages (Python, Bash).
- Infrastructure as Code (Terraform, Ansible).
- Observability and monitoring tools.
Soft Skills :
- Strong cross-functional collaboration and communication across engineering and compliance teams.
- Rigorous attention to detail with a proactive approach to risk and failure detection.
- Ability to perform under pressure and respond decisively during incidents and regulatory deadlines.
Education & Experience :
- Bachelor's degree in Computer Science, Information Technology, or related field.
- 5 - 7 years in Site Reliability Engineering, DevOps, or Platform Engineering within financial services or fintech.
- Hands-on experience maintaining reliability for real-time transaction systems, mobile banking, or payment gateways.
- Familiarity with regulatory compliance requirements and their operational implications for production systems.
Our Core Values : What Drives Us :
- Where Everybody is Somebody : We prioritize inclusivity and respect, ensuring that every team member's voice matters. We foster a collaborative environment where everyone can thrive.
- We Say "Yes" When Others Say "No" : Our team is friendly and compassionate, dedicated to delivering an uplifting customer experience. We go the extra mile to provide solutions that truly make a difference, no matter where our customers are on their financial journey.
- We Have a Passion for Our Purpose : We act with integrity and continually seek ways to grow and innovate. With the largest breadth of products in our industry and new offerings in development, we challenge the status quo to better serve our customers.
- We Win as a Team : We believe in the power of teamwork and diversity, celebrating the unique talents of every team member. Hard work is recognized and valued, creating an environment where everyone contributes to our collective success.
Why Join MFSG Technologies India?
Employee Benefits :
- Competitive Employee Benefits.
- Comprehensive Medical Coverage: Includes employee, spouse, two children, and two parents.
- Coverage Amount : 6 lakhs annually.
- 5 lakhs top-up is available if you would like to opt, with thousand rupees per month deduction.
- Provident Fund Contribution: Employer contribution of 1,800 per month to support long-term financial security.
- Performance-Based Incentives: Annual bonus of up to 5% of base salary, based on performance and company success.
Reward & Recognition :
- Bucketlist Rewards Platform : Recognizing outstanding contributions through a point-based rewards system. Redeem points for gift cards, experiences, and personalized perks that matter most to you.
- Culture of Appreciation : Celebrating milestones, achievements, and impact across teams.
Work-Life Balance & Flexibility :
- Flexible Work Model : Currently remote, with a future transition to a hybrid model.
- Paid Leave Entitlements : 15 Annual Leaves, 6 Casual Leaves, 12 Sick Leaves.
Learning & Career Growth :
- Continuous Learning Culture : Access to upskilling programs, mentorship, and professional development resources. Encouragement to learn, adapt, and innovate in a fast-growing financial services landscape.
- Career Growth Opportunities : Internal mobility programs to help employees advance in their careers. Support for leadership development, industry certifications, and specialized training.
- Collaborative Knowledge Sharing : Work with global teams, gaining exposure to international best practices. Encouragement to share ideas, experiment, and take ownership of impactful projects.
Key Requirements :
- AWS (cloud watch, ECS/EKS, EC2, Lambda, VPC, IAM, Route53/DNS and S3)
- Observability and production reliability and cloud migration (on-prem to AWS)
- Terraform and advanced SQL experience with cloud watch.
The job is for:
Did you find something suspicious?
Posted by
Posted in
DevOps / SRE
Functional Area
Site Reliability Engineering
Job Code
1677357