Posted on: 10/08/2026
Job title : Staff Software Engineer Platform & Distributed Systems
Location : Hyderabad (Hybrid) (Hybrid After 3 Months - 3 Days WFO/2 Days WFH)
Experience : 8+ Years
Notice Period : Immediate to 15 Days
Role Summary :
Enhance an existing AI platform by instrumenting it to capture signals, events, and feedback, enabling continuous learning all while the system remains live and serving customers.
Roles & Responsibilities :
- Design and implement event-driven architecture across APIs and services
- Define scalable event schemas that capture full context and intent
- Build reliable Kafka-based pipelines with replay and failure handling
- Instrument React/Next.js frontend to capture user interactions and feedback
- Ensure end-to-end event flow : UI - Backend - Event Bus - Storage
- Refactor systems using : Strangler pattern, Feature flags, Zero-downtime deployments
- Implement observability (logs, metrics, and traces) from day one
- Migrate historical data into new event schemas
- Provide technical leadership, design reviews, and documentation
Required Skills :
- 8+ years of software engineering experience (Staff-level)
- Strong distributed systems knowledge (CAP theorem, consistency models)
- Proven experience with event-driven systems (Kafka, RabbitMQ, etc.)
- Full-stack expertise : Backend : Python, Frontend : React / Next.js
- Experience transforming live production systems
- Hands-on experience with incremental rollouts and feature flags
Strong Plus (Nice to Have) :
- Observability tools (Open Telemetry, Datadog, Prometheus)
- AI/ML or LLM platform experience
- Event sourcing or CQRS
- High-scale analytics instrumentation
- Monolith-to-microservices migration experience
Tech Stack : Python, MongoDB, PostgreSQL, Redis, Kafka, React, Next.js, Docker, Kubernetes, Next.js, React, PostgreSQL, MongoDB, Redis, Kafka, Python
Did you find something suspicious?
Posted by
Posted in
Full Stack
Functional Area
Backend Development
Job Code
1661974