Posted on: 18/09/2026
Role Overview :
Join a high-impact team building scalable AI systems that power conversational interfaces, speech recognition, and large-scale data transformation initiatives.
What You'll Do :
- Architect and operate end-to-end AI systems for conversational interfaces and speech recognition.
- Drive accuracy and cost trade-offs with ground-truth metrics.
- Build and optimize self-hosted serving (vLLM, Triton class) for model serving and infrastructure.
- Own latency, throughput, and cost per user.
Core Technical Requirements :
- 4 to 9 years as an AI/ML engineer with production systems behind you.
- Depth across the modern AI stack : LLMs, speech models, vector retrieval, model serving.
- Strong software engineering skills.
- Fluency in Python and the production ML ecosystem.
Did you find something suspicious?