Posted on: 02/04/2026
Position Details :
- Experience : 5-10 Years
- Notice Period : Immediate joiners only
- Location : Mumbai, Hyderabad, Pune, Bangalore, Chennai, Noida, Gurugram, Kolkata
- Work Mode : Hybrid (3 days office) - Full-Time (FTE)
As a Generative AI Engineer, you will design, develop, and deploy AI-powered solutions using large language models (LLMs), vector databases, and cloud-based infrastructure. Youll be responsible for building and fine-tuning generative AI models, developing retrieval-augmented generation (RAG) systems, and integrating GenAI capabilities into scalable applications.
Key Responsibilities :
- Develop, fine-tune, and deploy LLM-based applications using frameworks like LangChain, Semantic Kernel, or LlamaIndex.
- Design and implement RAG pipelines using vector databases (Azure Cognitive Search, Pinecone, FAISS, Weaviate).
- Build chatbots, copilots, and content generation tools using OpenAI, Azure OpenAI, or Hugging Face models.
- Containerize and deploy GenAI services using Docker, FastAPI, and Azure Functions.
- Collaborate with cross-functional teams to identify and operationalize GenAI use cases.
- Implement monitoring, evaluation, and prompt optimization workflows.
- Manage prompt engineering, model selection, and performance tracking in production.
Qualifications :
- Bachelors or Masters degree in Computer Science, AI, Data Science, or related field
- 5-10 years of experience in AI/ML with exposure to Generative AI
- Strong experience with LLMs (GPT, LLaMA, etc.)
- Proficiency in Python and frameworks like LangChain, LlamaIndex, or Semantic Kernel
- Experience with vector databases (Pinecone, FAISS, Weaviate, Azure Cognitive Search)
- Understanding of RAG architectures, embeddings, and semantic search
- Experience with cloud platforms (Azure, AWS, or GCP)
- Familiarity with REST APIs, FastAPI/Flask, and microservices
- Experience in Docker and CI/CD pipelines
Did you find something suspicious?