Posted on: 17/08/2026
Role Overview :
As a Founding Product Architect (Infrastructure) for our Enterprise GenAI platform, you will serve as the technical backbone for our next-generation AI initiatives. You will be responsible for designing and scaling the underlying infrastructure that powers complex LLM applications, ensuring high availability, low latency, and robust security for enterprise-grade deployments. Working closely with the founding team, data scientists, and product managers, you will bridge the gap between raw AI research and production-ready software. Your work will directly dictate the performance and reliability of our AI agents, enabling our clients to automate critical business workflows with confidence and precision. This role offers a unique opportunity to shape the technical roadmap of a high-growth startup from the ground up, operating in a flexible Hyderabad-based or remote environment.
Key Responsibilities :
- Architect and deploy scalable cloud infrastructure that supports high-throughput LLM inference and complex RAG pipelines to ensure seamless performance for enterprise users.
- Design and maintain high-performance VectorDB clusters and data ingestion layers to enable real-time information retrieval and context-aware AI responses.
- Develop and optimize backend services using Python and FastAPI to create reliable APIs that serve as the interface for our Agentic AI frameworks.
- Implement robust monitoring, logging, and security protocols across the IT infrastructure to protect sensitive enterprise data and ensure compliance with industry standards.
- Collaborate with cross-functional teams to translate business requirements into modular system architectures that can evolve alongside rapidly advancing GenAI technologies.
- Lead the technical strategy for cloud computing resource allocation to balance computational efficiency with the cost-intensive nature of large-scale model deployment.
Required Skillset :
- Demonstrated expertise in designing and managing complex cloud infrastructure and enterprise-grade system architectures that handle high-concurrency workloads.
- Proven ability to build and maintain production-ready RAG systems and VectorDB implementations, ensuring high accuracy and retrieval speed for LLM-driven applications.
- Strong proficiency in Python and FastAPI for developing scalable, maintainable, and high-performance backend services.
- Deep understanding of Agentic AI workflows and the ability to integrate autonomous agents into existing enterprise IT ecosystems.
- Exceptional communication skills, with the ability to articulate complex technical trade-offs to non-technical stakeholders and lead technical discussions with clarity.
- A proactive, problem-solving mindset with the ability to thrive in a remote or hybrid work environment, demonstrating high levels of autonomy and ownership.
- A strong academic background in Computer Science or a related field, complemented by 7 to 10 years of hands-on experience in building and scaling enterprise software systems.
Did you find something suspicious?