HamburgerMenu
hirist

Job Description

Position : AI Solution Architect

Company Description :


NStarX is a practitioner-led, AI-first, and Cloud-first engineering services provider with a deep-rooted DNA in AI Engineering and Cloud-Native technologies. We specialize in delivering transformative business value through innovative technology integration and strong execution capabilities in software engineering.


With over a decade of industry expertise, we have supported numerous organizations in their digital transformation journeys, enabling them to achieve accelerated and sustainable outcomes.

Job Description :

We are looking for a highly skilled AI Architect with deep expertise in Generative AI, LLMs, Video Models (Digital Humans / Avatars), and end-to-end AI product architecture.

This role requires a hands-on technologist and strategic thinker who can design scalable AI systems, guide development teams, interact with global clients, and drive high-impact AI initiatives across cloud, on-prem GPU servers, and edge devices (AI PCs).

You will architect solutions that span the spectrum-from H200-class GPU compute for very large LLM workloads to lightweight, optimized models that run efficiently on edge devices.

This is a senior, high-ownership role for someone passionate about building real-world AI products at scale.

1. AI & GenAI Architecture :

- Design and architect LLM-based systems using both open-source (Llama, Mistral, etc.) and proprietary (OpenAI, Azure OpenAI, Anthropic, etc.) models.

- Architect Video-based AI systems, including Digital Human Avatars, Video Gen/Video2Text, and multimodal pipelines.

- Build end-to-end GenAI pipelines including data ingestion, preprocessing, retrieval, fine-tuning (LoRA/QLoRA/DAPT), evaluation, guardrailing, and deployment.

2. Core ML & Data Engineering :

- Define and orchestrate data pipelines, ML workflows, vector search architecture, and embedding strategies.

- Build scalable, secure ML engineering wrappers around models (e.g., inference servers, orchestration layers, API microservices).

- Oversee experimentation frameworks, evaluation methodologies, and MLOps integration.

3. Cloud & Infrastructure :

- Architect AI solutions on AWS and Azure (preferred), including GPU clusters, model hosting, DevOps/MLOps, and autoscaling.

- Work with Nvidia GPU server stacks (DGX/H200/H100/L40S) and edge AI systems (Intel/AMD/Qualcomm AI PCs).

- Optimize AI workloads across heterogeneous compute environments.

4. Product & Delivery Ownership :

- Lead AI architecture across POC - MVP - GA - Production-scale phases.

- Contribute to roadmap planning, feasibility analysis, and technical risk assessment.

- Ensure performance, scalability, cost efficiency, and robustness of AI products.

5. Governance, Security & Compliance :

- Embed data privacy, security controls, Responsible AI, and governance frameworks into product design.

- Ensure adherence to enterprise AI policies, guardrails, and regulatory requirements.

6. Client Engagement & Communication :

- Interact with global clients (North America & Europe) to understand requirements, present architectures, and provide expert guidance.

- Create clear architecture diagrams, documentation, and high-quality technical specifications for developers and stakeholders.

- Serve as the technical face of the project in client discussions.

7. Leadership & Collaboration :

- Collaborate with AI Engineers, Data Scientists, Product Owners, Cloud Architects, and MLOps teams.

- Mentor teams in AI design patterns, best practices, and solution development.

- Conduct architecture reviews, code/design audits, and knowledge-sharing sessions.

Required Skills & Qualifications Technical Skills :

- Deep expertise in Generative AI : LLMs, Vision/Video models, Digital Avatars, RAG systems, multimodal architectures.

- Strong experience in ML engineering, data pipelines, and scalable model APIs.

- Hands-on experience with Nvidia GPU systems, CUDA stack, TensorRT, VLLM/Ollama, and model optimization.

- Experience building AI on edge devices (Intel/AMD/Qualcomm NPUs, AI PCs).

- Proficient in AWS/Azure cloud ecosystems, including GPU-based deployments.

- Strong hold on Python, ML frameworks (PyTorch/TensorFlow), model serving frameworks, and MLOps tools.

Professional Requirements :

- Minimum 10 years of experience in ML/AI solution architecture.

- Proven track record of architecting POC/MVP/Production AI products.

- Strong architectural documentation and diagramming skills (Mermaid, Draw.io, Lucidchart, ArchiMate, etc.).

- Excellent communication skills for client presentations and internal leadership discussions.

- Ability to work in a fast-paced, multi-project environment across global teams.

Preferred Qualifications :

- Graduation from a Tier-1 institute (IIT/NIT/IIIT or equivalent).

- Certifications in AI/ML Architecture, Solution Architecture, or Cloud Architecture (AWS/Azure).

- Experience with enterprise AI governance and generative AI compliance frameworks.

Why Join Us?

- Work on cutting-edge GenAI products involving the latest LLMs, avatars, multimodal AI, and GPU/edge architectures.

- Collaborate with global teams and enterprise clients across North America and Europe.

- Shape the AI strategy of high-impact products used at scale.

- Join a passionate AI engineering ecosystem with strong innovation culture.

The job is for:

Women candidates preferred
Differently-abled candidates preferred
For women joining back the workforce
info-icon

Did you find something suspicious?

Similar jobs that you might be interested in

Loading chat...