HamburgerMenu
hirist

Data Architect - Databricks/AWS

Haparz
8 - 12 Years
Multiple Locations

Posted on: 03/09/2026

Job Description

Role : Data Architect - Databricks / AWS

Job Summary :

We are looking for an experienced Data Architect to define and drive the target data architecture for the Horizon MVP and its future evolution. The role will be responsible for designing a scalable, secure, governed cloud data platform covering ingestion, storage, processing, analytics, APIs, and downstream data consumption.

The architect will work closely with Data Engineering, Backend, DevOps, QA, and business stakeholders to establish architecture standards and ensure the platform is ready for advanced analytics, AI/ML, vector storage, and future LLM-based capabilities.

Role Details :

- Experience : 8+ Years

- Relevant Architecture Experience : 3+ Years in Data Architecture

- Location : Chennai / Pune

- Work Mode : Hybrid - 3 Days WFO

- Payroll : Haparz

- Notice Period : Immediate / Short Notice Preferred

Key Responsibilities :

- Define the target data architecture for the Horizon MVP and establish an architecture roadmap for future scalability.

- Design end-to-end architecture covering data ingestion, storage, processing, serving, reporting, APIs, and downstream applications.

- Establish canonical data models and schemas for travel signals, corridors, sources, evidence, scores, and generated insights.

- Define data normalization strategies for structured, semi-structured, and unstructured data from multiple external sources.

- Design and govern the Databricks platform architecture, including Unity Catalog, data schemas, access controls, and governance standards.

- Establish data-retention, lineage, data-quality, security, privacy, and compliance controls.

- Define secure integration patterns between Databricks, AWS PRODIGY, SharePoint, external APIs, and downstream applications.

- Design scalable data processing for 30-day signal windows, convergence/divergence scoring, corridor ranking, and spike detection.

- Define architecture patterns that support future vector storage, embeddings, LLM integration, and multi-year analytics.

- Design reliable batch and API-driven ingestion frameworks for structured and unstructured data.

- Review technical designs, identify architectural risks, and provide technical direction to engineering teams.

- Guide backend, data engineering, DevOps, and QA teams in implementing architecture standards.

- Ensure architecture decisions align with enterprise security, RBAC, PII handling, privacy, and operational requirements.

- Communicate architecture decisions, trade-offs, and technical recommendations effectively to technical and business stakeholders.

What We're Looking For :

- 8+ years of experience in data engineering, data platforms, or data architecture, with at least 3+ years in a Data Architect capacity.

- Strong hands-on experience designing cloud-based data platforms, lakehouses, or analytical platforms.

- Advanced knowledge of Databricks, Apache Spark/PySpark, Delta Lake, and Unity Catalog.

- Strong understanding of AWS data services, IAM, networking, and secure cloud integration patterns.

- Strong expertise in data modelling, metadata management, data lineage, data quality, retention, and governance.

- Experience architecting batch and API-based ingestion pipelines for structured, semi-structured, and unstructured data.

- Understanding of AI/ML workloads, feature pipelines, vector databases, embeddings, and LLM integration patterns.

- Experience designing APIs and downstream data-serving architectures.

- Strong knowledge of PII protection, RBAC, data privacy, and enterprise security controls.

- Excellent architectural communication and stakeholder-management skills.

info-icon

Did you find something suspicious?

Similar jobs that you might be interested in

Loading chat...