Posted on: 19/05/2026
Description :
- Python & PySpark
- Apache Spark & Spark Performance Optimization
- Databricks, Delta Lake & Distributed Data Processing
- GitLab CI/CD Pipelines & Automation
- Databricks REST APIs & SQL Optimization
- Scalable Data Platform Architecture
- Vector Search & Vector-Space Architectures
- AI/ML Workflows, LLMOps & RAG Pipelines
- Data Governance, Metadata & Data Quality
Key Responsibilities :
- Develop and optimize large-scale distributed data pipelines and processing applications
- Perform Spark tuning, query optimization, and troubleshooting for performance improvements
- Build and maintain enterprise-grade CI/CD pipelines using GitLab and automation tools
- Work on Databricks automation, REST APIs, and deployment workflows
- Support advanced AI/ML initiatives including LLMOps, RAG pipelines, vector search, and model evaluation
- Implement data governance, security, metadata management, and quality standards
- Mentor junior engineers and drive engineering best practices across teams
Please share the following details along with your resume :
- Current CTC :
- Expected CTC :
- Experience :
- Notice Period :
Did you find something suspicious?
Posted by
Recruiter
Last Active: NA as recruiter has posted this job through third party tool.
Posted in
Data Engineering
Functional Area
Data Engineering
Job Code
1637034