HamburgerMenu
hirist

Python/PySpark Developer - Data Pipeline

SysMind
6 - 14 Years
Multiple Locations

Posted on: 09/06/2026

Job Description

Job Summary:

We are looking for a skilled Python / PySpark Developer building scalable data processing solutions. The ideal candidate should have strong expertise in Python and PySpark, with experience handling large datasets and developing efficient data pipelines.

Key Responsibilities:

- Design, develop, and maintain data pipelines using Python and PySpark

- Process and analyze large-scale structured and unstructured data

- Build and optimize ETL/ELT workflows

- Write efficient and optimized data transformation logic

- Ensure data quality, integrity, and performance

- Collaborate with data engineers, analysts, and stakeholders

- Troubleshoot and resolve data processing issues

- Participate in code reviews and follow best practices

Required Skills:

- Strong proficiency in Python (mandatory)

- Hands-on experience with PySpark / Apache Spark

- Experience in building data pipelines and ETL processes

- Strong SQL knowledge

- Understanding of distributed computing concepts

- Experience with version control tools (Git)

Preferred Skills:

- Experience with cloud platforms (AWS / Azure / GCP)

- Familiarity with big data tools and ecosystems

- Knowledge of data warehousing concepts

- Exposure to orchestration tools (Airflow)

- Basic understanding of Databricks

Soft Skills:

- Strong analytical and problem-solving skills

- Good communication and teamwork abilities

- Ability to work in an agile environment

info-icon

Did you find something suspicious?

Similar jobs that you might be interested in

Loading chat...