Posted on: 10/10/2026
Job Description :
ZAGENO is a digital platform focused on simplifying the procurement of life-sciences research products. We are looking for a hands-on Data Engineer with strong PySpark expertise and proven production pipeline ownership to strengthen the reliability, scalability, and correctness of our data platform.
What you will do :
- Design, develop, maintain, and optimize scalable data pipelines using PySpark, Python, and SQL.
- Take primary ownership of production pipelines, from data ingestion and transformation through validation, deployment, and ongoing support.
- Troubleshoot unreliable pipelines, resolve recurring failures, and eliminate fragile implementations and manual workarounds.
- Improve pipeline performance, execution efficiency, reliability, and maintainability.
- Implement incremental processing, efficient transformations, and appropriate data partitioning strategies.
- Build and maintain change data capture (CDC) workflows, low-latency APIs, and data synchronisation processes.
Tech Stack :
Python, SQL, PySpark, Databricks, Delta Lake, Apache Airflow, Change Data Capture (CDC)
Requirements :
- 3+ years of experience in Data Engineering.
- Strong expertise in PySpark and proven production pipeline ownership.
- Ability to independently investigate failures, identify causes, implement lasting fixes, and continuously improve data quality.
- Experience collaborating with cross-functional teams like Analytics, Data Science, and Catalog Operations.
Did you find something suspicious?
Posted by
Posted in
Data Engineering
Functional Area
Data Engineering
Job Code
1677784