Familiarity with Python data libraries and processing approaches (e.g., Pandas, batch processing patterns).
Exposure to orchestration/scheduling concepts and operationalizing pipelines for reliability and observability.
Experience working with large datasets and optimizing end-to-end pipeline performance (I/O, SQL tuning, compute efficiency).
Proven ability to collaborate with stakeholders, translate requirements into technical solutions, and deliver within timelines. Good to have skills: Pandas, NumPy, Apache Airflow, Spark (PySpark), Linux/Shell Scripting
A free Jobstore account is required to proceed to the employer site.