About the Role
We are seeking a Python-based platform engineer to design and build a containerized API layer that abstracts and governs interactions with engines such as Apache Spark through a well-defined API contract. This role focuses on building platform capabilities, not simply consuming existing data tools—enabling consistent, secure, and scalable access to Spark and non-Spark based data pipelines.
The ideal candidate has strong experience developing production-grade APIs in Python that interface with data frameworks or pipeline orchestration systems, packaging services using containers (Docker/Kubernetes), and operating them as reusable platform services. A proven background in CI/CD automation using GitHub Actions is required, along with solid software engineering practices around testing, versioning, and deployment.
Responsibilities
Required Skills
Testing
Languages
Frameworks and Platforms