Role Overview
Builds performant data processing pipelines, manages relational databases, and constructs cloud-portable data lake environments using modern open storage formats and embedded analytical engines.
What You Will Do
Construct resilient data processing pipelines in Python using Medallion architecture principles, standardize data storage using open formats, leverage DuckDB for fast and lightweight data processing, and architect and optimize production PostgreSQL schemas.
Why It Might Be a Fit
Requires advanced SQL proficiency, deep PostgreSQL expertise, and working knowledge of SQLite, with a proven track record designing data lakes using the Medallion data processing architecture.
Requirements
- Advanced SQL proficiency with deep PostgreSQL expertise and working knowledge of SQLite.
- Hands-on experience building robust analytical pipelines using Python and DuckDB.
- Mastery of open table and storage formats (Parquet, Apache Iceberg, Delta Lake).
- Proven track record designing data lakes using the Medallion data processing architecture.
- Familiarity with the Databricks / PySpark ecosystem for distributed processing.
To apply for this job please visit wakapi.teamtailor.com.

Follow us on social media