uild Reliable Ingestions: Build and maintain robust data ingestion scripts to pull data from APIs, servers, or files into raw/staging tables. You ensure correct handling of retries, incremental/idempotent loads, and graceful error management.
Model with dbt: Grow our analytics-ready data layer by building and maintaining clean staging and mart models with clear naming, documentation, and automated tests, focusing on simplicity and readability before optimisation.
Support Orchestration: Extend and troubleshoot our existing workflow orchestration frameworks (Airflow). You will add tasks to existing DAGs, fix failing runs, and grow into mastering dependency, scheduling, and retry behaviour.
Champion Data Quality: Own pipeline and model health by adding and monitoring data quality checks (freshness, volume anomalies, null/duplicate validations) and investigating data discrepancies flagged by different stakeholders.
Document & Accelerate: Write clean code, maintain useful READMEs, and leverage AI-assisted development tools to speed up drafting while taking full ownership of verifying correctness before pushing to production.
Grow your ownership: Build a solid track record with data modelling, testing, and pipeline reliability as a foundation for growing into orchestration design, platform topics, and broader engineering responsibilities over time.