Data Engineer · Diametral
- Designed and built a configuration-driven Python framework for large-scale Oracle data migrations, using declarative YAML and SQL definitions with multiple loading strategies (catalogue, SCD Type 2, fact, aggregate), auditing and concurrency control
- Implemented a control-table-driven scheduler and a reporting module for pipeline observability, orchestrating daily and weekly batch cycles on a Medallion (Bronze/Silver/Gold) architecture with PySpark and Starburst (Trino)
- Monitor production data pipelines and resolve incidents end to end in an enterprise financial-services environment: root cause analysis, permanent fixes and documentation, working directly with DBA and DevOps teams
Python · SQL · PySpark · Oracle · Starburst/Trino · Airflow · GCP · Databricks · IBM Cloud