Data Engineer β bridging classic data warehousing with the modern lakehouse.
I spent ~5 years deep in the Oracle DWH ecosystem (PL/SQL, ODI, dimensional modeling, SCD) building and optimizing enterprise data pipelines, primarily in the insurance domain. Now I'm carrying that foundation into the modern data engineering stack β CDC, streaming, and lakehouse architecture.
- Data Warehousing Β· Oracle, PL/SQL, ODI, star schema, SCD Type 2, MLOG-based CDC β production ETL/ELT at enterprise scale
- Modern Data Engineering Β· Kafka, Debezium (CDC), Spark Structured Streaming, dbt, Apache Iceberg, Airflow
- Lakehouse Architecture Β· End-to-end pipelines from source CDC to a queryable, ACID lakehouse with BI + ML on top
- Domain Expertise Β· Deep experience in insurance (claims, provisions, benefits, reconciliation)
π Featured Project β ecommerce-realtime-pipeline
A self-hosted, real-time e-commerce lakehouse, built end to end:
Postgres β Debezium (CDC) β Kafka β Spark Structured Streaming β MinIO/Iceberg β dbt β Airflow β Superset
- Real-time CDC with soft-delete handling and LSN-based deduplication
- Medallion architecture (bronze/silver/gold) on Apache Iceberg β ACID, schema evolution, time travel
- ML layer β fraud detection, demand forecasting, customer segmentation, churn prediction (feature store in dbt, orchestrated by Airflow)
- AI access layer β an MCP server enabling natural-language analytics over the lakehouse
Oracle PL/SQL ODI Python SQL dbt Apache Spark Apache Kafka Debezium Apache Iceberg Apache Airflow Superset Qlik Sense PostgreSQL Docker MinIO
Most of my earlier work lived in enterprise Oracle/ODI repositories β my recent GitHub activity reflects my move into the open, modern data stack.
LinkedIn: https://linkedin.com/in/deniz-isik-ofc/
E-mail: denizsk977@gmail.com