At MIGRYX, I build production AWS and Databricks pipelines for large-scale data migration and modernization, processing datasets exceeding 150M records.
I've migrated 150+ legacy SAS files with 100K+ lines of code into scalable PySpark and Spark SQL implementations, preserving business logic and validating results through 100% row-level reconciliation. I re-engineered a repeatedly executed workflow into a single-pass process, reducing runtime from about six hours to roughly 20 minutes.
I work across Delta Lake, Unity Catalog, Medallion Architecture, Snowflake, Informatica IDMC, and Databricks Jobs, with a focus on reusable transformations, data quality, and distributed Spark performance.

