I'm currently building and operating Scala/Spark encounter-data processing frameworks at Intellectsoft, supporting regulated Medicaid reporting dependencies. I own ingestion, transformation, batch delivery, Hive schemas, Drools business rules, and production reliability across Cloudera Data Platform environments.
I improve pipeline resilience through idempotent ingestion, schema governance, automated validation, reconciliation logic, and restart-safe processing. I troubleshoot distributed Spark failures and tune skew, serialization, shuffles, joins, and query plans to keep critical downstream batch consumers reliable.
Previously at CreditKey, I designed distributed ingestion and transformation flows for a B2B BNPL platform, processing account, credit, and repayment events for settlement-ready datasets and scoring workflows. I also built maintainable ETL contracts, data models, SQL transformations, operational runbooks, and monitoring-ready production processes.
At Google, I built large-scale analytics and data-processing components, converting raw events into dependable structured datasets. Across my work, I bring 11+ years of experience with Scala, Spark, Hive, SQL, data quality, production operations, and scalable batch systems.
