At Bluetooth SIG, I architect and implement PySpark pipelines processing 2+ TB of data daily, and design data models for Microsoft Fabric Lakehouse and Warehouse workloads. I also develop shared semantic models and Power BI reports, and support migration of Hadoop HDFS data assets and reports to Microsoft Fabric.
Previously, at Infonomics, I led the migration of 4PB+ of data from a legacy Oracle database to big data platforms and built pipelines with PySpark, Spark SQL, Hive, and Airflow. Across my work in data engineering and architecture, I’ve developed cloud data platforms, ETL workflows, and analytics solutions on Azure, AWS, and GCP.

