At HSBC, I lead data engineers and own architecture for scalable ETL/ELT workflows supporting multiple business domains.
I've built cloud-native batch and real-time pipelines processing millions of daily records with Snowflake, Databricks, PySpark, AWS Glue, Apache Airflow, Apache Beam, GCP Dataflow, and BigQuery. I optimized 30+ pipelines, reducing end-to-end execution time by 40%.
I also build analytics-ready dbt models, data-quality frameworks, observability standards, and Terraform-based CI/CD deployments. Earlier at Datamatics, I developed NLP data-processing pipelines using AWS Comprehend for medical-record entity recognition.

