Skip to main content
PB
Open to opportunities

Prudhvi Bollepalli

@prudhvibollepalli

Senior data engineer building scalable ETL and analytics pipelines across cloud platforms for banking-grade insights.

Zimbabwe
Message

What I'm looking for

I want to keep building secure, scalable ETL and real-time data pipelines on AWS/Azure/GCP, with strong data quality and governance. I’m excited to collaborate in Agile teams and mentor others while delivering reliable analytics for business and risk needs.

I’m a Senior Data Engineer with 5+ years of experience designing and implementing large-scale data pipelines, ETL workflows, and analytics solutions across banking, retail, and financial services. I focus on building scalable data architectures that support business intelligence and data-driven decision-making.

At TD Bank, I designed and developed real-time streaming and batch pipelines using PySpark, Apache Airflow, Apache Kafka, and Spark Streaming. I optimized Spark jobs and SQL queries, reducing data processing time by 45%, and I built dimensional data models in Snowflake for customer analytics, transaction reporting, and credit risk assessment.

I also built automated data quality frameworks using Great Expectations, and I implemented data governance controls like data lineage tracking, metadata management, and PII data masking to ensure regulatory compliance. I migrated legacy ETL processes from on-premise Teradata to AWS cloud, improving scalability and reducing operational costs.

Previously at Capital One, I developed pipelines with Python and PySpark for credit card and banking data, built ETL workflows with Apache Airflow, and supported analytics in Redshift for fraud detection, credit risk, and customer insights. I’m committed to Agile collaboration, production support, mentoring, and delivering secure, well-documented data platforms end-to-end.

Experience

Work history, roles, and key accomplishments

TD Bank logoTB
Current

Senior Data Engineer

Mar 2024 - Present (2 years 4 months)

Designed and developed large-scale PySpark and Apache Airflow data pipelines for banking transactions and financial analytics. Built real-time Kafka/Spark Streaming pipelines, implemented ETL with AWS Glue, and optimized Spark/SQL performance while applying data governance and PII masking.

Capital One logoCO

Data Engineer

Sep 2020 - Aug 2023 (2 years 11 months)

Developed Python and PySpark data pipelines and ETL workflows using Apache Airflow for credit card transactions and customer banking data. Implemented Redshift data models for analytics, built batch processing jobs, and supported production monitoring, security controls, and data governance.

Education

Degrees, certifications, and relevant coursework

University of Memphis logoUM

University of Memphis

Data Science

Pursuing Data Science studies at the University of Memphis, with completion in May 2025.

Get matched with your dream remote job

Sign up now and join over 250,000+ remote workers who receive personalized job alerts, curated job matches, and more for free!

Sign up
Himalayas profile for an example user named Frankie Sullivan