Skip to main content
PS
Open to opportunities

Pooja Suvarna

@poojasuvarna

I’m a Senior Data Engineer building scalable ETL/ELT pipelines and cloud data platforms for performance, reliability, and cost efficiency.

India
Message

What I'm looking for

I’m looking for a data engineering role where I can build and optimize scalable AWS-based ETL/ELT pipelines, improve distributed Spark/Hive performance, and deliver reliable warehouse and migration solutions with a strong focus on cost and quality.

I’m a Data Engineer with 6 years of experience in Big Data Engineering, ETL/ELT pipeline development, and cloud data platform design. I focus on building scalable data pipelines, data warehousing solutions, and data migration workflows that perform well under real-world load.

In my recent work at Carelon Global Solutions, I designed and developed batch processing data pipelines on Amazon EMR using Apache Spark and Python. I built AWS Glue ETL jobs with PySpark to move data from Amazon S3 into curated datasets, leveraging the AWS Glue Data Catalog for metadata and schema discovery.

I’ve also strengthened reliability and speed by optimizing distributed processing—tuning Spark configurations, partitioning strategies, shuffle behavior, and memory allocation to resolve bottlenecks. I’ve implemented Snowflake bulk-loading with data validation, access control, and performance tuning, and automated ingestion and orchestration with robust error handling, logging, and reusable modules.

Earlier, at LKQ India Private Limited and Bullfinch Software, I worked across end-to-end integration and migration, using Sqoop for large-scale migrations to Hadoop clusters and Hive/Spark for efficient querying and warehousing. I’ve consistently applied performance tuning techniques and workflow orchestration (including AWS Step Functions and Airflow) to reduce overhead and deliver measurable cost savings.

Experience

Work history, roles, and key accomplishments

CS
Current

Data Engineer

Carelon Global Solutions

Oct 2025 - Present (10 months)

Designed and developed batch processing ETL/ELT data pipelines on AWS EMR using Apache Spark and Python, including AWS Glue jobs and data loading into Snowflake. Optimized Spark performance and automated data ingestion, reporting, and workflow orchestration with monitoring and alerting.

LL

Data Engineer

LKQ India Private Limited

Nov 2023 - Oct 2025 (1 year 11 months)

Built PySpark and Python data pipelines on AWS EMR for processing S3 data and orchestrated parallel job execution using AWS Step Functions. Implemented end-to-end data integration and migration using Sqoop and optimized querying with Hive features and schema evolution.

Education

Degrees, certifications, and relevant coursework

KS

KVGCE, Sullia

Bachelor of Engineering, Computer Science and Engineering

2016 - 2020

Grade: CGPA: 8.21

Bachelor of Engineering in Computer Science and Engineering from KVGCE, Sullia (2016 to 2020).

Get matched with your dream remote job

Sign up now and join over 250,000+ remote workers who receive personalized job alerts, curated job matches, and more for free!

Sign up
Himalayas profile for an example user named Frankie Sullivan