Skip to main content
ankit vishwakarmaAV
Looking for a job

ankit vishwakarma

@ankitvishwakarma3

I build reliable cloud data pipelines for analytics and identity products.

India
Message

What I'm looking for

I'm looking to continue owning production-grade cloud data pipelines, collaborating across analytics, engineering, and client teams, and applying modern cloud architectures and generative AI tooling to improve data workflows.

At Publicis Sapient, I build automated streaming and batch deduplication frameworks that prevent redundant records from reaching identity-generation workflows and protect downstream identity graph accuracy. I also consolidate, cleanse, reconcile, and deduplicate data from 15 upstream tables for cross-product attribution.

Previously at LTIMindtree and IHS Markit, I automated survey and analytics data ingestion into ADLS, built API pipelines for Mixpanel events, and delivered PySpark migration pipelines from Hadoop to AWS S3. I've also automated reconciliation and PGP encryption workflows through scheduled EMR jobs.

My work spans PySpark, Spark SQL, Azure Databricks, ADLS, AWS, Hive, Airflow, and SAS. I bring consistent ownership of production-grade ETL pipelines while collaborating across analytics, engineering, and client teams.

Experience

Work history, roles, and key accomplishments

PS
Current

Senior Data Engineer

Dec 2023 - Present (2 years 9 months)

Engineered an automated streaming/batch deduplication framework to eliminate redundant file records before transmitting identity generation keys to IDM on a 15-minute cadence. Extracted, cleansed, and reconciled heterogeneous data across 15 upstream tables to generate a consolidated, attribution-ready unified table.

LT

Data Engineer

Jun 2022 - Dec 2023 (1 year 6 months)

Automated store-level customer experience survey data ingestion from client SFTP servers into Azure Data Lake Storage (ADLS), implementing schema mapping and validation. Architected automated API ingestion fetching analytics event payloads from Mixpanel REST endpoints into ADLS and executed transformations to match existing schema structures.

IHS Markit logoIM

Sr. Data Analyst (Automotive)

IHS Markit

Jan 2019 - Jun 2022 (3 years 5 months)

Delivered an integrated data pipeline using PySpark, ingesting five legacy source datasets from Hadoop, performing transformations and data cleansing, and outputting consolidated results to AWS S3. Constructed automated pipeline security routines to encrypt outbound deliverables into PGP format using Python cryptographic libraries, eliminating manual bottlenecks via scheduled EMR jobs.

Conduent logoCO

Associate Business Analyst

Jun 2017 - Apr 2019 (1 year 10 months)

Processed multi-source client datasets across varying raw structures, handling missing value imputation, identifying duplicate records, executing business cleansing rules, and staging final datasets within SAS repositories.

Education

Degrees, certifications, and relevant coursework

CSJM University logoCU

CSJM University

Bachelor of Computer Applications, Computer Applications

Completed a Bachelor of Computer Applications degree in 2013.

CC

Christ Church Inter College

Intermediate, General Studies

Completed Intermediate (Class XII) education.

Get matched with your dream remote job

Sign up now and join over 250,000+ remote workers who receive personalized job alerts, curated job matches, and more for free!

Sign up
Himalayas profile for an example user named Frankie Sullivan