Skip to main content
SV
Open to opportunities

Santiago venegas

@santiagovenegas

I build reliable data pipelines for customer engagement and financial services.

Colombia
Message

What I'm looking for

I'm looking to build reliable, scalable data pipelines and improve data quality, governance, and customer-facing analytics. I value collaborating with business and marketing stakeholders to translate complex data work into accurate, actionable outcomes.

At EPAM supporting Mastercard, I build daily multi-step Impala SQL pipelines for card engagement, customer segmentation, consent and eligibility rules, journey state, and campaign-ready communication exports.

I design Python orchestration with checkpoints, retries, and logging so warehouse extracts and loads can resume without full restarts. I've also migrated engagement business logic from Impala to PostgreSQL, preserving history and exit rules while improving production reliability.

Previously, at Finaipro for Bancolombia, I built Spark, Python, and Apache Sqoop ETL pipelines on Hadoop for large-scale ingestion and daily delivery. My background also includes PySpark ETL, Tableau and Power BI reporting, A/B testing, predictive-model monitoring, Salesforce data quality, web scraping, and SQL automation.

I work closely with marketing, business, agency, and partner stakeholders to reconcile counts, resolve duplicates and consent edge cases, and document logic that keeps customer communications accurate and actionable.

Experience

Work history, roles, and key accomplishments

EP
Current

Data Engineer

Mar 2024 - Present (2 years 5 months)

Built daily multi-step Impala SQL pipelines for card engagement, including segmentation, consent/eligibility rules, and campaign-ready exports. Designed Python orchestration with checkpoints, retries, and logging for reliable warehouse extracts and loads.

FI

Data Engineer

Finaipro

May 2023 - Feb 2024 (9 months)

Built ETL pipelines with Spark, Python, and Apache Sqoop to automate ingestion at scale on the Hadoop stack. Managed large-scale file processing on Hadoop HDFS and deployed scheduled ETL packages for daily execution.

iFood logoIF

Data Analyst

Aug 2022 - Nov 2022 (3 months)

Designed and executed PySpark ETL processes, curating large-scale datasets for analysis. Built Tableau dashboards that turned operational data into clear, actionable views and led A/B testing experiments.

FO

Demand Data Analyst

Foodology

Feb 2022 - Jul 2022 (5 months)

Improved predictive model accuracy in Jupyter using Pandas cleansing and key variable selection. Applied statistical techniques and Power BI reporting to strengthen model monitoring and performance.

SunRun logoSU

Back Office Data Analyst

Apr 2020 - Nov 2021 (1 year 7 months)

Built dashboards that highlighted operational failure points for faster remediation. Implemented Salesforce CRM data quality and cleansing procedures, streamlining data used for decisions.

TA

Marketing Research Expert

Tapclicks

Mar 2019 - Aug 2019 (5 months)

Generated qualified leads via web scraping and SQL querying. Preprocessed data with SQL and automated email generation to cut manual outreach effort.

Education

Degrees, certifications, and relevant coursework

Udemy logoUD

Udemy

Certificate, SQL and Databases

2023 - 2023

Complete SQL and Databases Bootcamp: Zero to Mastery.

Coursera logoCO

Coursera

Specialization, Big Data

2022 - 2022

Specialized Program in NoSQL, Big Data, and Spark Foundations.

AH

Acamica / Digital House

Data Scientist, Data Science

2021 - 2021

Data Scientist program from Acamica / Digital House.

UC

University Foundation CAFAM

Bachelor of Business Administration, Hospitality Business Administration

2013 - 2017

Bachelor in Hospitality Business Administration from University Foundation CAFAM.

Get matched with your dream remote job

Sign up now and join over 250,000+ remote workers who receive personalized job alerts, curated job matches, and more for free!

Sign up
Himalayas profile for an example user named Frankie Sullivan