Santiago venegas
@santiagovenegas
I build reliable data pipelines for customer engagement and financial services.
What I'm looking for
At EPAM supporting Mastercard, I build daily multi-step Impala SQL pipelines for card engagement, customer segmentation, consent and eligibility rules, journey state, and campaign-ready communication exports.
I design Python orchestration with checkpoints, retries, and logging so warehouse extracts and loads can resume without full restarts. I've also migrated engagement business logic from Impala to PostgreSQL, preserving history and exit rules while improving production reliability.
Previously, at Finaipro for Bancolombia, I built Spark, Python, and Apache Sqoop ETL pipelines on Hadoop for large-scale ingestion and daily delivery. My background also includes PySpark ETL, Tableau and Power BI reporting, A/B testing, predictive-model monitoring, Salesforce data quality, web scraping, and SQL automation.
I work closely with marketing, business, agency, and partner stakeholders to reconcile counts, resolve duplicates and consent edge cases, and document logic that keeps customer communications accurate and actionable.
Experience
Work history, roles, and key accomplishments
Built daily multi-step Impala SQL pipelines for card engagement, including segmentation, consent/eligibility rules, and campaign-ready exports. Designed Python orchestration with checkpoints, retries, and logging for reliable warehouse extracts and loads.
Data Engineer
Finaipro
May 2023 - Feb 2024 (9 months)
Built ETL pipelines with Spark, Python, and Apache Sqoop to automate ingestion at scale on the Hadoop stack. Managed large-scale file processing on Hadoop HDFS and deployed scheduled ETL packages for daily execution.
Designed and executed PySpark ETL processes, curating large-scale datasets for analysis. Built Tableau dashboards that turned operational data into clear, actionable views and led A/B testing experiments.
Demand Data Analyst
Foodology
Feb 2022 - Jul 2022 (5 months)
Improved predictive model accuracy in Jupyter using Pandas cleansing and key variable selection. Applied statistical techniques and Power BI reporting to strengthen model monitoring and performance.
Built dashboards that highlighted operational failure points for faster remediation. Implemented Salesforce CRM data quality and cleansing procedures, streamlining data used for decisions.
Marketing Research Expert
Tapclicks
Mar 2019 - Aug 2019 (5 months)
Generated qualified leads via web scraping and SQL querying. Preprocessed data with SQL and automated email generation to cut manual outreach effort.
Education
Degrees, certifications, and relevant coursework
Udemy
Certificate, SQL and Databases
2023 - 2023
Complete SQL and Databases Bootcamp: Zero to Mastery.
Coursera
Specialization, Big Data
2022 - 2022
Specialized Program in NoSQL, Big Data, and Spark Foundations.
Acamica / Digital House
Data Scientist, Data Science
2021 - 2021
Data Scientist program from Acamica / Digital House.
University Foundation CAFAM
Bachelor of Business Administration, Hospitality Business Administration
2013 - 2017
Bachelor in Hospitality Business Administration from University Foundation CAFAM.
Tech stack
Software and tools used professionally
Availability
Location
Authorized to work in
Portfolio
github.com/Codemonster808Job categories
Interested in hiring Santiago?
You can contact Santiago and 90k+ other talented remote workers on Himalayas.
Message SantiagoGet matched with your dream remote job
Sign up now and join over 250,000+ remote workers who receive personalized job alerts, curated job matches, and more for free!
