Rajesh Jaiswal
@rajeshjaiswal
AI/ML & data architect building scalable AWS lakehouse and enterprise ETL/ELT platforms for predictive analytics.
What I'm looking for
I’m an AI/ML & Data Architect with 13+ years of experience building and scaling enterprise data platforms across AWS ecosystems, healthcare, technology, manufacturing, and analytics. I focus on production-grade ETL/ELT, lakehouse architecture, petabyte-scale warehousing, predictive analytics, anomaly detection, and cloud migration—always with an eye toward performance engineering and cost optimisation.
I design metadata-driven, event-driven pipelines for batch and near-real-time workloads, and I turn complex business needs into reliable enterprise solutions. My work includes secure FHIR and EDI data processing, building AI-ready data platforms, and implementing data quality and governance so teams can trust downstream KPIs and machine learning outcomes.
In recent roles, I’ve optimized AWS Glue/EMR and orchestration reliability, and built observability with Grafana and alerting across SES/SNS/Lambda/EventBridge. I also launched lakehouse architectures (e.g., S3/Glue/Delta Lake) that improved query performance by 30% and reduced storage cost by 15%, migrated ETL to Dagster for 27% faster throughput, and supported large-scale production operations such as an Airflow platform monitoring 500+ pipelines and a 4.1 PB Redshift warehouse driving 13M annual queries.
Experience
Work history, roles, and key accomplishments
Associate Technical Architect
Impetus Technologies (India) Pvt. Ltd.
Dec 2025 - Present (7 months)
Architect cloud-native AWS data solutions for secure FHIR/EDI processing and design metadata-driven, event-driven pipelines. Optimize Glue/EMR/CloudWatch and build observability and alerting for production workloads.
Launched an AWS S3 and Glue lakehouse using Delta Lake to improve query performance and reduce storage cost. Migrated healthcare ETL to Dagster and led a cross-functional data, BI, and data-science team.
Designed an Apache Airflow platform and onboarded teams from legacy ETL tools, adding monitoring for 500+ production pipelines. Built and supported a 4.1 PB Amazon Redshift warehouse and mentored Data and BI engineers to improve productivity and reduce errors.
Built real-time and daily petabyte-scale pipelines using Scribe, Scuba, Hive, Presto, Spark, and Python. Developed predictive, time-series, perspective, and anomaly-detection KPIs and delivered decision-support visualizations with React JS, Tableau, and Power BI.
Data Analyst
Caterpillar (Aditi Staffing)
Sep 2017 - Jun 2018 (9 months)
Built Python time-series models to improve forecast accuracy and developed a telematics-based vehicle-health model to reduce unplanned maintenance. Automated vehicle data-plan selection, integrated feeds from Vodafone, Verizon, Iridium, and ORBCOMM, and delivered $1.4M in savings.
Delivered finance and supply-chain data solutions using MySQL and Hadoop and implemented SAS-based sentiment analysis and reporting automation. Improved inventory planning accuracy and reduced maintenance cost while increasing productivity and shortening project timelines.
Education
Degrees, certifications, and relevant coursework
University of Illinois Springfield
Master of Science, Computer Science
2016 - 2017
Grade: GPA 3.8
Earned an M.S. in Computer Science from the University of Illinois Springfield (Aug 2016–Jul 2017) with a GPA of 3.8.
University of Pune
Bachelor of Engineering, Computer Engineering
2008 - 2012
Grade: GPA 3.8
Earned a B.E. in Computer Engineering from the University of Pune (Jun 2008–Jun 2012) with a GPA of 3.8.
Tech stack
Software and tools used professionally
Availability
Location
Authorized to work in
Job categories
Skills
Interested in hiring Rajesh?
You can contact Rajesh and 90k+ other talented remote workers on Himalayas.
Message RajeshGet matched with your dream remote job
Sign up now and join over 250,000+ remote workers who receive personalized job alerts, curated job matches, and more for free!
