I've built and automated data pipelines and platforms for Alfa-Bank, Sberbank, JSC NIIAS, and Russian Railways, working across banking, fintech, and transportation.
At Alfa-Bank, I developed a Python, watchdog, and Apache Airflow file orchestrator with four DAGs for all 10 regulatory reports. I also reverse-engineered Greenplum, ClickHouse, and MS SQL sources, prepared CDC integration requirements for a Debezium → Kafka → Flink architecture, and built four SAP OData parsers for the DWH team.
At Sberbank, I optimized PySpark ETL on YARN by addressing data skew and redesigned customer identification logic, eliminating 180,000 duplicate or invalid records. I also built an incremental SCD Type 2 pipeline and maintained Airflow DAGs supporting a customer-verification product for more than 4 million customers.
Earlier, I built a PostgreSQL DWH and Python ETL pipelines at JSC NIIAS, eliminating manual Excel processing and reducing legacy dashboard load times from 45–60 seconds to 8–10 seconds. At Russian Railways, I automated production file deployment and PDF data extraction, turning hours of schedule preparation into 3–10 minutes.

