I've built and operated scalable Big Data and Data Platform solutions for Azercell and Prodata, supporting telecommunications and agriculture workloads.
At Azercell, I administer VMware Tanzu Kubernetes clusters and enterprise data applications including Trino and Apache Airflow. I manage CI/CD with GitLab, upgraded Apache Spark from 2.4.7 to 3.3.2, and maintain governance through Apache Ranger.
I design hybrid ML pipelines with AWS SageMaker Pipelines and built an AI agent for Data Engineers using AWS Lambda, EMR, Step Functions, S3, Bedrock, Glue, and Athena.
Earlier, I developed near-real-time telecom CDR streaming pipelines with NiFi, Kafka, and Spark Streaming, optimized Spark and SQL workloads, and built PySpark ETL workflows, dimensional models, and data-quality test suites.
