I engineered a maritime AIS and weather pipeline for my MSc Big Data Project, using Python and GeoPandas to combine vessel trajectories with meteorological observations.
I deployed the pipeline with Docker and MongoDB, then used geospatial indexing and latency benchmarks to assess query performance.
For a distributed search project, I built a PySpark Locality-Sensitive Hashing pipeline for high-dimensional similarity search and benchmarked its execution and recall across configurations.
I also designed an Airflow project that ingests data from REST APIs, transforms it into Parquet, and integrates AWS S3 with PySpark. Earlier, as a Logistics & Inventory Assistant at Ydrothermiki Prevezas, I improved tracking workflows and audited data integrity.

