
Ser Cos
@sercos
I build AI/ML, cloud, governance, and knowledge-graph data pipelines.
What I'm looking for
I've delivered enterprise data governance and Microsoft Fabric integration for Argenx, building pipelines that connect Microsoft Purview and Alation for cataloguing, lineage, and bi-directional metadata consistency across 1,000+ assets.
I build ingestion, quality, and reporting workflows with Python, PySpark, Azure Data Factory, Synapse Analytics, Azure Key Vault, Blob Storage, and Power BI. My work supports secure, scalable processing and stakeholder visibility into metadata coverage, lineage completeness, and data quality.
At GSK, I've worked across Anzo, Neo4j, Azure, and Hortonworks projects, using NLP, semantic technologies, machine learning, and knowledge graphs to support analytics, classification, inference, customer segmentation, safety-stock optimisation, and manufacturing anomaly prediction.
I've also built data lakes, ETL/ELT pipelines, warehouses, APIs, and real-time streaming solutions for SSE and other organisations. I enjoy turning structured, semi-structured, and unstructured data into reliable platforms that teams can use for operational and business decisions.
Experience
Work history, roles, and key accomplishments
AI/ML Data Engineer
London Bridge IT Ltd
Sep 2016 - Present (10 years)
Freelance contractor delivering data engineering and AI/ML solutions across multiple sectors including banking, telco, energy, healthcare, and agronomy.
Enterprise Data Governance & Microsoft Fabric Integration
Argenx
Jan 2024 - Jan 2026 (2 years)
Architected end-to-end data governance pipeline on Microsoft Fabric integrating Microsoft Purview with Alation for enterprise-wide data cataloging and lineage tracking. Built automated data ingestion workflows and implemented Azure-native data pipelines for scalable data processing.
Graph DB Anzo & Neo4j
GSK
Jan 2021 - Jan 2024 (3 years)
Worked on NLP, machine learning, and semantic text mining to translate data into machine understandable representations. Developed end-to-end data pipelines with Python and used Cypher for transforming relational datasets into Neo4j.
Azure Data Engineer
GSK
Jan 2019 - Jan 2021 (2 years)
Worked on the Safety Stock Optimizer project using AI/ML techniques to quantify demand and supply uncertainties. Used Azure Data Factory for end-to-end data pipelines and Azure Databricks for AI/ML model development.
Big Data Engineer
GSK
Jan 2018 - Jan 2019 (1 year)
Worked on AI Prediction Model and Root Cause Analysis project to identify key indicators of anomalies in manufacturing performance. Created holistic data infrastructure and data lake/marts for manufacturing lines.
Data Warehouse Engineer
SSE
Jan 2017 - Jan 2018 (1 year)
Worked on Smart Meter Transformation project and created/maintained optimal Data Lake and data pipeline architecture. Optimized Report Data Warehouse and Data Mart construction for performance, access, and integration.
Education
Degrees, certifications, and relevant coursework
University of London
Bachelor of Science, Computer Science
Pursued higher education in London, United Kingdom.
Tech stack
Software and tools used professionally
AWS Glue
Druid
Dremio
Superset
GitHub
Kubernetes
Jenkins
Jupyter
NumPy
Pandas
PySpark
Dask
DB
Sqoop
MySQL
PostgreSQL
SQLite
Cassandra
Hadoop
HBase
Sybase
Django
Databricks
Git Flow
Neo4j
Terraform
Azure DevOps
Jira
JSON
XML
TensorFlow
PyTorch
MLflow
Keras
Kubeflow
Kafka
Ambari
Zookeeper
Ubuntu
CentOS
Solr
Airflow
Graph Engine
SQL
Azure Cosmos DB
Azure Blob Storage
SciPy
Calendly
Seldon
Cosmos
Microsoft Fabric
Bridge
Factory
Microsoft Purview
X++
Availability
Location
Authorized to work in
Portfolio
github.com/sercostrSalary expectations
Social media
Job categories
Skills
Interested in hiring Ser?
You can contact Ser and 90k+ other talented remote workers on Himalayas.
Message SerGet matched with your dream remote job
Sign up now and join over 250,000+ remote workers who receive personalized job alerts, curated job matches, and more for free!
