At Compass Mining, I designed a log-based diagnostics system that maps device log patterns and vendor API errors to plain-language fault diagnoses. I also built the fleet incident classification model, covering 18 issue types across six priority levels.
As Monitoring Team Lead, I led 24/7 monitoring and incident response for 50,000 devices across 18 sites, with alert-to-action response time at two minutes. I introduced structured incident reviews and runbook updates, and built real-time site-level status monitoring from device telemetry.
Earlier at Compass Mining, I built an internal monitoring platform, automated remote reboot logic, and streamlined downtime credit calculations and customer reporting. At WATTUM, I remotely troubleshot ASIC and GPU fleets, deployed equipment, and led a small remote team.
I run a self-built homelab with a three-node Proxmox VE cluster and a Kubernetes cluster provisioned with Terraform and Ansible. I’m a Certified Kubernetes Administrator (CKA), and I document the lab with network diagrams and a hardware breakdown.

