At Adobe, I lead Platform Reliability for 400+ Kubernetes clusters across AWS, Azure, and on-premise environments, guiding 35 engineers across four scrum teams and a $25M annual infrastructure budget.
I've delivered more than $1M in annual savings through right-sizing, private endpoint routing, and NAT and egress optimization. I also introduced AI-augmented SRE operations that improved the cluster-to-engineer ratio from 8:1 to 12:1, while automation reduced cluster build time from 18 days to 3 days and support-case lifespan from 35 days to 14 days.
My background spans API Gateway reliability at 50 billion daily calls across nine geographies, secure build infrastructure for Acrobat and Flash Player, enterprise incident operations, and network infrastructure at IBM Canada. I coach managers, build operational excellence mechanisms, and connect technical risk and capacity planning to business priorities.
