At QuestionPro, I raised platform availability from 99.5% to 99.9% by engineering fault tolerance into the application tier and horizontally scaling a bottlenecked on-premises database.
I designed a fault-tolerant, containerized DNS platform across multiple data centers and architected Azure networking with Terraform. I also consolidated the L7 load-balancing layer into a templated Nginx tier.
I own reliability, monitoring, and incident response for business-critical platforms, serving as the L3 escalation point and leading incidents through root cause analysis to long-term corrective actions. Earlier, at Tata Consultancy Services, I managed Linux environments and automated recurring operational tasks with shell scripts.

