At IBM, I support enterprise Kubernetes and OpenShift platforms across development through production, including cluster upgrades, operator-health validation, incident resolution, and root cause analysis.
I built a single Ansible playbook to deploy the MongoDB Atlas–Instana integration across six production environments in one run, eliminating manual per-environment setup. I also build and maintain Grafana and Prometheus dashboards, SSL-expiry monitoring, and automation that helps detect platform issues before production impact.
I troubleshoot Kafka data-ingestion failures and AWS Redshift query-performance issues affecting reporting pipelines, while driving ElastiCache compliance remediation and Terraform-based infrastructure updates with cloud and architecture teams.
During my IBM SRE internship, I onboarded Guardium Data Security Center into IBM Concert, automated onboarding workflows, and contributed to early Grafana, Prometheus, and Ansible work. I enjoy building secure, repeatable operational automation with Python, Shell scripting, Docker, Kubernetes, and least-privilege RBAC.
