Skip to main content
Kaleb ClementKC
Open to opportunities

Kaleb Clement

@kalebclement

Site Reliability Engineer focused on scalable observability, incident automation, and multi-cloud reliability.

Indonesia
Message

What I'm looking for

I’m looking for a Site Reliability/Cloud role where I can build observability and automation that prevent incidents, improve SLO/SLA outcomes, and enable safe multi-cloud migrations—working with engineering teams to raise reliability through measurable operational excellence.

I’m a Site Reliability Engineer building and operating critical middleware and reliability systems at ByteDance, where I serve as the primary maintainer for 14 core platforms and lead 24/7 on-call operations and cross-functional incident responses. I delivered 99.99% reliability while architecting major improvements to monitoring, stability, and migration safety.

I built Diary Platform, an internal SRE observability and service-metadata single source of truth for ~600 services, and developed Openclaw to automate root-cause analysis—reducing engineering workload by 50–70% and cutting triage time by nearly 70%. I’ve also driven zero-downtime migrations of 60+ services across GCP, AWS, and Alibaba Cloud, and improved job scheduling stability by eliminating ~100,000 monthly failures and reducing MTTR by 25%.

Experience

Work history, roles, and key accomplishments

ByteDance logoBY
Current

Global E-Commerce Service Architecture

Feb 2024 - Present (2 years 5 months)

- Serve as the primary maintainer for 14 core middleware platforms (including storage, caching, and scheduling), delivering 99.99% reliability by leading 24/7 on-call operations and major cross-functional incident responses for Global E-commerce teams distributed across the US, Asia, Europe, and the Rest of the World.

- Architected and built Diary Platform, an internal SRE observability and serv

Tokopedia logoTO

Cloud Platform Engineer

Aug 2021 - Feb 2024 (2 years 6 months)

- Led the enterprise-wide migration from Jenkins to GitHub Actions, standardizing and centralizing CI/CD pipelines (including Ansible, Packer, and Terraform) across Tokopedia and its subsidiaries — significantly enhancing pipeline scalability and maintainability for 5,000+ engineers.

- Enabled engineers to develop CI/CD use cases using self-hosted GitHub Runners across Tokopedia and its subsidia

PT.Tricada Intronik logoPI

Backend Engineer

Apr 2021 - Aug 2021 (4 months)

- Created, designed, and maintained service applications according to client requirements.
- Deployed and monitored applications automatically using Jenkins, Kubernetes, and other deployment tools.
- Designed and executed queries for flow processing jobs using KSQL and Kafka Designed and composed product API's between services for frontend and UI/UX teams.
- Implemented an alert and notification s

Education

Degrees, certifications, and relevant coursework

PU

President University

Bachelor's degree, Information Technology

2018 - 2022

Get matched with your dream remote job

Sign up now and join over 250,000+ remote workers who receive personalized job alerts, curated job matches, and more for free!

Sign up
Himalayas profile for an example user named Frankie Sullivan