
Artem Babkin
@artembabkin
I build zero-downtime Kubernetes platforms, infrastructure automation, and observability systems at production scale.
What I'm looking for
I've rebuilt production infrastructure at Rostelecom, migrating more than 100 services to Kubernetes with zero downtime while redesigning logging, monitoring, load balancing, and operating systems.
I design and operate RKE2 platforms with Ceph, Cilium, PostgreSQL, Kafka, Vault, Terraform, Ansible, and GitOps workflows through ArgoCD and Fleet. My work has increased deployment speed by 40–50%, reduced false alerts by roughly 70%, and freed 40–50% of compute resources.
I lead two engineers, run workshops for development teams, and manage incident response and on-call processes. I like infrastructure that helps teams ship reliably instead of getting in the way.
Experience
Work history, roles, and key accomplishments
Architecture and Design of a Production Kubernetes Platform:
- Designed and deployed a production RKE2 cluster from scratch with automated node scaling and a dedicated etcd cluster
- Executed a full, zero-downtime migration from a legacy cluster (v1.20). Implemented HPA for critical services and leveraged affinity/anti-affinity, taints/tolerations, and resource requests/limits to ensure performanc
Automation, Security, and Infrastructure Management:
- Deployed a dedicated Kubernetes cluster for infrastructure services and migrated near by 70% of services into it, freeing up 40-50% of computational resources, enhancing their availability, and preserving existing data.
- Implemented a GitOps approach using ArgoCD for the infrastructure cluster.
- Conducted a full refactoring of Nginx/Angie co
Infrastructure Transformation and Kubernetes Implementation:
- Executed complete legacy infrastructure transformation: migrated 30+ bare-metal servers and 100+ virtual machines to Infrastructure as Code state using Ansible and Terraform
- Designed and implemented company's first Kubernetes clusters for running production and infrastructure workloads
- Automated service lifecycle through GitLab CI/
IT Infrastructure Administration and Monitoring:
- Deployed and maintained centralized monitoring system based on Zabbix 5.0 (CentOS 7 + MySQL + Apache2 + phpMyAdmin)
- Developed custom scripts for monitoring proprietary applications and system metrics
- Created and maintained Active Directory user accounts and groups, Exchange mailboxes
Automation and Process Optimization:
- Developed automation
User Support and System Administration:
- Provided comprehensive technical support for users, peripherals, and video surveillance systems
- Managed server maintenance and network infrastructure deployment
- Automated troubleshooting procedures for common user PC issues
Process Automation and Monitoring:
- Developed Bash scripts to resolve user issues including display resolution problems and DRM
Education
Degrees, certifications, and relevant coursework
Independent University of Moscow
IT Specialist, Network and System Administration/Administrator
Moscow Instrument-Making Technical School (MITS) of the Plekhanov Russian University of Economics
Specialist, Network and System Administration/Administrator
2014 - 2018
Tech stack
Software and tools used professionally
Availability
Location
Authorized to work in
Salary expectations
Social media
Job categories
Skills
Interested in hiring Artem?
You can contact Artem and 90k+ other talented remote workers on Himalayas.
Message ArtemGet matched with your dream remote job
Sign up now and join over 250,000+ remote workers who receive personalized job alerts, curated job matches, and more for free!
