Production Kubernetes
CCE provisioning, upgrades, node pools, scheduling, resources, RBAC, Helm and troubleshooting.
Read case studyOpen to remote DevOps / Platform / SRE roles
DevOps engineer experienced in Kubernetes, Infrastructure as Code, CI/CD, cloud networking, observability and data/AI-platform infrastructure.
Most recently: almost two years at Sber in the GigaData team, working with Cloud.ru Advanced and Evolution. Earlier: self-managed Kubernetes in on-premise environments.
Selected experience
Sanitized accounts of real work without internal addresses, configurations or employer data.
CCE provisioning, upgrades, node pools, scheduling, resources, RBAC, Helm and troubleshooting.
Read case studyOn-premise clusters with Kubespray and kubeadm, HAProxy/Keepalived, Harbor, FluxCD and GitLab CI/CD.
Read case studyrclone/S3, resource and concurrency tuning, retries, recovery, verification, DNS and private endpoints.
Read case studyTerraform/OpenTofu, VPC, peering, routing, NAT/SNAT, DNS, load balancers and private connectivity.
Read case studyPrometheus, VictoriaMetrics, Grafana, remote write, alerting, cardinality and incident investigation.
Read case studyGitLab CI/CD, Airflow, Ray and Spark-related workloads from the infrastructure and platform side.
Read case studyCareer
A practical progression from development and infrastructure support to production Kubernetes and Platform Engineering.
Production Kubernetes/CCE, Cloud.ru, Terraform/OpenTofu, GitLab CI/CD, data-platform infrastructure, observability and S3/OBS.
Self-managed Kubernetes, Kubespray, kubeadm, GitLab CI/CD, FluxCD, Harbor, HAProxy/Keepalived and multi-environment delivery.
Python, Node.js, Vue.js, Docker, Nginx, Linux, databases, monitoring and application deployment.
Technology map
Kubernetes · CCE · kubeadm · Kubespray · Docker · Helm · FluxCD
Terraform · OpenTofu · Cloud.ru Advanced / Evolution · Yandex Cloud · Ansible
GitLab CI/CD · Runner · Kubernetes Agent · Jenkins · OpenShift
Prometheus · VictoriaMetrics · Grafana · PromQL / MetricQL
VPC · Peering · Routing · NAT/SNAT · DNS · Load Balancers · Private Endpoints
Airflow · Ray · Spark workloads · Kafka · PostgreSQL · Redis · ClickHouse · S3/OBS
Scope note: my work with Ray, Spark, Kafka and databases was primarily from the DevOps/platform side: deployment, Kubernetes integration, networking, access, monitoring, resources and infrastructure troubleshooting.
Runnable evidence
An independently recreated example using synthetic data: copy, safe rerun, unavailable destination, recovery, corruption detection and integrity verification.
Contact
Remote work from Volgograd is preferred. I am considering DevOps, Platform, SRE and infrastructure-focused MLOps roles.