Core infrastructure
Kubernetes, GCP, Azure, Linux, Docker, PostgreSQL, multi-region and hybrid-cloud production systems
Lead Platform Engineer · AI Infrastructure & Reliability
Lead Platform Engineer with 5+ years of experience building reliable infrastructure for international AI, HealthTech, TravelTech, InsurTech, and banking products. Operated 8 Kubernetes clusters and 30+ microservices across five continents, supporting products used by 333K+ active users and approximately 400 clinics while reducing release time by 80%, cloud spend by 40%, and incident response time by 50%.
Technical profile
Kubernetes, GCP, Azure, Linux, Docker, PostgreSQL, multi-region and hybrid-cloud production systems
Terraform, Terragrunt, Ansible, Helm, ArgoCD, GitLab CI/CD, Jenkins, GitHub Actions, GitOps
Prometheus, Grafana, Loki, Alertmanager, Sentry, Trivy, SAST/DAST, RBAC, ingress and certificate management
vLLM, Ollama, ClearML, KubeRay, GPU workloads, Python, MongoDB, Redis, Kafka, RabbitMQ
Professional experience
Lead Platform Engineer
Senior DevOps Engineer
Senior DevOps Engineer
DevOps Engineer · Contract
DevOps Engineer · Software Engineering Intern
Selected additional engagements
Infrastructure consulting
Production reliability and infrastructure work for a power-bank rental platform in Georgia: Docker-based service operations, observability, deployment safety, and investigation of database and application-performance incidents.
Post-employment consulting · Project-based
Returned after the full-time role for a separate infrastructure engagement, providing delivery continuity and applying deep knowledge of the company’s multi-cloud production environment.
Volunteer experience
Sep 2015 — Dec 2017 · Human Rights
Kutafin Moscow State Law University (MSAL)
Provided pro bono legal assistance to socially vulnerable communities.
Sep 2017 — Jan 2018 · Science and Technology
Created, introduced, maintained, and processed technical documentation for the organization.
Sep 2021 — Present · Science and Technology
Gareni / Pamix / Tabler
Advise colleagues on programming languages, databases, automation, and optimization.
May 2023 · Social Services
Supported a one-month social-services distribution initiative in Paris.
Jun 2023 · Arts and Culture
Provided short-term administrative support at the Paris startup campus.
Jun 2023 · Science and Technology
Supported on-site operations at the Viva Technology event in Paris.
Selected complex systems work
01
Diagnocat · Scale · Reliability
Architected and maintained eight Kubernetes clusters across GCP and Yandex Cloud, supporting more than 30 microservices in five continental regions with high-availability and disaster-recovery requirements.
Outcome: release cycle reduced from 4–6 hours to under one hour; incident response time reduced by 50%.
02
Confidential AI Platform · PostgreSQL · Azure
Responded to an Azure database outage that caused mass 5xx responses and broken authentication, then delivered a highly available PostgreSQL production setup and improved database visibility and maintenance controls.
Closed production work: HA PostgreSQL, replica-load analysis, Grafana datasource and environment cronjob alignment.
03
MLOps · GPU · AIOps
Combined model-serving operations with automated incident analysis: production LLM inference at Diagnocat; GPU/NVML recovery, readiness controls, HolmesGPT, and MCP-connected service anomaly analysis at a confidential AI platform.
Outcome at Diagnocat: four or more models in production and 4× lower external AI API costs.
04
Confidential AI Platform · Security · Delivery
Migrated delivery workflows to GitLab, hardened ingress client-IP trust, introduced managed credential sharing and database RBAC, and investigated WAF protection for the production edge.
Security work covered the full path from source control and secrets to ingress, rate limits, database access, and production networking.
05
Selected product · Full stack · AI
Designed and built a production personal-growth platform across React, Node.js, Flutter, MongoDB, Redis/BullMQ, Qdrant, and OpenAI APIs, including semantic search, asynchronous AI workflows, encryption, and data export.
Demonstrates product ownership beyond infrastructure: application architecture, mobile and web delivery, AI pipelines, privacy controls, and self-hosted K3s operations.
Latest writing
Systems Thinking
Why real stability comes from personal discipline, responsibility, and the ability to keep acting when circumstances change.
Software Engineering
What two difficult OCaml projects at École 42 taught me about expressions, types, state, and becoming a more disciplined engineer.
AI Infrastructure
A practical, low-cost architecture connecting observability, source control, an LLM, and team chat to answer infrastructure questions automatically.
Education
2022 — 2025
RNCP Level 7 · Expert in IT Architecture — Information Systems & Networks
Project-based computer science training centered on independent problem solving, peer learning, software engineering, and systems thinking.
2019 — 2022
Software Engineering Program
Peer-to-peer, project-based training in software engineering, systems programming, and collaborative development.
Graduated 2018
Law
Legal education supporting structured analysis, risk assessment, governance, and compliance-minded engineering.
Russian · Tajik · English
Next opportunity
I bring a builder’s mindset, production ownership, and a bias toward clear, automated, maintainable infrastructure.