Hello
Skilled and experienced Platform/Infrastructure Engineer with a strong focus on Kubernetes, GPU-accelerated AI/ML infrastructure, and the Cloud-Native ecosystem. With five years of hands-on experience, I specialize in deploying and operating large-scale GPU clusters, enabling AI/ML model serving workloads on bare-metal and cloud Kubernetes platforms, and building the observability and security foundations that production AI infrastructure demands. I have deep expertise in NVIDIA GPU stack management — including GPU Operator, Multi-Instance GPU (MIG), DCGM telemetry, and high-performance interconnects — across air-gapped and sovereign cloud environments.
