0%

SAMEER
BANCHHOR

DevOps Engineer with 3 years of experience designing, automating, and managing cloud-native infrastructure & AI platforms. Specialized in Kubernetes, Infrastructure as Code, vLLM & GPU workloads, CI/CD automation, observability, and high-availability operations.

/ DEVOPS ENGINEER   / AI INFRASTRUCTURE ENGINEER   / CLOUD & PLATFORM   / SYSTEM DESIGN
KUBERNETES EKS & GKE AWS / AZURE / GCP TERRAFORM IAC ARGOCD GITOPS vLLM & OLLAMA NVIDIA GPUs & CUDA PROMETHEUS & GRAFANA DOCKER CONTAINERIZATION GITHUB ACTIONS VECTOR DATABASES HASHICORP VAULT PYTHON & BASH & GO HELM & KUSTOMIZE TRITON INFERENCE KUBERNETES EKS & GKE AWS / AZURE / GCP TERRAFORM IAC ARGOCD GITOPS vLLM & OLLAMA NVIDIA GPUs & CUDA PROMETHEUS & GRAFANA DOCKER CONTAINERIZATION
99.95%
Infrastructure Uptime SLA
HA Kubernetes & AWS cloud platform operations
30%
Cloud Cost Savings
Resource optimization & Terraform automation
>70%
Deployment Time Reduction
ArgoCD GitOps & automated CI/CD pipelines
90%
Workflows Automated
End-to-end cloud & app deployment automation
3 Yrs
DevOps & Cloud Operations
Cloud-native & AI infrastructure experience

ENGINEERING CAPABILITIES

Production infrastructure architectures engineered for reliability, automation, and scale.

// Cloud & Orchestration
cloud.providers → AWS, Azure, GCP
containers.k8s → Docker, Kubernetes, Helm, ArgoCD, Istio, Kustomize

DEVOPS & AI INFRASTRUCTURE EXPERIENCE

3 years of experience designing, automating, and scaling production cloud infrastructure and AI model serving.

CORE ROLE • CLOUD & PLATFORM ENGINEERING

DevOps Engineer

2023 – Present • Full-Time

Designing, automating, and managing production-grade cloud-native infrastructure on AWS. Built reliable, scalable, secure, and cost-efficient platforms enabling engineering teams to deploy multiple times daily.

Built production-grade HA Kubernetes clusters and automated cloud infrastructure using Terraform & AWS.
Designed scalable CI/CD pipelines with GitHub Actions, GitLab CI, and Jenkins; implemented GitOps using ArgoCD.
Configured full observability with Prometheus & Grafana; automated backups, disaster recovery, and incident RCA.
Impact: Maintained 99.95% uptime SLA, reduced cloud costs by 30%, automated 90% of deployment workflows, and cut deployment time by over 70%.
AWS Kubernetes Terraform ArgoCD Docker Helm Prometheus Grafana GitHub Actions
SPECIALIZED CAPABILITY • LLM & GPU PLATFORMS

AI Infrastructure & Model Serving Platforms

2023 – Present • Production Systems

Engineering AI infrastructure for LLM applications, GPU workloads, inference deployments, vector databases, and scalable containerized platforms.

Deployed LLM inference servers using vLLM & Ollama and hosted open-source models on Kubernetes using NVIDIA GPU Operator.
Managed GPU workloads, optimized GPU utilization, model startup times, and implemented autoscaling for AI workloads.
Integrated Vector Databases for Retrieval-Augmented Generation (RAG) applications and built scalable inference APIs.
Managed model versioning and deployment with MLflow & Hugging Face, enabling developer self-service model serving platforms.
NVIDIA GPUs CUDA vLLM Ollama Triton Server Hugging Face MLflow Ray Vector DBs LangChain Infra

SYSTEM DESIGN KNOWLEDGE

Proven expertise in blueprinting distributed, fault-tolerant, and high-throughput architectures.

Systems Comfortable Designing
URL Shortener Chat Application Notification Service Distributed Cache Rate Limiter API Gateway Authentication Service File Storage System Video Streaming Platform Event Driven Architecture Distributed Logging Kubernetes Architecture High Availability Systems Microservices AI Inference Platform Multi-Tenant SaaS RAG Infrastructure CI/CD Platform Design
Core Distributed Concepts
CAP Theorem Consistency Models Horizontal Scaling Vertical Scaling Load Balancing Database Sharding Replication Caching Strategies Message Queues Event Streaming Fault Tolerance Consensus Algorithms Service Discovery Reverse Proxy Distributed Tracing

CI/CD & DEPLOYMENT SIMULATORS

Experience real-time GitOps deployment pipelines, interactive terminal command execution, and live cloud infrastructure topology scaling.

LIVE ENGINE

GitOps Automated Pipeline Runner

1. SOURCE
git commit `a8f3b9`
READY
2. CI BUILD
Docker & PyTest
WAITING
3. GITOPS
ArgoCD Sync
WAITING
4. K8S DEPLOY
Rolling Pod Update
WAITING
5. TELEMETRY
Prometheus & SLA
WAITING
Live Pipeline Logs
// Select a preset pipeline above and click "TRIGGER GITOPS DEPLOYMENT" to watch real-time automated execution.
CLI SANDBOX

Interactive Terminal Shell

Run interactive infrastructure commands directly in-browser or click quick presets below.

sameer@infra-node-01:~$
# DevOps & AI Infra Interactive Sandbox v2.4
# Type 'help' or click any quick command above.
sameer@infra-node-01:~$
INFRA CALCULATOR

AI Infra Topology & Cost Estimator

Simulate cluster size, GPU hardware acceleration, and multi-region specs for real-time AWS cost and throughput metrics.

NVIDIA A100/H100 GPU Pool 4 GPUs
EKS CPU Worker Nodes 8 Nodes
Multi-Region Active-Active Failover
Est. Cloud Cost $3,840/mo
LLM Throughput 1,450 tok/s
Uptime SLA 99.95%
GENERATED TERRAFORM (HCL)
resource "aws_eks_node_group" "gpu_pool" {
  cluster_name    = "ai-production-cluster"
  node_group_name = "vllm-a100-nodes"
  instance_types  = ["p4d.24xlarge"]
  desired_size    = 4
}

FEATURED PROJECTS

Production Kubernetes platforms, AI inference servers, GitOps pipelines, IaC automation, and observability stacks.

CONTAINERS & ORCHESTRATION • AWS & TERRAFORM

Production Kubernetes Platform

2023 – Present • Production

High-Availability production Kubernetes cluster engineered with Terraform IaC, AWS cloud services, Helm charts, and GitOps delivery.

HA Kubernetes cluster setup on AWS cloud with automated cluster provisioning via Terraform.
Declarative GitOps deployment workflows, monitoring with Prometheus & Grafana, log aggregation, autoscaling, and disaster recovery.
Kubernetes Terraform AWS Helm Prometheus Grafana
// Highlights: HA Kubernetes Cluster | GitOps Deployment | Prometheus & Grafana Monitoring | Centralized Logging | Horizontal Pod Autoscaling (HPA) | Automated Backup & Disaster Recovery.
AI INFRASTRUCTURE • LLM & GPU SERVING

AI Inference Platform

2024 – Present • Production

Scalable AI model inference and GPU workload management platform powering open-source LLMs and vector search backends.

High-throughput LLM hosting using vLLM, Docker, and Kubernetes GPU scheduling with NVIDIA Operator.
Model management, API gateway integration, authentication, Redis caching, and autoscaling for AI inference workloads.
vLLM Docker Kubernetes NVIDIA GPU FastAPI Redis
// Highlights: LLM Hosting | GPU Scheduling | Auto Scaling | Model Versioning & Management | API Gateway | Authentication & Rate Limiting.
CONTINUOUS DELIVERY • ARGOCD & HELM

GitOps Platform

2024 • Production

Declarative multi-environment continuous delivery platform enabling zero-downtime rollouts and automated rollbacks across clusters.

Multi-environment deployment pipelines utilizing ArgoCD GitOps, Helm charts, and GitHub Actions.
Automated rollbacks, canary deployment strategies, and zero-downtime deployment workflows.
ArgoCD Helm GitHub Actions Kubernetes
// Highlights: Multi-environment deployment | Automated rollbacks | Canary deployment | Zero downtime deployment.
INFRASTRUCTURE AS CODE • TERRAFORM & ANSIBLE

Infrastructure Automation

2023 – 2025

End-to-end cloud infrastructure automation provisioning VPC networking, compute, storage, databases, and IAM security models.

Complete cloud provisioning on AWS using modular Terraform IaC and Ansible automation playbooks.
IAM automation, VPC networking, EC2 compute instances, RDS databases, S3 buckets, and EKS Kubernetes clusters.
Terraform AWS Ansible VPC EC2 RDS S3 EKS
// Highlights: Complete cloud provisioning | IAM automation | Custom VPC Networking | EC2, RDS & S3 automation | EKS Cluster provisioning.
OBSERVABILITY • PROMETHEUS & GRAFANA

Monitoring Stack

2023 – Present

Full-stack metrics, logging, and alerting platform delivering real-time operational visibility across cloud microservices.

Infrastructure metrics collection, application performance monitoring, and automated alerting via Alertmanager.
Custom Grafana dashboards, log aggregation with Loki & ELK Stack, and OpenTelemetry instrumentation.
Prometheus Grafana Loki Alertmanager ELK Stack OpenTelemetry
// Highlights: Infrastructure metrics | Application monitoring | Alertmanager notification routing | Custom Dashboards | Centralized Log Aggregation.
PRODUCTION API • GO 1.24 & DUCKDB 1.5

Durg Voter Production REST API & Analytics Engine

2026 • Production Live

High-performance Go REST API & Analytics Engine serving 1M+ records with sub-15ms latency. Deployed on Google Cloud Run (voter.sameerbanchhor.com).

DuckDB 1.5 CGO engine executing sub-15ms aggregations over 1.04M records.
Enterprise middleware: Token Bucket rate limiter, Request ID tracing, JSON logging.
AI transliteration pipeline (English ↔ Hindi) via Gemini API with <50ms fast search.
Go 1.24 DuckDB 1.5 Google Cloud Run Gemini API Docker OpenAPI 3.0
AUDIO INFRASTRUCTURE • NEURAL SYNTHESIS

Chhattisgarhi Neural Audio Synthesis Pipeline (VITS Deep TTS)

2025 – 2026

End-to-end VITS neural audio synthesis pipeline and model serving framework for Chhattisgarhi dialect (18M+ speakers).

VITS Neural Network PyTorch Model Serving Audio Processing HuggingFace Hub

CERTIFICATIONS & GITHUB HIGHLIGHTS

Industry certifications, open-source contributions, technical writings, and active code repositories.

Certifications & Degrees

AWS Certified Solutions Architect
AWS Certified Developer
AWS Certified SysOps Administrator
Certified Kubernetes Administrator (CKA)
Certified Kubernetes Application Developer (CKAD)
HashiCorp Terraform Associate
Microsoft Azure Administrator
RHCSA (Red Hat Certified System Administrator)

Open Source & Technical Blogs

Contributed to Kubernetes ecosystem & published reusable Helm charts & Terraform modules.
Created GitHub Actions workflows, DevOps automation scripts, and AI infrastructure templates.
Technical Blogs: Kubernetes Deep Dive, DevOps Best Practices, AI Infrastructure, System Design, Cloud Architecture, Terraform, GitOps, Docker Internals, Linux Performance, Distributed Systems.

Featured GitHub Repositories