Skip to content
duglee_labs

gaurav sharma

AI-first DevOps & Platform Engineer

AWS Solutions Architect · HashiCorp Terraform Associate

summary

AI-first engineer built on ~10 years of full-stack, cloud, infra, DevOps, and FinOps foundations — deliberately blurring the lines between what infrastructure engineers used to do and what one engineer can now ship. Architects multi-cloud environments (AWS — Certified Solutions Architect, GCP, Azure) with Terraform (HashiCorp Certified), Docker, and Kubernetes orchestration; CI/CD across Jenkins, GitHub Actions, FluxCD, and ArgoCD. Delivers full-scale product development end to end, and builds AI processes and tooling — including an open-source CLI coding agent — that compress the path from idea to shipped software. Brings a FinOps mindset to every project: peak performance at optimized cost, including a 60% infrastructure cost reduction (~$100K/year).

experience

Site Reliability Engineer · Kochava

Aug 2024 — Present · Remote

  • Operate hybrid multi-cloud infrastructure for cross-cloud products and services — AWS and GCP alongside on-prem, including Kubernetes clusters that span all three environments.
  • Own the cross-cloud network fabric: site-to-site VPN tunnel architecture, routing, and connectivity between AWS, GCP, and on-prem.
  • Deploy and run self-hosted platform services end to end (Airflow, databases) — provisioning, upgrades, and reliability — where self-hosting beats managed offerings on cost or control.
  • Drive FinOps across both clouds: cost optimization folded into architecture decisions up front, not retrofitted after the bill arrives.
  • Own reliability, monitoring, and observability across cloud and on-prem systems; land defaults that are secure, compliant, and cost-effective at the same time.
  • Partner with ML teams on MLOps infrastructure: Dask clusters for distributed compute and Airflow-orchestrated model-training pipelines.
  • Build platform automation — from infrastructure workflows to onboarding scripts that turn manual setup into repeatable, auditable runs.
  • Embed AI-led development practices and workflows into infrastructure engineering, working across cross-region teams.

DevOps Lead — DevOps Engineer III · PepperContent Global Pvt. Ltd.

Jan 2021 — Jul 2024 · Remote

  • Scaled infrastructure from a single monolith to a network of 56 microservices as one of the earliest infra hires, owning platform growth from day one.
  • Drove FinOps practices that cut infrastructure costs 60% — roughly $100K/year in savings.
  • Migrated infrastructure to Terraform with 96% coverage.
  • Implemented DevSecOps controls: SSO, least-privilege cloud access, and periodic security & vulnerability audits.
  • Hardened Docker images for security and size; established VCS, branch-protection, review, and SSH-access processes to ISO compliance.
  • Introduced version-controlled infrastructure diagrams and partnered with developers to optimize code and queries for lower resource consumption.

Full-Stack Developer · DevOps Engineer · Sourcefuse Technologies

Feb 2020 — Jan 2021

  • Built a Kubernetes orchestration platform at Rakuten with a drag-and-drop UI for composing and orchestrating Kubernetes YAML.
  • Collaborated with a 21-developer team at RelaySolutions on a US-healthcare appointment & ride-booking product.
  • Worked extensively with Docker manifests and image layering.

Backend Team Lead — Senior Backend & Cloud Developer · Appknit

Feb 2018 — Nov 2020

  • Architected backends and scalable infrastructure for ~30 mobile apps across fitness, education, e-commerce, and ed-tech.
  • Led a team of 5 developers and 2 interns, mentoring them into full-stack engineers.
  • Owned end-to-end infrastructure for multiple clients and helped establish early-stage development processes, project management, and team coordination.

SDE I — Internship · Smartdata Enterprises

Feb 2017 — Jan 2018

  • Stepped in for senior developers to close long-pending, high-priority modules — including a stalled payments module — on schedule.
  • Contributed to MultusMedical (Arizona-based), rendering DICOM medical images in the browser.

independent projects

copair — local-first CLI coding agentcopair.dugleelabs.io

TypeScript · Node.js · OpenAI-compatible & AWS Bedrock APIs · Agentic Development · Open-Weight Models · Release Management · Research & Case Studies

Model-agnostic coding agent with smart model-tier routing, a unified API-key layer, and prompt caching. Open-core: public cli-core via git subtree with GitHub Actions sync and ESLint import-boundary enforcement. Built under dugleelabs (dugleelabs.io).

HLS-transcoding-nodejs — adaptive bitrate streaminggithub.com/dugleelabs/HLS-transcoding-nodejs

Node.js · Apple HLS · S3

Open-source reference implementation of adaptive-bitrate video streaming in Node.js: transcoding to Apple HLS (m3u8) for bandwidth-aware video-on-demand over HTTP. 80+ stars on GitHub, with a companion engineering write-up.

skills

cloud
AWS, GCP, Azure
iac
Terraform, Pulumi, Ansible
containers
Docker, Kubernetes
ci/cd
GitHub Actions, Jenkins, FluxCD, ArgoCD, CircleCI, TravisCI, JFrog Artifactory
observability
Datadog, Prometheus, Grafana, New Relic, ELK
data
MongoDB, Redis, ClickHouse, Kafka, RabbitMQ
security
DevSecOps, HashiCorp Vault, SSO, ISO, SOC 2, Auditing
languages
Python, Go, JavaScript, Shell
backend
Node / Express, Flask, Loopback, Angular, React
ai engineering
Agentic Development, Open-Weight Models, Research & Case Studies
practices
FinOps, MLOps (Dask, Airflow), Release Management, Incident Response, Monitoring & Alerting, Team Leadership

education

Master of Computer Applications (MCA) · Lovely Professional University

2015 — 2017

Bachelor of Computer Applications (BCA) · Himachal Pradesh University

2012 — 2015

next

Deepening work on agentic coding tools and small-model capability research through dugleelabs — and looking for platform problems where infrastructure depth and AI tooling expertise compound.

← back to profile