Skip to content
View Gaurav-Kaushal's full-sized avatar

Block or report Gaurav-Kaushal

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Gaurav-Kaushal/README.md

Gaurav Kaushal

Senior DevOps Engineer  |  CI/CD  |  Kubernetes  |  Infrastructure Automation  |  Cloud

     


About

Senior DevOps Engineer with 8+ years in IT, currently deployed at Optum / UHG via Artech LLC. I build and manage large-scale infrastructure — physical server fleets, automated CI/CD pipelines, artifact management, and internal tooling that teams rely on.

My focus areas: infrastructure reliability, pipeline automation, and container orchestration. Writing about DevOps at gauravkaushal.tech.


Tech Stack

CI/CD & GitOps

Jenkins GitHub Actions GitLab CI ArgoCD CircleCI

Containers & Orchestration

Docker Kubernetes Helm Amazon ECS OpenShift

Infrastructure as Code

Terraform Ansible Chef Puppet CloudFormation

Cloud Platforms

AWS Azure Google Cloud

Monitoring & Observability

Prometheus Grafana Loki ELK Stack Datadog Nagios

Artifact & Source Management

JFrog Git GitHub Bitbucket

Scripting & Languages

Bash Python Groovy Go

Web Servers & Proxy

Nginx Apache HAProxy

Databases

MySQL PostgreSQL MongoDB Redis DynamoDB

Collaboration & Tooling

Jira Confluence Linux


Projects

Enterprise CI/CD Pipeline Modernization

Jenkins GitHub Actions Docker Groovy

Designed and maintained multi-stage Jenkins pipelines using shared Groovy libraries, serving multiple development teams across a large-scale on-premises environment. Migrated select workflows to GitHub Actions for teams adopting modern SCM practices, establishing reusable workflow templates to enforce consistency across repositories. Integrated automated build, test, and artifact promotion gates — reducing manual intervention in the release process and enforcing quality checks before any code reached staging or production.

  • Shared library architecture for Jenkins — reusable stages across 10+ projects with minimal duplication
  • Parallel execution stages cut average pipeline runtime by ~65%, from 40+ minutes to under 15
  • Branch-based promotion strategy: dev → staging → production with mandatory approval gates
  • Groovy-based notification hooks wired into team channels, reducing delayed-failure detection

Artifact Lifecycle Management — JFrog Artifactory

JFrog Jenkins Bash

Owned the JFrog Artifactory setup supporting multi-team build artifact management across a large physical server fleet. Configured local, remote, and virtual repositories for Maven, Docker, and generic artifact types. Implemented automated cleanup policies and retention rules to manage storage at scale. Integrated Artifactory as the single artifact source of truth for all Jenkins pipelines — eliminating unversioned binaries and reducing environment-specific deployment failures.

  • Repository layout and permission model supporting 10+ teams with isolated access controls
  • Automated promotion of release candidates between Artifactory repositories via Jenkins pipelines
  • AQL-based cleanup policies reclaimed significant storage across a 500+ server fleet
  • Eliminated deployment failures caused by untracked or mismatched binary versions

Infrastructure Automation — Terraform & Ansible

Terraform Ansible Linux Python

Built Terraform modules for repeatable infrastructure provisioning and Ansible playbooks for configuration management across a large fleet of on-premises Linux servers. Standardized server bootstrapping — OS hardening, package installation, user management, and service configuration — into idempotent playbooks that could be applied consistently across new and existing nodes. Eliminated manual SSH-based configuration drift that had accumulated over years of ad-hoc changes.

  • Modular Terraform structure with reusable components, managing 500+ on-premises Linux servers
  • Ansible roles for OS baseline, application deployment, and compliance enforcement across the fleet
  • Dynamic inventory eliminated all static host files, cutting onboarding time for new nodes
  • Drift detection runs integrated into the CI pipeline, flagging config divergence before it reached production

Observability Platform — Prometheus, Grafana & ELK

Prometheus Grafana Loki ELK Stack

Designed a centralized observability stack covering metrics, logs, and alerting for on-premises infrastructure and containerized workloads. Deployed Prometheus with custom exporters for application and system-level metrics. Built Grafana dashboards tailored to infrastructure teams, application owners, and on-call engineers — each with different signal density and alert context. Integrated Loki for log aggregation alongside ELK for long-term log search and audit requirements.

  • Custom recording and alerting rules in Prometheus covering 30+ services, with tiered severity routing
  • Grafana dashboards for 20–30 services, tailored to infra teams, app owners, and on-call engineers
  • Alertmanager escalation policies reduced mean time to acknowledge critical alerts
  • ELK pipeline handling structured log ingestion, parsing, and retention for audit compliance

Containerization & Kubernetes Workload Management

Kubernetes Docker Helm ArgoCD

Led containerization efforts for application workloads — writing production Dockerfiles with multi-stage builds, managing Helm charts for repeatable Kubernetes deployments, and enforcing resource requests/limits, health probes, and rolling update strategies across services. Implemented GitOps-based continuous delivery using ArgoCD, syncing cluster state with Git as the source of truth and eliminating manual kubectl apply workflows from production.

  • Multi-stage Dockerfiles reducing final image sizes significantly, with non-root user enforcement
  • Helm chart structure with environment-specific value overrides across dev/staging/prod
  • ArgoCD GitOps delivery eliminated manual kubectl apply workflows from production entirely
  • Kubernetes RBAC, namespace isolation, and NetworkPolicy enforced across multi-team clusters

GitHub Stats

GitHub Stats

GitHub Streak

Top Languages


Senior DevOps Engineer  |  Open to global roles  |  Based in India, relocation-ready  |  Get in touch

Popular repositories Loading

  1. aws-instance-schedular aws-instance-schedular Public

    An automated EC2 instance scheduler using AWS Lambda and EventBridge. It helps reduce AWS costs by automatically starting and stopping non-production instances based on defined working hours. Addit…

    3

  2. Gaurav-Kaushal Gaurav-Kaushal Public

    Personal Repo

    1

  3. YouTube-Video-Downloader YouTube-Video-Downloader Public

    Download Your Video from YouTube directly.

    Python 1

  4. Google-Forms-Bot Google-Forms-Bot Public

    Python 1

  5. onprem-infra-automation-ansible onprem-infra-automation-ansible Public

    Automating On-Premises Infrastructure Setup for Scientific Research Servers

    1

  6. xgame xgame Public