Skip to content
View ZettStai's full-sized avatar

Block or report ZettStai

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
ZettStai/README.md

Hi, I’m Zett 👋

Site Reliability Engineer | Platform Engineer | DevOps & Infrastructure Consultant

Experience supporting startups, enterprise platforms, and globally distributed engineering teams.

9 years building and modernizing production infrastructure, reduce operational noise, and build reliable platforms through Kubernetes, automation, observability, and cloud engineering.

Currently available for:

  • Contract SRE / Platform work
  • Infrastructure modernization
  • Kubernetes migrations
  • Reliability consulting

I help teams:

  • Modernize legacy infrastructure
  • Reduce alert fatigue
  • Migrate to Kubernetes
  • Improve observability and incident response
  • Reduce cloud spend without sacrificing reliability

Worked with:

Palo Alto Networks • Hopin • DeepDyve • Policy Networks • Startups & SMBs

Core technologies

Platform Kubernetes • Helm • Terraform

Observability Prometheus • Grafana

Cloud & Security GCP • Cloudflare

Data PostgreSQL • Redis • Elasticsearch

Automation GitHub Actions • Python • Go • Bash

Production Experience

  • Led observability modernization from legacy based Nagios to Prometheus/Grafana across production environments
  • Reduced alert fatigue through smarter routing, automation, and operational tuning
  • Migrated legacy VM-based infrastructure into containerized Kubernetes platforms
  • Designed resilient multi-cluster environments for disaster recovery and operational continuity
  • Reduced cloud and infrastructure spend through modernization and platform optimization

Featured Case Studies

  • 🚀 Nagios → Prometheus Migration (case study underway)
  • Kubernetes Platform Modernization (case study underway)
  • 📉 Alert Fatigue Reduction Initiative (case study underway)
  • 🔐 Disaster Recovery & Multi-Cluster Resilience (case study underway)

🔗 LinkedIn: Zett Stai

Pinned Loading

  1. terraform-gcp-k8s-demo terraform-gcp-k8s-demo Public

    HCL